ragflow

mirror of https://github.com/infiniflow/ragflow.git synced 2026-07-04 01:29:35 +08:00

Author	SHA1	Message	Date
Panda Dev	2fd8cdc3cc	fix(go): wire CheckConnection to ListModels in ollama, lm-studio, and vllm (#14614 ) ### What problem does this PR solve? Three Go drivers had `CheckConnection` returning a hardcoded `no such method` error, even though each one already has a working `ListModels` that hits the configured base URL with the configured API key. So the "Check connection" button in the model provider UI always failed for these three providers, even when the underlying setup was fine. Affected drivers: - `internal/entity/models/ollama.go` - `internal/entity/models/lmstudio.go` - `internal/entity/models/vllm.go` This is a real user-facing gap because Ollama and LM Studio are two of the most popular local LLM runners, and vLLM is widely used for self-hosted deployments. ### What this PR includes For each of the three drivers, replace the stub with a small implementation that calls `ListModels` and returns its error: ```go func (o OllamaModel) CheckConnection(apiConfig APIConfig) error { _, err := o.ListModels(apiConfig) return err } ``` This is the exact pattern that xai, moonshot, deepseek, aliyun, and gitee already use for the same method. No JSON change. No factory change. No interface change. ### Type of change - [x] Bug Fix (non-breaking change which fixes an issue) ### How was this tested? - `go build ./internal/entity/models/...` in a clean go 1.25 image (the go.mod minimum) returns exit 0. - The full ModelDriver interface still resolves on each driver (NewInstance, Name, ChatWithMessages, ChatStreamlyWithSender, Encode, Rerank, ListModels, Balance, CheckConnection). - Pattern parity with the existing xai, moonshot, deepseek, aliyun, and gitee CheckConnection methods. Closes #14609	2026-05-08 12:00:10 +08:00
Panda Dev	bb10b83e61	Go: implement Rerank in ZhipuAI driver (#14608 ) ### What problem does this PR solve? The ZhipuAI Go driver had a stub Rerank method that returned "not implemented", even though conf/models/zhipu-ai.json already ships glm-rerank as a rerank model and the rerank URL suffix is already wired in url_suffix: ```json "url_suffix": { ... "rerank": "rerank" }, "models": [ {"name": "glm-rerank", "model_types": ["rerank"]}, ... ] ``` So the config was ready but the driver was not. A tenant who picked glm-rerank in the Go layer could not actually run a rerank call. This PR fills the gap so the listed model works end to end. ### What this PR includes - `internal/entity/models/zhipu-ai.go`: real implementation of `ZhipuAIModel.Rerank`, plus two small local types (`zhipuRerankRequest`, `zhipuRerankResponse`) that mirror the standard OpenAI-compatible rerank shape used by SiliconFlow. No factory change. No JSON change. No interface change. ### How the driver works - POST to `${BaseURL}/${URLSuffix.Rerank}` (resolves to `https://open.bigmodel.cn/api/paas/v4/rerank` with the default config), reusing the existing httpClient on the driver. - Validate apiConfig and the API key, validate the model name, and resolve the region. Return a clear local error before any HTTP call when something is missing. - Send `{model, query, documents, top_n, return_documents: false}` in the body, the same shape the SiliconFlow driver already uses. - Walk `results[].relevance_score` and copy each score into the output slice indexed by `results[].index`, so the output order matches the input order even if the API returns results in a different order. - Empty `texts` input returns an empty `[]float64` with no HTTP call. - Non-200 responses propagate the upstream status line and body. ### Type of change - [x] New Feature (non-breaking change which adds functionality) ### How was this tested? - `go build ./internal/entity/models/...` in a clean go 1.25 image (the go.mod minimum) returns exit 0. - The full method set on `ZhipuAIModel` still matches the `ModelDriver` interface (NewInstance, Name, ChatWithMessages, ChatStreamlyWithSender, Encode, ListModels, Balance, CheckConnection, Rerank). - Pattern parity with the existing SiliconFlow Rerank implementation (`internal/entity/models/siliconflow.go`). Closes #14607	2026-05-07 17:56:30 +08:00
Jin Hai	94324afee9	Go: fix auth issue in hybrid mode (#14611 ) ### What problem does this PR solve? Since secret key get and set logic is updated, the go server also need to update. ### Type of change - [x] Bug Fix (non-breaking change which fixes an issue) --------- Signed-off-by: Jin Hai <haijin.chn@gmail.com>	2026-05-07 17:14:22 +08:00
Haruko386	078ea3bf4a	Go: implement provider: Nvidia (#14623 ) ### What problem does this PR solve? 1. Implement `Nvidia` Provider: Fully support NVIDIA NIM APIs with robust parameter handling (including the `thinking` parameter) and safe URL merging in `NewInstance`. 2. Fix Misleading CLI Errors: Corrected a bug in `common_command.go` where failed chat requests inaccurately reported `failed to list instance models`. ### Type of change - [x] Bug Fix (non-breaking change which fixes an issue) - [x] New Feature (non-breaking change which adds functionality)	2026-05-07 14:17:57 +08:00
Panda Dev	b8b741555f	Go: implement provider: OpenAI (#14605 ) ### What problem does this PR solve? Add a Go driver for OpenAI (GPT models). The config file conf/models/openai.json has been in the repo for a while with the full GPT-5 model list, but internal/entity/models/factory.go had no case for "openai". So any tenant that configured OpenAI as a model provider in the Go layer fell through to the default branch and got the dummy driver. Chat, list models, and check connection all returned dummy responses instead of reaching the API. OpenAI is the most commonly requested provider and the JSON config already ships with the repo, so this gap is high impact even though the JSON has been there for some time. ### What this PR includes - New file internal/entity/models/openai.go with an OpenAIModel that implements the ModelDriver interface. - factory.go: route the "openai" provider name to NewOpenAIModel. - conf/models/openai.json: add "models": "models" under url_suffix so ListModels can hit /v1/models with no hardcoded fallback. ### How the driver works - OpenAI exposes the canonical OpenAI-compatible API at https://api.openai.com/v1. - ChatWithMessages and ChatStreamlyWithSender post to /chat/completions in the same shape the moonshot, vllm, and xai drivers use. - ListModels and CheckConnection call /models to list available ids and confirm the API key works. - reasoning_content is passed through for the o-series and other reasoning models, in both the non-stream and stream paths. - Encode (embeddings) is left as "not implemented" for now, the same way the other recent provider drivers do it. Rerank and Balance are not part of OpenAI's public API surface in this layer and return a clear "not implemented" or "no such method" error. ### Type of change - [x] New Feature (non-breaking change which adds functionality) ### How was this tested? - go build ./internal/entity/models/... in a clean go 1.25 image (the go.mod minimum) returns exit 0 with no errors. - Method set of OpenAIModel matches the ModelDriver interface: NewInstance, Name, ChatWithMessages, ChatStreamlyWithSender, Encode, Rerank, ListModels, Balance, CheckConnection. - Pattern parity with the merged moonshot (#14433), volcengine (#14460), minimax (#14478), vllm (#14532), xai (#14550), and lm-studio (#14586) PRs. Closes #14604	2026-05-07 13:09:51 +08:00
Haruko386	dd7a0ce1d3	Go: implement provider: lm-studio (#14586 ) ### What problem does this PR solve? implement `lm-studio` provider ### Type of change - [x] New Feature (non-breaking change which adds functionality) - [x] Refactoring	2026-05-06 19:23:11 +08:00
Jack Storment	c2ad672c09	Go: implement provider: xAI (#14550 ) Closes #14552 ### What problem does this PR solve? Add a Go driver for xAI (Grok models). The config file conf/models/xai.json has been in the repo since the early Go provider work, but internal/entity/models/factory.go had no case for "xai". So any xAI request fell through to the dummy driver and never reached the API. This PR adds the missing driver and wires it up. ### What this PR includes - New file internal/entity/models/xai.go with an XAIModel that implements the ModelDriver interface. - factory.go: route the "xai" provider name to NewXAIModel. ### How the driver works - xAI exposes an OpenAI-compatible API at https://api.x.ai/v1. - ChatWithMessages and ChatStreamlyWithSender post to /chat/completions in the same shape the moonshot and deepseek drivers use. - ListModels and CheckConnection call /models to confirm the API key works and to list available model ids. - reasoning_content is passed through for grok-3-mini and other xAI reasoning models, both in the non-stream and stream paths. - Encode, Rerank, and Balance are not part of the public xAI API at the moment, so they return a clear "not implemented" or "no such method" error. ### Type of change - [x] New Feature (non-breaking change which adds functionality) ### How was this tested? - go build ./internal/entity/models/... in a clean go 1.25 image (the go.mod minimum) returns exit 0 with no errors. - Method set of XAIModel matches the ModelDriver interface: NewInstance, Name, ChatWithMessages, ChatStreamlyWithSender, Encode, Rerank, ListModels, Balance, CheckConnection. - Pattern parity with the merged moonshot (#14433), volcengine (#14460), minimax (#14478), and vllm (#14532) PRs. --------- Co-authored-by: Jin Hai <haijin.chn@gmail.com>	2026-05-06 12:16:37 +08:00
Haruko386	cd54c08e84	Go: implement provider: Ollama (#14580 ) ### What problem does this PR solve? implement `Ollama` provider ### Type of change - [x] New Feature (non-breaking change which adds functionality) - [x] Refactoring --------- Co-authored-by: Jin Hai <haijin.chn@gmail.com>	2026-05-06 12:03:58 +08:00
qinling0210	7335916868	Use GetChatModel, remove duplicate functions in model_service.go (#14546 ) ### What problem does this PR solve? Use GetChatModel, remove duplicate functions in model_service.go ### Type of change - [x] Refactoring Co-authored-by: Jin Hai <haijin.chn@gmail.com>	2026-05-06 11:33:32 +08:00
Jin Hai	aa57b5bd8b	Go: move logger to common module (#14545 ) ### What problem does this PR solve? As title ### Type of change - [x] Refactoring Signed-off-by: Jin Hai <haijin.chn@gmail.com>	2026-05-06 10:41:58 +08:00
Jin Hai	3a51c27a75	Go: CLI chat with text, image, video (#14573 ) ### What problem does this PR solve? ``` RAGFlow(user)> chat with 'glm-4.6v-flash@test@zhipu-ai' message 'What are the pics talk about?' image 'https://cdn.bigmodel.cn/static/logo/register.png' 'https://cdn.bigmodel.cn/static/logo/api-key.png' Answer: The first picture shows a login/register modal with options for phone number login, account login, and WeChat QR code login, along with a prompt for new users to get a 20 million tokens experience package. The second picture displays the API keys management page of a platform, including a warning about API key security and a table listing existing API keys with details like creation time and usage history. Time: 31.600545 RAGFlow(user)> chat with 'glm-4.6v-flash@test@zhipu-ai' message 'What are the video talk about?' video 'https://cdn.bigmodel.cn/agent-demos/lark/113123.mov' Answer: Based on the sequence of frames provided, the video is a demonstration of a web search and navigation process. 1. The video starts with a blank Google search page. 2. The user types "智谱" (which is the Chinese name for the company Zhipu AI) into the search box. 3. The search is initiated and the page shows "About 0 results". 4. The search results load, showing information about Zhipu AI, including its website. 5. The user clicks on the main website link (www.zhipuai.cn). 6. The video ends by showing the homepage of Zhipu AI's website, titled "Z.ai GLM Large Model Open Platform". In summary, the video is about searching for the company "智谱" (Zhipu AI) on Google and then navigating to its official website. Time: 76.582520 ``` ### Type of change - [x] New Feature (non-breaking change which adds functionality) Signed-off-by: Jin Hai <haijin.chn@gmail.com>	2026-05-05 18:14:39 +08:00
qinling0210	12af73f2ca	Support stream for multimodal chat (#14537 ) ### What problem does this PR solve? Support stream for multimodal chat ### Type of change - [x] Refactoring	2026-04-30 19:33:57 +08:00
Haruko386	93f3b90121	Go: implement provider: Vllm (#14532 ) ### What problem does this PR solve? Implement the vLLM model provider for RAGFlow to fully support local and self-hosted open-source models (e.g., Qwen, GLM, Llama) via the vLLM framework, and fix several critical bugs related to model instance management and API requests. Key changes and fixes: 1. Added Standard vLLM Provider (`vllm.go`, `vllm.json`): - Implemented `VllmModel` driver strictly adhering to the OpenAI API specification. - Removed hardcoded and dangerous routing logic (e.g., forcing `AsyncChat` for Qwen/GLM prefixes), ensuring standard `/v1/chat/completions` compatibility. - Refactored `ListModels` to use safe JSON parsing (resolving nil pointer panics) and standard `GET` requests without bodies. - Added `APIConfig.Region` fallback logic to prevent empty `base_url` fetching when checking models. 2. Fixed `ChatToModelStreamWithSender` Bug (`model_service.go`): - Resolved the `model is disabled` error when streaming chat with local database-saved models. - Added the missing `if modelInfo.Status == "active"` block to correctly invoke `NewInstance` and inject the dynamic `base_url` into the provider driver before starting the SSE stream. 3. Fixed `ListSupportedModels` Bug (`model_service.go`): - Added dynamic `NewInstance` injection for `base_url`. Previously, the list models function used the static JSON config without injecting user-configured dynamic URLs from the database, resulting in an `unsupported protocol scheme ""` error. ### Type of change - [x] Bug Fix (non-breaking change which fixes an issue) - [x] New Feature (non-breaking change which adds functionality)	2026-04-30 16:30:14 +08:00
qinling0210	265f92c83e	Simplify chat and support multimodal chat (#14523 ) ### What problem does this PR solve? Simplify chat and support multimodal chat ### Type of change - [x] Refactoring	2026-04-30 15:25:01 +08:00
Yingfeng	4ee0702aed	Feat: add skills space to context engine (#13908 ) ### What problem does this PR solve? issue #13714 ### Type of change - [x] New Feature (non-breaking change which adds functionality)	2026-04-30 12:36:03 +08:00
Jin Hai	261be81127	Go: add drop instance models (#14485 ) ### What problem does this PR solve? 1. drop instance model 2. Fix issue of drop instance but not drop models. ### Type of change - [x] New Feature (non-breaking change which adds functionality) Signed-off-by: Jin Hai <haijin.chn@gmail.com>	2026-04-29 19:18:49 +08:00
Haruko386	0e1477eb23	Go: implement provider: MiniMax (#14478 ) ### What problem does this PR solve? implement MiniMax provider ### Type of change - [x] Bug Fix (non-breaking change which fixes an issue) - [x] New Feature (non-breaking change which adds functionality)	2026-04-29 19:06:40 +08:00
Jin Hai	bb05a8bd7e	Update create model instance command (#14441 ) ### What problem does this PR solve? 1. support command: ``` RAGFlow(user)> create provider 'vllm' instance 'test' key 'test-key' url 'base-url' region 'abc'; SUCCESS RAGFlow(user)> list instances from 'vllm'; +----------+----------------------------------------+----------------------------------+--------------+----------------------------------+--------+ \| apiKey \| extra \| id \| instanceName \| providerID \| status \| +----------+----------------------------------------+----------------------------------+--------------+----------------------------------+--------+ \| test-key \| {"base_url":"base-url","region":"abc"} \| 40213c89430311f1a7cf38a74640adcc \| test \| b4d40e6142d311f1a4f938a74640adcc \| enable \| +----------+----------------------------------------+----------------------------------+--------------+----------------------------------+--------+ ``` 2. support add vllm model ``` RAGFlow(user)> add model 'Qwen/Qwen2-0.5B' to provider 'vllm' instance 'test' with tokens 131072 chat; SUCCESS ``` 3. add vllm chat ### Type of change - [x] New Feature (non-breaking change which adds functionality) - [x] Refactoring --------- Signed-off-by: Jin Hai <haijin.chn@gmail.com>	2026-04-29 17:05:08 +08:00
qinling0210	486ca463aa	Port PR14454 to GO (PruneDeletedChunks) (#14463 ) ### What problem does this PR solve? Port PR14454 to GO (PruneDeletedChunks) ### Type of change - [x] Bug Fix (non-breaking change which fixes an issue)	2026-04-29 17:04:22 +08:00
Haruko386	decf673049	Go: implement provider: volcengine (#14460 ) ### What problem does this PR solve? implement `volcengine` provider ### Type of change - [x] New Feature (non-breaking change which adds functionality)	2026-04-29 15:45:08 +08:00
qinling0210	f3c232cf47	Remove model_bundle.go, modify chat_session.go (#14458 ) ### What problem does this PR solve? Remove model_bundle.go, modify chat_session.go ### Type of change - [x] Refactoring	2026-04-29 14:44:12 +08:00
Jin Hai	b493a33316	Go: update chat URL (#14453 ) ### What problem does this PR solve? Update the URL to: /api/v1/chat/completions ### Type of change - [x] Refactoring Signed-off-by: Jin Hai <haijin.chn@gmail.com>	2026-04-29 11:45:06 +08:00
qinling0210	dcce864d4c	Simplify Encode (#14437 ) ### What problem does this PR solve? Simplify Encode ### Type of change - [x] Refactoring	2026-04-28 18:07:42 +08:00
Haruko386	4e5a093ac5	Go: implement provider: Moonshot (#14433 ) ### What problem does this PR solve? implement `Moonshot` provider ### Type of change - [x] New Feature (non-breaking change which adds functionality)	2026-04-28 18:06:25 +08:00
Jin Hai	f670913bb4	Refactor model type to model class (#14426 ) ### What problem does this PR solve? As title ### Type of change - [x] Refactoring Signed-off-by: Jin Hai <haijin.chn@gmail.com>	2026-04-28 16:05:15 +08:00
Jin Hai	7c25870923	Go: update db model (#14423 ) ### What problem does this PR solve? As title. ### Type of change - [x] Refactoring Signed-off-by: Jin Hai <haijin.chn@gmail.com>	2026-04-28 16:04:55 +08:00
Jin Hai	ae420f6358	Go: fix compilation (#14418 ) ### What problem does this PR solve? Add methods to volcengine ### Type of change - [x] Bug Fix (non-breaking change which fixes an issue) Signed-off-by: Jin Hai <haijin.chn@gmail.com>	2026-04-28 13:21:05 +08:00
qinling0210	effc84a042	Refactor model in GO (#14398 ) ### What problem does this PR solve? Refactor model in GO ### Type of change - [x] Refactoring	2026-04-28 12:59:01 +08:00
Jin Hai	819257f257	Go: add volcengine (#14409 ) ### What problem does this PR solve? 1. Refactor server_main 2. Add volcengine ### Type of change - [x] Refactoring --------- Signed-off-by: Jin Hai <haijin.chn@gmail.com>	2026-04-28 12:12:58 +08:00
Jin Hai	965717c4fb	Go: add new provider: google (#14395 ) ### What problem does this PR solve? As title. ### Type of change - [x] New Feature (non-breaking change which adds functionality) --------- Signed-off-by: Jin Hai <haijin.chn@gmail.com>	2026-04-27 20:35:47 +08:00
Jin Hai	c3eac4103a	Go: aliyun model provider (#14379 ) ### What problem does this PR solve? As title. ### Type of change - [x] New Feature (non-breaking change which adds functionality) --------- Signed-off-by: Jin Hai <haijin.chn@gmail.com>	2026-04-27 14:53:33 +08:00
Jin Hai	1c244df90d	Go: add gitee and siliconflow as model provider (#14336 ) ### What problem does this PR solve? As title ### Type of change - [x] New Feature (non-breaking change which adds functionality) --------- Signed-off-by: Jin Hai <haijin.chn@gmail.com>	2026-04-24 20:59:30 +08:00
qinling0210	1473000135	Implement retrieval_test in GO (#14231 ) ### What problem does this PR solve? Implement retrieval_test in GO ### Type of change - [x] Refactoring	2026-04-24 15:30:14 +08:00
Jin Hai	2b029882d7	Go: add new provider minimax (#14296 ) ### What problem does this PR solve? 1. Add new provider minimax 2. Add new command: CHECK INSTANCE 'instance_name' FROM 'provider_name'; ``` RAGFlow(user)> check instance 'test' from 'minimax'; SUCCESS ``` ### Type of change - [x] New Feature (non-breaking change which adds functionality) Signed-off-by: Jin Hai <haijin.chn@gmail.com>	2026-04-23 10:16:20 +08:00
bohdansolovie	e0f0eb277d	Fix upload stream handling to prevent truncated files (#14267 ) ## Summary - Replace single `Read()` call in Go upload service with `io.ReadAll()`. - Prevent potential truncated/corrupted file content during multipart upload. - Keep existing API behavior unchanged while fixing data integrity risk. ## Root Cause `io.Reader.Read()` may return fewer bytes than requested without an error. The previous implementation read once into a full buffer and assumed all bytes were populated. ## Test plan - Upload files of multiple sizes and verify uploaded content integrity. - Confirm upload endpoint still returns successful responses. - Verify downstream document parsing works on uploaded files. ## Issues Closes #14266	2026-04-22 16:32:38 +08:00
Jin Hai	74b44e1aa3	Go: add balance command (#14262 ) ### What problem does this PR solve? ``` RAGFlow(user)> list supported models from 'moonshot' 'test'; +---------------------------------+ \| model_name \| +---------------------------------+ \| moonshot-v1-32k-vision-preview \| \| kimi-k2.6 \| \| moonshot-v1-8k \| \| moonshot-v1-auto \| \| moonshot-v1-128k \| \| moonshot-v1-32k \| \| kimi-k2.5 \| \| moonshot-v1-8k-vision-preview \| \| moonshot-v1-128k-vision-preview \| +---------------------------------+ RAGFlow(user)> show balance from 'moonshot' 'test'; +---------+----------+ \| balance \| currency \| +---------+----------+ \| 0 \| CNY \| +---------+----------+ ``` ### Type of change - [x] New Feature (non-breaking change which adds functionality) --------- Signed-off-by: Jin Hai <haijin.chn@gmail.com>	2026-04-21 21:31:50 +08:00
Jin Hai	e48d75987c	Go: add stream / think chat (#14242 ) ### What problem does this PR solve? 1. Supports stream and non-stream chat 2. Supports think and non-think chat 3. List supported models from DeepSeek service. (This command can be used to verify the API validity) ### Type of change - [x] New Feature (non-breaking change which adds functionality) --------- Signed-off-by: Jin Hai <haijin.chn@gmail.com>	2026-04-21 16:52:32 +08:00
Jin Hai	f269ee9739	Go: add thinking features to zhipu-ai (#14234 ) ### What problem does this PR solve? ``` RAGFlow(user)> list models from 'zhipu-ai'; +------------+------------+---------------+----------------+ \| features \| max_tokens \| model_types \| name \| +------------+------------+---------------+----------------+ \| [thinking] \| 128000 \| [chat] \| glm-4.7 \| \| [thinking] \| 128000 \| [chat] \| glm-4.5 \| \| [thinking] \| 128000 \| [chat vision] \| glm-4.6v-Flash \| \| [thinking] \| 128000 \| [chat] \| glm-4.5-x \| \| [thinking] \| 128000 \| [chat] \| glm-4.5-air \| \| [thinking] \| 128000 \| [chat] \| glm-4.5-airx \| \| [thinking] \| 128000 \| [chat] \| glm-4.5-flash \| \| [thinking] \| 64000 \| [vision] \| glm-4.5v \| \| \| 128000 \| [chat] \| glm-4-plus \| \| \| 128000 \| [chat] \| glm-4-0520 \| \| \| 128000 \| [chat] \| glm-4 \| \| \| 8000 \| [chat] \| glm-4-airx \| \| \| 128000 \| [chat] \| glm-4-air \| \| \| 128000 \| [chat] \| glm-4-flash \| \| \| 128000 \| [chat] \| glm-4-flashx \| \| \| 1000000 \| [chat] \| glm-4-long \| \| \| 128000 \| [chat] \| glm-3-turbo \| \| \| 2000 \| [vision] \| glm-4v \| \| \| 8192 \| [chat] \| glm-4-9b \| \| \| 512 \| [embedding] \| embedding-2 \| \| \| 512 \| [embedding] \| embedding-3 \| \| \| 4096 \| [asr] \| glm-asr \| \| \| 0 \| [tts] \| glm-tts \| \| \| 0 \| [ocr] \| glm-ocr \| \| \| 0 \| [rerank] \| glm-rerank \| +------------+------------+---------------+----------------+ ``` ### Type of change - [x] New Feature (non-breaking change which adds functionality) Signed-off-by: Jin Hai <haijin.chn@gmail.com>	2026-04-20 21:53:27 +08:00
Jin Hai	af2ed416a7	Add extra field to model instance (#14203 ) ### What problem does this PR solve? Now each model support region with different URL ### Type of change - [x] New Feature (non-breaking change which adds functionality) --------- Signed-off-by: Jin Hai <haijin.chn@gmail.com>	2026-04-20 15:31:12 +08:00
Jin Hai	94106646e7	Go: set and list default models (#14191 ) ### What problem does this PR solve? ``` RAGFlow(user)> set default vlm "zhipu-ai" "ccc" "glm-4.6v-flash"; SUCCESS RAGFlow(user)> list default models; +--------+----------------+----------------+----------------+------------+ \| enable \| model_instance \| model_name \| model_provider \| model_type \| +--------+----------------+----------------+----------------+------------+ \| true \| ccc \| glm-4.6v-flash \| zhipu-ai \| llm \| \| true \| ccc \| glm-4.6v-flash \| zhipu-ai \| image2text \| +--------+----------------+----------------+----------------+------------+ ``` ### Type of change - [x] New Feature (non-breaking change which adds functionality) Signed-off-by: Jin Hai <haijin.chn@gmail.com>	2026-04-17 18:05:33 +08:00
Jin Hai	e03212fd7a	Fix go cli models command and api (#14166 ) ### What problem does this PR solve? ``` RAGFlow(user)> list providers; +--------------------------------------+----------+-------------------------------------------+--------------+ \| base_url \| name \| tags \| total_models \| +--------------------------------------+----------+-------------------------------------------+--------------+ \| https://open.bigmodel.cn/api/paas/v4 \| ZHIPU-AI \| LLM,TEXT EMBEDDING,SPEECH2TEXT,MODERATION \| 21 \| \| https://api.x.ai/v1 \| xAI \| LLM \| 6 \| +--------------------------------------+----------+-------------------------------------------+--------------+ RAGFlow(user)> show provider 'zhipu-ai'; +--------------------------------------+----------+-------------------------------------------+--------------+ \| base_url \| name \| tags \| total_models \| +--------------------------------------+----------+-------------------------------------------+--------------+ \| https://open.bigmodel.cn/api/paas/v4 \| ZHIPU-AI \| LLM,TEXT EMBEDDING,SPEECH2TEXT,MODERATION \| 21 \| +--------------------------------------+----------+-------------------------------------------+--------------+ RAGFlow(user)> delete provider 'zhipu-ai'; SUCCESS RAGFlow(user)> add provider 'zhipu-ai'; SUCCESS RAGFlow(user)> create provider 'zhipu-ai' instance 'ccc' 'ccxxccxx'; SUCCESS RAGFlow(user)> list instances from 'zhipu-ai'; +---------------------------------------------------+----------------------------------+--------------+----------------------------------+--------+ \| apiKey \| id \| instanceName \| providerID \| status \| +---------------------------------------------------+----------------------------------+--------------+----------------------------------+--------+ \| ccxxccxx \| 640dd7ee398711f1bdd838a74640adcc \| ccc \| d1d59de5398411f1bdd838a74640adcc \| active \| +---------------------------------------------------+----------------------------------+--------------+----------------------------------+--------+ RAGFlow(user)> list models from 'zhipu-ai'; +----------+------------+---------------+---------------+ \| features \| max_tokens \| model_types \| name \| +----------+------------+---------------+---------------+ \| map[] \| 128000 \| [chat] \| glm-4.7 \| \| map[] \| 128000 \| [chat] \| glm-4.5 \| \| map[] \| 128000 \| [chat] \| glm-4.5-x \| \| map[] \| 128000 \| [chat] \| glm-4.5-air \| \| map[] \| 128000 \| [chat] \| glm-4.5-airx \| \| map[] \| 128000 \| [chat] \| glm-4.5-flash \| \| map[] \| 64000 \| [image2text] \| glm-4.5v \| \| map[] \| 128000 \| [chat] \| glm-4-plus \| \| map[] \| 128000 \| [chat] \| glm-4-0520 \| \| map[] \| 128000 \| [chat] \| glm-4 \| \| map[] \| 8000 \| [chat] \| glm-4-airx \| \| map[] \| 128000 \| [chat] \| glm-4-air \| \| map[] \| 128000 \| [chat] \| glm-4-flash \| \| map[] \| 128000 \| [chat] \| glm-4-flashx \| \| map[] \| 1000000 \| [chat] \| glm-4-long \| \| map[] \| 128000 \| [chat] \| glm-3-turbo \| \| map[] \| 2000 \| [image2text] \| glm-4v \| \| map[] \| 8192 \| [chat] \| glm-4-9b \| \| map[] \| 512 \| [embedding] \| embedding-2 \| \| map[] \| 512 \| [embedding] \| embedding-3 \| \| map[] \| 4096 \| [speech2text] \| glm-asr \| +----------+------------+---------------+---------------+ RAGFlow(user)> disable model 'glm-4.5-flash' from 'zhipu-ai' 'ccc'; SUCCESS RAGFlow(user)> drop instance 'ccc' from 'zhipu-ai'; SUCCESS RAGFlow(user)> list instances from 'zhipu-ai'; No data to print ``` Signed-off-by: Jin Hai <haijin.chn@gmail.com>	2026-04-17 09:55:25 +08:00
chanx	1031aebc8f	feat(file): Add file ancestor directory lookup feature by go (#14037 ) ### What problem does this PR solve? feat(file): Add file ancestor directory lookup feature by go ### Type of change - [x] New Feature (non-breaking change which adds functionality)	2026-04-14 15:22:03 +08:00
chanx	6aec8058bb	refactor: Remove knowledge base-related API handlers that are already included in the dataset. (#14094 ) ### What problem does this PR solve? refactor: Remove knowledge base-related API handlers that are already included in the dataset. ### Type of change - [x] Refactoring	2026-04-14 15:19:31 +08:00
Jin Hai	3e787b3b09	Go: update search (#14023 ) ### What problem does this PR solve? Update search ### Type of change - [x] New Feature (non-breaking change which adds functionality) Signed-off-by: Jin Hai <haijin.chn@gmail.com>	2026-04-13 15:07:04 +08:00
Jin Hai	cfc2928de2	Go: remove unused API route (#14028 ) ### What problem does this PR solve? As title ### Type of change - [x] Refactoring --------- Signed-off-by: Jin Hai <haijin.chn@gmail.com>	2026-04-10 18:00:41 +08:00
Jin Hai	3d59448b0d	Go: add parameter parsing of list chats (#14026 ) ### What problem does this PR solve? As title. ### Type of change - [x] New Feature (non-breaking change which adds functionality) Signed-off-by: Jin Hai <haijin.chn@gmail.com>	2026-04-10 14:33:32 +08:00
Jin Hai	a37605cbd2	Go: add get chat (#14025 ) ### What problem does this PR solve? As title ### Type of change - [x] New Feature (non-breaking change which adds functionality) --------- Signed-off-by: Jin Hai <haijin.chn@gmail.com>	2026-04-10 13:06:51 +08:00
chanx	4538910b52	feat: Implement file-related functionality (#14011 ) ### What problem does this PR solve? feat: Implement file-related functionality - Implement file deletion API and business logic - Add context support for file deletion operations and prevent root folder deletion - Implement file move functionality - Add File Download API Endpoints and Utility Functions ### Type of change - [x] New Feature (non-breaking change which adds functionality) --------- Co-authored-by: Yingfeng <yingfeng.zhang@gmail.com>	2026-04-10 12:15:27 +08:00
Jin Hai	cd04467b9b	Go: add delete search (#14014 ) ### What problem does this PR solve? As title. ### Type of change - [x] New Feature (non-breaking change which adds functionality) --------- Signed-off-by: Jin Hai <haijin.chn@gmail.com>	2026-04-10 09:42:37 +08:00
Jin Hai	5951e2b564	Go: Add create search (#13998 ) ### What problem does this PR solve? As title ### Type of change - [x] New Feature (non-breaking change which adds functionality) --------- Signed-off-by: Jin Hai <haijin.chn@gmail.com>	2026-04-09 20:04:06 +08:00

1 2 3

115 Commits