WeKnora
Tencent
Open-source LLM knowledge platform: turn raw documents into a queryable RAG, an autonomous reasoning agent, and a self-maintaining Wiki.
PROJECT TOPICS
INSTALL REFERENCE
dsh plugin --profile web add github:baldovinmarques391-design/dsh-plugin-glm-vision
该命令指向仓库当前默认分支;尚无绑定当前 commit 的完整验证结果。
PROJECT README
给 DSH (DeepSeek Harness) 加上"看图"能力的插件。
即使你用的模型本身不支持图片(比如纯文本的 DeepSeek),装了这个插件后,用户发送的图片会被自动转发给智谱 GLM-4V-Flash 视觉模型,生成的文字描述会注入对话中,让模型"看到"图片内容。
dsh plugin --profile web add git+https://github.com/baldovinmarques391-design/dsh-plugin-glm-vision.git
安装后需要手动将
dsh-plugin-glm-vision添加到 bundles 列表。编辑$DSH_HOME/profiles/web/package.json,在dsh.profile.bundles数组中加入"dsh-plugin-glm-vision":"dsh": { "profile": { "bundles": [ "@deepseek-ai/dsh-base", "@deepseek-ai/dsh-web-app", "dsh-plugin-glm-vision" ] } }
在 $DSH_HOME/.credentials.yaml 中添加:
GLM_API_KEY: 你的智谱API密钥
重启后,控制台应出现以下日志,表示插件加载成功:
dsh-plugin-glm-vision: loaded (model: GLM-4V-Flash, cache: 0 entries).
dsh-plugin-glm-vision: image translation layer active.
dsh-plugin-glm-vision: all modules active.
插件安装后会自动在 DSH 设置界面中显示配置项。也可通过 cordis.patch.yml 手动配置:
| 配置项 | 默认值 | 说明 |
|---|---|---|
glmApiKeyEnv |
GLM_API_KEY |
存放 API Key 的环境变量名 |
toolTimeoutMs |
120000 |
image_query 工具超时时间(毫秒) |
model |
GLM-4.1V-Thinking-Flash |
使用的 GLM 模型名 |
autoTranslate |
true |
是否自动翻译图片 |
enableTool |
true |
是否注册 image_query 工具 |
用户发送图片
↓
插件拦截消息,提取图片
↓
调用 GLM-4V-Flash API 生成图片描述
↓
用文字描述替换原始图片
↓
模型收到文字描述,正常回复
同时,插件会注册 image_query 工具,模型可以在任何时候主动调用来分析图片。
| 场景 | 结果 |
|---|---|
| 单张图片 | ✅ 正确识别并描述 |
| 两张图片同时发送 | ✅ 分别描述每张图片 |
| 旧对话中发新图片 | ✅ 描述图片 + 保持上下文 |
| 纯文本对话 | ✅ 正常工作,不影响 |
| 进程稳定性 | ✅ 无崩溃 |
MIT
CLASSIFICATION EVIDENCE
系统优先读取 GitHub Topics,再与站内分类词典和词根规则比对。当前命中: multimodal、vision。