文件与数据
技能
The first vision plugin for DeepSeek Harness, and the vision bridge for every text-only coding agent. Paste an image, get structured JSON evidence (OCR, layout, semantics). | 全网最强 DeepSeek Harness 外挂视觉插件,为 DeepSeek、GLM 等纯文本模型外挂视觉能力,粘贴图片即得结构化 JSON 证据(OCR、版面、语义)。
agent-skills
claude-code
claude-skills
codex
文件与数据
技能
为纯文本模型"看图“设计更好的视觉工具箱和技能,支持多图理解,图片问答,前端UI还原、GUI 自动化等,并可选无缝接入多个主流agent,直接识别粘贴图片| A vision toolkit and skill designed for text-only llms — image Q&A, long-screenshot OCR, frontend UI restoration, and GUI automation, with optional seamless integration for Codex, Claude Code, Pi, Oh My Pi, and OpenCode
agent
agent-skills
claude-code
codex
文件与数据
插件
TongFlow — multimodal workflow studio and engine (canvas + Python plugin engine) and dsh-tongflow, the DeepSeek Harness studio plugin
3d
agent
ai
ai-tools
文件与数据
插件
Eyes for text-only DeepSeek Harness agents: built-in free vision chain (no key) + pixel-level vision tools (Q&A, grounding, crop, pixel diff, colors, OCR, SVG trace, cutout, screenshots). One-command install, no Python, image turns work like ordinary tool-calling turns.
deepseek-harness
dsh
dsh-plugin
multimodal
文件与数据
插件
Near-native image understanding for DeepSeek Harness
deepseek-harness
dsh-plugin
image-understanding
multimodal
文件与数据
插件
给 DeepSeek 补上「眼睛和耳朵」的多模态视觉插件:看图 / OCR / 物体检测 / 视频理解 / 语音转写 / 截图直读,一键安装(DSH 插件)。
dashscope
deepseek
dsh-plugin
mcp
文件与数据
插件
DeepSeek Harness 插件:DeepSeek 大脑 + 自动识图。GUI 附加图片自动经 OpenAI 兼容 VLM 转译成文字后交给 DeepSeek 作答;支持百炼/智谱/OpenRouter 等任意 OpenAI 兼容端点(默认 qwen3.7-flash),无 key 自动探测本地 Ollama(图片不出本机);安装时有一问式确认
dashscope
deepseek-harness
dsh-plugin
image-understanding
文件与数据
插件
Dsh-visual-plugin.Give your text-only model eyes: forward user images to any OpenAI-compatible vision model and see the results in a Web UI right panel
deepseek
deepseek-harness
dsh-plugin
image-description
文件与数据
插件
Vision for DeepSeek Harness: Doubao Web by default (zero-cost, no API key), Antigravity IDE quota (flash/pro), any IDE CLI, Gemini — auto detail escalation, evidence memory
antigravity
doubao
dsh
dsh-plugin
文件与数据
完整应用
Vision-language gateway plugin for DeepSeek Harness - paste an image, DeepSeek sees text
coding-agent
deepseek
deepseek-harness
dsh
文件与数据
插件
DeepSeek Harness plugin: describe_image — give a text-only model vision through an OpenAI-compatible VLM endpoint
deepseek
deepseek-harness
describe-image
dsh
文件与数据
插件
给 DeepSeek 安装一双眼睛和一支画笔:会话里直接贴截图/图片,GLM 视觉模型先精确转写图片内容(报错信息、代码、界面逐字保留),然后 DeepSeek 继续处理你的问题——同一轮完成,全程无感;需要配图时,DeepSeek 自动调用文生图后端出图并显示在会话中。
deepseek-harness
dsh
dsh-plugin
dsh-plugins
文件与数据
插件
一个工具 = MiniMax 全部多模态能力:DSH 纯文本模型看图/画图/生视频/说话/唱歌/翻唱/搜索/查额度 | One mmx_bridge tool = all MiniMax multimodal (VLM/image/video/speech/music/cover/search/quota) for DeepSeek Harness (DSH)
agent-tool
ai-agent
cordis
deepseek-harness
文件与数据
插件
DeepEye vision plugin for DeepSeek Harness (DSH): image description, OCR, VQA, UI layout, and clipboard analysis.
cordis
deepseek-harness
dsh
dsh-plugin
文件与数据
插件
该仓库暂未提供项目说明。
deepseek
deepseek-harness
dsh
dsh-plugin
文件与数据
插件
Paste images into DeepSeek Harness with a four-model vision race, OCR, and an automatic text bridge.
deepseek
deepseek-harness
dsh-plugin
image-to-text
文件与数据
插件
Config-only DeepSeek Harness bundle for OpenAI-compatible vision models.
deepseek-harness
dsh
dsh-plugin
multimodal
文件与数据
插件
Image understanding, OCR, and persistent visual evidence for text-only DeepSeek Harness models
ai-agents
computer-vision
deepseek
deepseek-harness
文件与数据
插件
DeepSeek Harness 视觉桥接插件:上传/粘贴图片,通过 OpenAI 兼容视觉 API 生成英文描述,再交给纯文本模型。
agent
ai
attachment
chat
文件与数据
插件
Native vision, gpt-image-2 and Seedance plugins for DeepSeek Harness via Xiapan Cloud
deepseek-harness
dsh-plugin
image-generation
multimodal
文件与数据
插件
给纯文本大模型装上原生视觉:流式真实思考链 · 跨轮次无感重看 · 像素级证据与 SVG 图元 · OpenAI/Anthropic/Responses 三协议兼容 | Native vision for text-only LLMs: streaming real thinking chain, cross-turn re-view, pixel-level evidence & SVG primitives, OpenAI/Anthropic/Responses compatible.
ai-proxy
claude-code
codex
deepseek
文件与数据
插件
DSH plugin: text-only models (e.g. DeepSeek-V4) automatically see images via a vision model. Official surface-replace, cache-friendly, human transcript untouched. 纯文本模型自动识图桥
deepseek-harness
dsh-plugin
multimodal
vision
文件与数据
插件
Vision sidecar for DeepSeek Harness: accept pasted images on text-only models.
deepseek-harness
dsh-plugin
vision
文件与数据
插件
Native interactive visual-reasoning plugin for DeepSeek Harness: precise pixel grounding (SOM grid / zoom / annotate / measure / diff / color / OCR) + MiMo V2.5 multimodal backend, zero external MCP servers.
ai-agents
computer-vision
cordis
deepseek-harness