improved_vision_for_deepseek
zyh20041227
Full-coverage image tiling for DeepSeek Harness vision models, dense-text OCR, and document AI
DSH-PLUGIN STORE / LIVE CATALOG
聚合 GitHub 上的 DSH 插件,打造 DeepSeek Harness 生态的一站式目录。
179 个项目,匹配「vision」
zyh20041227
Full-coverage image tiling for DeepSeek Harness vision models, dense-text OCR, and document AI
1HelloMan1
DeepSeek Harness usage dashboard with API balance, daily spend, external vision-call accounting, per-model stats, call logs, cache rate, TTFT, and CSV export.
3361805598-gif
DSH Markdown sidebar viewer and editor with block- and text-range annotations for structured revision requests.
Altairpaca
Windows Computer Use for DeepSeek Harness (DSH): window-bound screenshot/OCR/click with verification loop, pure-OCR mode, pluggable vision models.
Einskyle
DeepSeek vision bridge for dsh: route image attachments to a vision model (Qwen3-VL via pi-ai/llama.cpp) and continue on a text-only LLM (DeepSeek)
Favio8
DeepEye vision plugin for DeepSeek Harness (DSH): image description, OCR, VQA, UI layout, and clipboard analysis.
MC5lan
给 DeepSeek 安装一双眼睛和一支画笔:会话里直接贴截图/图片,GLM 视觉模型先精确转写图片内容(报错信息、代码、界面逐字保留),然后 DeepSeek 继续处理你的问题——同一轮完成,全程无感;需要配图时,DeepSeek 自动调用文生图后端出图并显示在会话中。
Spirit4471
Multimodal bridge for any LLM: qwen_vision (Qwen-VL) + qwen_generate (Qwen-Image) — an MCP server for Kimi Code & Claude Code, and a DeepSeek Harness plugin (dsh-multimodal-bridge). 给任何 LLM 补上 Qwen 视觉与出图能力:MCP server(Kimi Code / Claude Code)+ DeepSeek Harness 插件双形态。
Terry12138qy
DeepSeek Harness 识图插件:为不具备原生识图能力的模型提供识图能力(阿里云百炼 qwen3.5-omni-plus,失败自动切换智谱 glm-4.6v-flash)。由 claude-vision-skill 移植适配。 | Vision tool for DeepSeek Harness
b8yg7vjstj-ctrl
DeepSeek Harness plugin: local llama.cpp (llama-server) as a first-class model provider — process/router management, catalog sync with mmproj vision pairing, auto-start, stop-then-load switching, and a sidebar terminal monitor panel.
beihzb
Native Jupyter-style notebook for DeepSeek Harness: real ipykernel sidecar + VS Code-aligned cell UI, tqdm progress, inline figures, per-cell AI revision.
beijingwahw
Vision-only desktop automation agent plugin for DeepSeek Harness (DSH) | 纯视觉桌面自动化 Agent 插件:SoM grounding · Planner-Actor · effect verification · skill library
c-ling
【已停止维护】DeepSeek Harness 视觉桥接插件:新版 Harness 已原生支持图片识别,请使用原生能力,本仓库仅作历史参考。
datit309
Full 2-way Telegram Remote Control, Vision, Voice, Shell, Diff & Sound Notifications for DeepSeek Harness (DSH)
dongsheng123132
Deterministic revision-pinned benchmarks and regression evidence for DeepSeek Harness
haiziyao
Vision routing and image generation for DeepSeek Harness through a fixed Mix model.
jin123-alpha
External vision proxy extension for DeepSeek Harness, enabling text-only models to analyze and understand images via OpenAI-compatible vision APIs.
libinyam
Config-only DeepSeek Harness bundle for OpenAI-compatible vision models.
limccn
Give DeepSeek (text-only) models **vision** in Claude Code and Codex by routing image files to any OpenAI-compatible vision endpoint (OpenRouter, SiliconFlow, DashScope, Ollama, llama.cpp, vLLM, LM Studio, …). Zero runtime dependencies, MIT licensed.
meimiaoji-creator
meow-vision 是 DeepSeek Harness 的一款视觉插件,解决纯文本模型无视觉。另一方面vue页面开发视觉验证不闭环的问题。
moduqishi
给纯文本大模型装上原生视觉:流式真实思考链 · 跨轮次无感重看 · 像素级证据与 SVG 图元 · OpenAI/Anthropic/Responses 三协议兼容 | Native vision for text-only LLMs: streaming real thinking chain, cross-turn re-view, pixel-level evidence & SVG primitives, OpenAI/Anthropic/Responses compatible.
moon09300731
DeepSeek Harness 视觉能力全家桶:vision_understand 工具 + 粘贴/拖拽/按钮三入口识图
moton16
Give your text-only LLM eyes: zero-dependency MCP server that reads images via external OpenAI-compatible vision models with automatic provider fallback, plus a one-command patch for DSH image-attachment degradation.
secretxuan
Native-vision Windows computer-use for DeepSeek Harness: screenshots, UIA, OCR, and approval-gated input