DSH-PLUGIN STORE / LIVE CATALOG

DSH插件商店

聚合 GitHub 上的 DSH 插件,打造 DeepSeek Harness 生态的一站式目录。

已收录
6553
功能分类
11
更新时间
10/07 02:01

80 个项目,匹配「ocr」

文件与数据 插件

dsh-pseudo-vision

DDDFXYqiming

Local OCR, color-statistics, pixel-scan, and metadata bridge for text-only DeepSeek Harness models; no external vision API.

deepseek-harness dsh-plugin ocr vision
文件与数据 技能

dsh-vision-skill

DDDFXYqiming

Vision skill plugin for DeepSeek Harness (image analysis and OCR)

agent-skills deepseek-harness dsh dsh-plugin
模型与 MCP 插件

Low-token, low-latency Windows computer-use MCP with learned shortcuts, UIA/CDP/OCR routing, and DeepSeek Harness support

computer-use deepseek-harness dsh-plugin mcp
模型与 MCP 插件

auto-mouse

Fish121380

Windows desktop UI context picker for AI agents: select windows, UI elements, or screen regions with hover highlighting, UI Automation, screenshots, local OCR, and user-approved MCP output. Works with OpenAI Codex, DeepSeek Harness, and other MCP-compatible clients.

accessibility agent-tools ai-agents codex-plugin
文件与数据 插件

dsh-vision-bridge

GooDAnDReaDY

Universal vision bridge for DeepSeek Harness: attachments with native models, 40+ tools, PDF/OCR/diagrams.

deepseek-harness dsh dsh-plugin multimodal
开发工具 完整应用

dsh-desktop

Plocr

DeepSeek Harness 桌面工作台:Electron 原生壳 + 内嵌 harness 运行时(离线、免装 Node),壳仅保留桌面原生能力,与 harness 间以 bridge 插件通信

deepseek-harness desktop-app dsh dsh-plugin
界面增强 技能

dsh-skill-manager

ZBCs-StudioCr-CN

DeepSeek Harness 可视化侧边栏 Skill 管理器:启用/禁用、默认加载、分类管理、智能归类、SKILL.md 可视化编辑——无需手动操作文件

ai-tools deepseek-harness dsh dsh-plugin
文件与数据 插件

glm4v-vision-mcp

ethanweave

GLM-4.6V 图像理解 MCP:识图/OCR/图表解析,原生接入 DeepSeek Harness(dsh-mcp-client),也兼容 Codex/Cline 等

codex dsh-plugin glm image-understanding
文件与数据 插件

DeepSeek Harness vision plugin: analyze_image (structured OCR evidence) + capture_image (USB camera visual loop). 摄像头视觉闭环 + 结构化证据,支持 Ollama / DeepSeek / Xiaomi 三后端。

camera deepseek-harness dsh-plugin image-to-text
其他 待识别

dsh-paddle-ocr

omdsh-dev

DSH plugin for PaddleOCR-VL document layout parsing: convert PDFs and images to Markdown with async jobs, progress tracking, and workspace export.

dsh-plugin
文件与数据 插件

dsh-md-convert

yakoylp

Convert Office documents and PDFs (incl. scanned, via CPU-first routing OCR with lightweight models: PP-DocLayout-L layout, RapidOCR text, SLANet tables, FormulaNet formulas) to structurally-formatted Markdown. CLI + dsh agent tool (md_convert).

cli cordis deepseek-harness document-conversion
文件与数据 插件

dsh-vision-primitives

zouyuanqing

Native interactive visual-reasoning plugin for DeepSeek Harness: precise pixel grounding (SOM grid / zoom / annotate / measure / diff / color / OCR) + MiMo V2.5 multimodal backend, zero external MCP servers.

ai-agents computer-vision cordis deepseek-harness
其他 插件

dsh-wsl-media

173787247

Local media/doc pipeline: ffprobe, extract, thumbnail, PDF, ASR, pandoc, OCR, exif.

deepseek-harness dsh-plugin wsl
文件与数据 插件

dsh-eye-vision

AlloyPlane

该仓库暂未提供项目说明。

ai deepseek-harness dsh dsh-plugin
文件与数据 插件

dsh-vision-analysis

Harvey-Will

DeepSeek Harness 图像理解插件 · 8 种分析模式 · 支持任意API接口 · 内置免费视觉模型 | DeepSeek Harness vision plugin · 8 analysis modes · works with any OpenAI- or Anthropic-compatible API · built-in free vision model

deepseek deepseek-harness dsh-plugin free
文件与数据 插件

easy-vision

Koreyer

A DeepSeek Harness tool plugin that lets text-only agents "see" local images — auto-detects the real format and returns a detailed text description via any OpenAI-compatible vision model.

cordis-plugin deepseek-harness dsh-plugin image
文件与数据 插件

dsh-eyes

Leeminjing

Give text-only DeepSeek models on-demand vision: upload images, DeepSeek answers by calling a view_image tool backed by any OpenAI-compatible vision endpoint (Qwen/DashScope by default).

dashscope deepseek-harness dsh-plugin multimodal
文件与数据 插件

vision-exp-tile

Nicholaskin

DSH(DeepSeek Harness)插件:800×800 无损切块 + 坐标标注 + 直连 DeepSeek 视觉 API 识别聚合,专治大图看不清

deepseek deepseek-harness dsh dsh-plugin
消息通讯 渠道适配

Chat with, monitor, and approve your DSH (DeepSeek Harness) agents from WeChat over the clawbot iLink gateway: two-way text/images/voice/files/video, native vision or OCR, context-rotation policies, reminders, and a standalone admin console.

agent bridge chat-bot cordis
文件与数据 技能

图片结构化分析技能:双引擎OCR+形状/表格/图标/布局识别,让纯文本模型看懂图片

agent-skills deepseek-harness dsh-plugin image-analysis
文件与数据 插件

dsh-document-drop

Su2uka111

Document drag-and-drop plugin for DeepSeek Harness: PDF/Word/Excel/text auto-parsed with dual-track context (inline + doc_read), native WinRT OCR for scans. Zero upstream changes, hot-injectable. | DeepSeek Harness 文档拖拽读取插件:PDF/Word/Excel/文本自动解析,双轨上下文(内联 + doc_read 按需检索),扫描件原生 WinRT OCR。零修改上游,运行时热注入。

deepseek-harness dsh dsh-plugin ocr
模型与 MCP 插件

shadow-vision

WardLu

Open-source MCP vision server that gives text-only LLMs and AI agents image understanding, OCR, visual analysis, UI inspection, and multimodal capabilities.

ai-agents computer-vision dsh-plugin image-analysis
文件与数据 插件

dsh-plugin-vision

aijunjiang

Give your DSH agent eyes via any OpenAI-compatible vision model - 11 provider presets (Doubao/Qwen-VL/GLM-V/OpenAI/Gemini/Ollama...), capability checkboxes that inject live prompt guidance, and analysis of images the user drops into the chat; the agent writes its own observation prompt, and base64 never enters its context.

deepseek-harness doubao dsh dsh-plugin
文件与数据 插件

Zero-dependency vision OCR/Q&A toolkit (CLI + local web GUI) for OpenAI-compatible VLMs: Zhipu GLM, Qwen, OpenAI, OpenRouter, SiliconFlow

deepseek-harness dsh-plugin glm llm