image-analysis-skill
SKL-666666
图片结构化分析技能:双引擎OCR+形状/表格/图标/布局识别,让纯文本模型看懂图片
DSH-PLUGIN STORE / LIVE CATALOG
聚合 GitHub 上的 DSH 插件,打造 DeepSeek Harness 生态的一站式目录。
57 个项目,匹配「ocr」
SKL-666666
图片结构化分析技能:双引擎OCR+形状/表格/图标/布局识别,让纯文本模型看懂图片
Su2uka111
Document drag-and-drop plugin for DeepSeek Harness: PDF/Word/Excel/text auto-parsed with dual-track context (inline + doc_read), native WinRT OCR for scans. Zero upstream changes, hot-injectable. | DeepSeek Harness 文档拖拽读取插件:PDF/Word/Excel/文本自动解析,双轨上下文(内联 + doc_read 按需检索),扫描件原生 WinRT OCR。零修改上游,运行时热注入。
aijunjiang
Give your DSH agent eyes via any OpenAI-compatible vision model - 11 provider presets (Doubao/Qwen-VL/GLM-V/OpenAI/Gemini/Ollama...), capability checkboxes that inject live prompt guidance, and analysis of images the user drops into the chat; the agent writes its own observation prompt, and base64 never enters its context.
jmjmj009gt
Zero-dependency vision OCR/Q&A toolkit (CLI + local web GUI) for OpenAI-compatible VLMs: Zhipu GLM, Qwen, OpenAI, OpenRouter, SiliconFlow
princefrogdida-ux
Windows-first vision suite with image understanding, OCR, screenshot diffing, and multi-provider routing for DeepSeek Harness.
baimaomaomao556
EagleEye MCP — pixel-accurate visual toolbox for Agents (screenshot, measure, OCR, regression)
leozou320-ai
Offline macOS Vision OCR for DeepSeek Harness — accurate, local, API-key free. | DeepSeek Harness 本地离线 OCR 插件
protoctistmoses143
Convert PDFs, Office docs, scanned images, and more to clean Markdown, JSON, or text locally with offline OCR—no servers, no API keys, fully private.
qizhen2021
该仓库暂未提供项目说明。