DSH-PLUGIN STORE / LIVE CATALOG

DSH插件商店

聚合 GitHub 上的 DSH 插件,打造 DeepSeek Harness 生态的一站式目录。

已收录
6553
功能分类
11
更新时间
10/07 02:01

7 个项目,匹配「llm-eval」

开发工具 技能

ouroboros

Q00

Agent OS: the agent gets smarter on its own. We just hold the line: Interview-gated, staged evaluation, budgeted evolution loop. MCP server, 14 runtimes: Claude Code, Codex CLI, Gemini CLI, OpenCode, Copilot, Kiro and more.

agent-os agentic-ai ai-agent ai-coding-agent
Agent 与会话 技能

SkillCorpus

EverMind-AI

Open-source infrastructure that turns scattered SKILL.md files into curated, retrieval-ready agent-skill corpora—with retrieval and evaluation tooling included.

agent-memory agent-skills ai-agents benchmark
文件与数据 插件

dsh-design-qa

sunxin-ai

Design-fidelity QA for DeepSeek Harness: lend any text-only model an eye, then judge whether the implementation matches the mock. Ships the benchmark behind that judgement — four fixtures, 23 injected defects, and every raw model transcript. Retires itself when DeepSeek ships vision.

benchmark deepseek-harness design-qa design-review
开发工具 技能

oh-my-knowledge

lizhiyao

OMK — Evidence-backed evaluation and observability for prompts, RAG, skills, agents, and workflows. Native Codex, Claude Code, and DeepSeek Harness support.

agent-evaluation ai benchmark bootstrap-ci
学习研究 插件

dsh-plugin-evaluation-standards

dsh-plugin-evaluation

Open evaluation datasets, test cases, and metrics for DSH plugins.

benchmarks deepseek-harness deepseek-harness-plugin dsh
部署运维 插件

Controlled request-surface replay and regression workbench for DeepSeek Harness

agent-observability deepseek deepseek-harness dsh
Agent 与会话 插件

Visual workflows and multi-model evaluation for DeepSeek Harness

agent-workflow deepseek deepseek-harness dsh-plugin