DSH-PLUGIN STORE / LIVE CATALOG

DSH插件商店

聚合 GitHub 上的 DSH 插件,打造 DeepSeek Harness 生态的一站式目录。

已收录
4403
功能分类
11
更新时间
08/18 02:30

15 个项目,匹配「benchmark」

学习研究 技能

DeepSeek V4 × J-Space capability realization report — benchmark evidence that J-Space reduces capability-realization loss on DeepSeek V4 (Flash/Pro).

agent-skills ai-agent benchmark deepseek
模型与 MCP 插件

flameox

morluto

Runtime evidence that helps agents trace, profile, and burn down hotspots in application and native code, GPU kernels, and inference stacks.

benchmarking coding-agents cordis debugging
安全与治理 插件

openguardrails

openguardrails

The vendor-neutral protocol for AI agent safety & security — and the neutral benchmark that ranks the vendors.

agents ai-safety ai-security dsh-plugin
其他 插件

Continual self-evolution plugin for DeepSeek Harness: versioned, auditable, rollback-safe harness state refined from session trajectories, with a benchmark-driven validation loop.

ai-agent deepseek-harness deepseek-harness-plugin dsh
学习研究 插件

dsh-excel-chat

hccccc01333

dsh-excel-chat — talk to Excel in DeepSeek Harness: create, edit, repair, and verify spreadsheets by conversation (cells, formulas, styles, filters, tables, charts); every edit is auto-validated.

agent benchmark deepseek-harness dsh-plugin
开发工具 插件

DSH-arena

Apageoflove

Local-first experiment and evaluation workbench plugin for DeepSeek Harness (DSH).

ai-agent arena benchmarking cordis
Agent 与会话 插件

dsh-ops-kit

LeslieWylie

A reusable DeepSeek Harness bundle for evidence-driven memory, orchestration, benchmark operations, and plugin release workflows.

agent-tools deepseek-harness dsh dsh-plugin
开发工具 插件

dsh-benchmark

dongsheng123132

Deterministic revision-pinned benchmarks and regression evidence for DeepSeek Harness

ai-agent benchmark deepseek-harness dsh
开发工具 插件

smokinggun

morluto

Help your agents find the smoking gun they're looking for. Optimization evidence for agents: find complexity hotspots.

ai-agents benchmarks code-optimization code-quality
其他 待识别

Benchmark-driven self-evolution for DeepSeek Harness · 冻结基准上的 Agent Profile 自我进化:评测 → 候选 → 严格接受/回滚

dsh dsh-plugin
Agent 与会话 插件

dsh-plan-lattice

1052326311

Execution-time drift firewall for long-running DeepSeek Harness agents. Real-Harness tests: unsafe stale mutations 12/12 native -> 0/12; valid controls 7/7 both; post-SIGKILL unsafe continuation 2/2 -> 0/2.

agent-harness agent-orchestration agent-planning agent-safety
学习研究 插件

DSH plugin: run a command N rounds, judge by median/distribution — 批量回归取统计结论

benchmarking deepseek-harness dsh-plugin regression-testing
部署运维 插件

dsh-budget

PerryLink

Cost governance for DeepSeek Harness: aggregated token/cost metering per model, session and day, budget caps with threshold alerts and over-limit policies, carbon footprint estimation, per-model latency benchmarks, a Settings budget tab, and the /budget command

budget carbon-footprint cordis cost-tracking
学习研究 插件

dsh-eval

hccccc01333

Agent evaluation platform for DeepSeek Harness: benchmark YAML, headless dsh orchestration, trace-based metrics, LLM judge, paired A/B, keyless replay, and cross-harness import.

agent-evaluation benchmark deepseek-harness dsh
Agent 与会话 插件

dsh-swarm-router

r600a-code

DSH plugin: sub-agent matrix swarm — routes heterogeneous tasks to the most suitable model (OpenRouter-like + cfgpu.com/llm/square), dispatches each via in-process subagents. 32/32 benchmark green.

cfgpu deepseek-harness dsh-plugin llm-routing