dshbase

Plugin directory / Developer / dsh-plugin-mlquant-benchmark

dsh-plugin-mlquant-benchmark

Unverified initial-d

✓ Actively maintained Builds on 1 official DSH packages

View on GitHub ↗ ← Back to plugin directory

2Stars
0Forks
0Open issues
Language
2026-08-24Last push
Cross-platformPlatform

What it does

DeepSeek Harness tools for reproducing the ml-quant-trading protocol v1 benchmark.

Our take
Unverified — not yet verified

DeepSeek Harness tools for reproducing the ml-quant-trading protocol v1 benchmark. Not yet verified — install and test it yourself.

“Unverified” means our automated CI has not yet installed this plugin. Feature descriptions and version compatibility are the author’s claims. This is not a security audit and not an endorsement of third-party code.

Plugin author? Get the “Verified” label — submit your own evidence (screenshots, logs, or a short demo) and we'll review and flip the badge.

Submit verification evidence ↗

README

dsh-plugin-mlquant-benchmark

CI
Listed on Awesome DSH Plugin

DeepSeek Harness tools for reproducing the
initial-d/ml-quant-trading
protocol v1 CPU benchmark.

The point is narrow: make a DSH agent able to run the existing benchmark, read
the machine-readable artifact, validate it against the benchmark protocol, and
draft an issue-ready report. This plugin does not add a trading agent, does not
call market data APIs, and does not configure any model provider.

Why this exists

ml-quant-trading is a good reproducibility target for agent harnesses:

  • deterministic synthetic benchmark input;
  • fixed protocol v1 command, seed, panel size, repetitions, and thread counts;
  • JSON artifact suitable for automated checking;
  • public issue template for DeepSeek Harness benchmark reports;
  • explicit boundary that benchmark throughput is not trading performance.

Challenge: can DeepSeek Harness reproduce a quant benchmark end to end, preserve
the evidence bundle, and avoid turning runtime numbers into alpha claims?

Listed in
awesome-dsh-plugin
via PR #2573.

Run-To-Report Path

  1. Install the plugin from GitHub.
  2. Open an initial-d/ml-quant-trading checkout in DSH.
  3. Ask DSH to run, validate, summarize, and draft a benchmark report.
  4. Submit the drafted report through the dedicated issue template.

That path is intentionally small: the plugin turns DSH attention into a
reproducible benchmark report, not an investment or leaderboard claim.

Tools

This package registers four DSH tools:

Tool Purpose
mlquant_benchmark_v1_cpu Run the fixed protocol v1 CPU benchmark and write artifacts/benchmark-v1.json.
mlquant_read_benchmark_json Read the JSON artifact and render a compact Markdown result table.
mlquant_validate_benchmark_json Check protocol v1 fields, expected cases, fixed parameters, and variance warnings.
mlquant_draft_github_issue Draft a DeepSeek Harness benchmark issue body from the JSON artifact. It does not post to GitHub.

Install

Install the package in a DeepSeek Harness profile or preset environment:

dsh plugin --profile web add github:initial-d/dsh-plugin-mlquant-benchmark

The package declares a dsh.bundle manifest that inserts:

- id: mlquant-benchmark
  name: dsh-plugin-mlquant-benchmark

If you use a local checkout while developing, add the same row manually:

- id: mlquant-benchmark
  name: file:/path/to/dsh-plugin-mlquant-benchmark

This package is intentionally not published to npm yet. GitHub distribution is
enough for the first DSH-facing benchmark reports; npm can come later if there
is real usage.

Suggested DSH prompt

Read AGENTS.md, docs/benchmarking.md, and docs/reality_check.md.
Use the mlquant benchmark tools to run the protocol v1 CPU benchmark, validate
and read the JSON artifact, and draft a DeepSeek Harness benchmark report. Keep
the result as an engineering reproducibility benchmark, not a trading-performance
claim.

Public report path

Post the drafted report through the main repository's dedicated template:

https://github.com/initial-d/ml-quant-trading/issues/new?template=deepseek_harness_benchmark.yml

Seed example:

https://github.com/initial-d/ml-quant-trading/issues/61

For context and agent-facing guardrails, read the main repository's
DeepSeek Harness Recipe
and
Quant Agent Reproducibility Target.

Development

npm install
npm test

The test loads the plugin with a mock ctx.tools.register, verifies that the
four tools register, reads and validates sample artifacts, and drafts an issue
body.

Non-goals

  • No investment advice.
  • No backtest-performance claim.
  • No hidden model provider configuration.
  • No posting to GitHub from the tool.
  • No private data or API keys in artifacts.

Install

🧩 Let your agent install it (recommended)

Install the catalog once, then DeepSeek Harness can find and install any plugin from this site automatically:

dsh plugin add dshbase-catalog

Then say "install dsh-plugin-mlquant-benchmark for me" — your agent finds it in the directory and installs it. Docs: dshbase-catalog · verified packs.

This plugin is GitHub source (not published to npm) — install it straight from the repo:

Web profile:

dsh plugin --profile web add github:initial-d/dsh-plugin-mlquant-benchmark

Headless (CLI) profile:

dsh plugin --profile headless add github:initial-d/dsh-plugin-mlquant-benchmark

Test report

Not yet L3-verified — see failure note below if we already ran it.

Status: pending · last test 2026-08-25
Note: 验证: runtime-fail Browse all pending failures →
Security: not yet scanned — our daily static scan will cover it shortly.

Share this badge

More in Developer

Browse all 7789 plugins →