Benchmark data
Download the canonical hibench export tables used by the dashboard. Files are generated fromhibench export and copied into/data/ on each site build.
Benchmark data last updated July 7, 2026 UTC
Dataset overview
- Schema version
- hibench.benchmark.v1
- Coverage
- 16 agents ยท 1132 captured versions
- License
- MIT
- Update cadence
- Refreshed when benchmark captures are exported and the site is rebuilt
See the methodology for capture rules, tokenizer notes, and how primary requests are selected. Source repository:hibenchmark/hibench.
Downloads
One row per canonical primary capture: agent, version, tokenizer totals, tool/skill/MCP/sub-agent counts, and capture metadata.
Declared tools per primary request with definition size, type, and MCP/sub-agent flags.
Bundled skills per run with token cost and preview text.
Declared sub-agent types per run with token cost and source classification.
MCP server declarations per run when present.
Tokenized text-field breakdown (instructions, injected context, user prompt, tool context, and related categories).
Export manifest with schema version, row counts, and file inventory.
Optional GitHub star metadata for agents with public repository links.