docketrouter
Models / Moonshot AI

Kimi K3

by Moonshot AI · moonshotai/kimi-k3

Kimi K3 is a 2.8T parameter open-weight multimodal reasoning model from Moonshot AI. It is suited for complex coding, knowledge work, and long-horizon agentic workflows, and is particularly strong at...

reasoningtool-usevisionreleased 2026-07-16
Legal score · raw
-
not yet benchmarked
Context
1.05M
max output 944K
Input
$3
per 1M tokens
Output
$15
per 1M tokens
Suite cost
-
run the suite to see

Benchmark results

TaskCategoryRawJuicedCorrectLatencyCostRan
Hearsay IdentificationEvidence------
Bluebook Citation FormatResearch & Writing------
Federal Civil ProcedureProcedure------
Limitations ArithmeticProcedure------
Contract Clause ClassificationContracts------
Citation Hallucination ResistanceReliability------

Measured by DocketBuster

These numbers come from DocketBuster's own legal battery, not from DocketRouter's suite. Latest run per metric, with n and a 95% Wilson interval where the source reports one. See docketbuster.com/benchmarks.

MetricValuenIntervalMeasured
Statute pinpoint, exact section (no retrieval)25.0% (75/300)30095% CI 20.4% to 30.2%2026-08-22
Say-nothing rate (declines to bluff when the answer is not in the record)100.0%5895% CI 93.8% to 100.0%2026-08-22
Abstained on statute pinpoint47.7% (143/300)300count, no interval reported2026-08-22
Coaching quality (GW-14x, 0 to 8)7.36 / 8-rubric mean, no interval reported2026-08-23

Source files: hard-llm-ortier-kimik3.json, gw14x-ortier-kimik3.json. Raw model name in source: moonshotai/kimi-k3.