Models / Moonshot AI
Kimi K3
by Moonshot AI ·
moonshotai/kimi-k3Kimi K3 is a 2.8T parameter open-weight multimodal reasoning model from Moonshot AI. It is suited for complex coding, knowledge work, and long-horizon agentic workflows, and is particularly strong at...
reasoningtool-usevisionreleased 2026-07-16
Legal score · raw
-
not yet benchmarked
Context
1.05M
max output 944K
Input
$1
per 1M tokens
Output
$12.50
per 1M tokens
Suite cost
-
no run recorded
Benchmark results
This model has not been run through the suite yet. See Humanity's Last Lawsuit for the benchmark models are ranked on.
Measured by DocketBuster
These numbers come from DocketBuster's own legal battery, not from DocketRouter's suite. Latest run per metric, with n and a 95% Wilson interval where the source reports one. See docketbuster.com/benchmarks.
| Metric | Value | n | Interval | Measured |
|---|---|---|---|---|
| Statute pinpoint, exact section (no retrieval) | 25.0% (75/300) | 300 | 95% CI 20.4% to 30.2% | 2026-08-22 |
| Say-nothing rate (declines to bluff when the answer is not in the record) | 100.0% | 58 | 95% CI 93.8% to 100.0% | 2026-08-22 |
| Abstained on statute pinpoint | 47.7% (143/300) | 300 | count, no interval reported | 2026-08-22 |
| Coaching quality (rubric, 0 to 8) | 7.36 / 8 | - | rubric mean, no interval reported | 2026-08-23 |