Skip to content
WorkBench.ai

Model

Kimi K3 evaluated on the hub's only benchmark

Released on 16 July 2026, Kimi K3 (Moonshot AI) joins the hub's rankings: its place on the Business Index, its strongest and weakest business functions, its cost and its hallucination rate.

Report written by the hub from the published rankings: every sentence is computed, none is an opinion.

  • 6th of 27 on the Business Index, at 99.8% (95% margin of error: ± 0.7 points).
  • 0.2 index points below Kimi K2.6, its predecessor at Moonshot AI, released on 20 April 2026. The gap is smaller than the margin of error: the two models cannot be told apart.
  • Mean cost per test: $0.021, the 22nd cheapest of the 27 models with a known price.
  • Mean hallucination rate: 1.0%, the share of items missing from the source for which it invented a value.
  • It took the hub's only published benchmark.

Scores by business function

The model's mean accuracy on the benchmarks it took in each business function.

Business functionScore
FinanceNot tested
Accounting99.8%
Human resourcesNot tested
LegalNot tested
SalesNot tested
MarketingNot tested
Customer serviceNot tested
Procurement & logisticsNot tested
ITNot tested
ManagementNot tested
Business Index99.8%