Skip to content
WorkBench.ai

Model

DeepSeek V4.1 Flash evaluated on the hub's only benchmark

Released on 10 September 2026, DeepSeek V4.1 Flash (DeepSeek) joins the hub's rankings: its place on the Business Index, its strongest and weakest business functions, its cost and its hallucination rate.

Report written by the hub from the published rankings: every sentence is computed, none is an opinion.

  • 2nd of 27 on the Business Index, at 100.0% (95% margin of error: ± 0.0 points).
  • Mean cost per test: $0.0025, the 10th cheapest of the 27 models with a known price.
  • Mean hallucination rate: 0.0%, the share of items missing from the source for which it invented a value.
  • It took the hub's only published benchmark.

Scores by business function

The model's mean accuracy on the benchmarks it took in each business function.

Business functionScore
FinanceNot tested
Accounting100.0%
Human resourcesNot tested
LegalNot tested
SalesNot tested
MarketingNot tested
Customer serviceNot tested
Procurement & logisticsNot tested
ITNot tested
ManagementNot tested
Business Index100.0%