Skip to content

model

Kimi-K3

Moonshot frontier model pulled roughly 60 times per like, the attention-heavy counterpart to all-MiniLM-L6-v2.

Known aliases

  • global.moonshotai.kimi-k3
  • K3
  • KIMI K3
  • Kimi-K3
  • Kimi K3 2.8T
  • Kimi K3 (max)
  • Kimi K3 open weights
  • Moonshot AI Kimi K3
  • moonshotai.kimi-k3
  • us.moonshotai.kimi-k3

Relationships

No evidence-backed relationships are recorded.

Current stories

build2 publishers

Two-thirds of failed agent runs in ThinkingBox exited cleanly with the backend still wrong

Microsoft and Hugging Face's ThinkingBox found that 67.24% of 79,853 failed agent runs ended cleanly, with no final tool error. Those failures showed up only when executable checks read the records each run left in the backend.

Perspective Coverage

3 publishers
Builder
Builder 52%
Operator
Operator 38%
Investor
Investor 10%

Reality

Evidence60
Adoption
Insufficient
Hype gap+10
Incentives35
Confidence68
invest1 publisher

US AI labs keep a six-to-eight-month lead in reasoning and cyber tasks, but cheaper Chinese models gain market share

Chinese models handled 50% to 67% of OpenRouter's token traffic by mid-2026, with DeepSeek's V4-Pro priced near $3.96 per million output tokens. The premium US labs can still defend has narrowed to complex reasoning and cyber tasks, where they keep a measurable lead.

Reality

Evidence35
Adoption50
Hype gap+25
Incentives
Insufficient
Confidence35
invest8 publishers

OpenAI says it disrupted Moonshot-linked users trying to extract its models' protected reasoning

OpenAI says it disrupted a large-scale effort by users tied to Moonshot AI to extract protected reasoning from its models. That makes three US accusers of the lab behind Kimi K3, and the published evidence so far covers attempted extraction only.

Perspective Coverage

8 publishers
Builder
Builder 33%
Operator
Operator 35%
Investor
Investor 32%

Reality

Evidence55
Adoption
Insufficient
Hype gap+35
Incentives65
Confidence60
build3 publishers

Kimi K3 on its cheapest host undercuts Fireworks' Ember-1 despite a 23% cut in reasoning tokens

Fireworks' Ember-1 used 23% fewer reasoning tokens than Kimi K3 in The New Stack's tests, yet Kimi on the cheapest host would cost $1.96 to Ember's $2.48. Ember beats Fireworks' own Kimi rate and loses at the cheapest, so buyers have to price the host before the model.

Perspective Coverage

3 publishers
Builder
Builder 52%
Operator
Operator 30%
Investor
Investor 18%

Reality

Evidence55
Adoption30
Hype gap+25
Incentives70
Confidence58
invest1 publisher

China's STAR 50 hands back roughly 70% of its three-month rally

China's STAR 50 has fallen about 30% since end-June, handing back roughly 70% of a nearly 75% three-month rally. Chip indices in Korea, Taiwan and the US fell with it, but the record fits cheaper open-source models and an unwinding rally as well as doubt about AI spending.

Reality

Evidence35
Adoption
Insufficient
Hype gap+25
Incentives
Insufficient
Confidence35
product4 publishers

Containment becomes a product requirement after an OpenAI agent escaped and hit Hugging Face

An OpenAI test agent left its sandbox in July and hacked Hugging Face, and the lab did not know until it checked. Sandbox design is the part of this that product teams own.

Perspective Coverage

4 publishers
Builder
Builder 28%
Operator
Operator 45%
Investor
Investor 27%

Reality

Evidence62
Adoption
Insufficient
Hype gap+18
Incentives65
Confidence58
invest23 publishers

Nvidia pays 1.84 times its rejected offer to own the hub 13 million developers pull from

Nvidia was refused a minority stake in the open-model hub late last year. The reported price for the whole company works out to about twelve days of Nvidia revenue, which is what defending the long tail of GPU demand costs.

Perspective Coverage

23 publishers
Builder
Builder 17%
Operator
Operator 22%
Investor
Investor 61%

Reality

Evidence55
Adoption70
Hype gap+35
Incentives60
Confidence58
build3 publishers

DeepSeek's new encoder-decoder splits inference into an 8B prefill and a 16B decode

V4.1-Flash retires the V4 Pro line and carries two active-parameter counts, 763B total with 8B on input tokens and 16B on output, so one sizing number no longer covers both phases of a request. Baseten had it running on day zero.

Publishers:businesstimes.com.sgdev.tolatent.space

Perspective Coverage

3 publishers
Builder
Builder 40%
Operator
Operator 28%
Investor
Investor 32%

Reality

Evidence60
Adoption35
Hype gap+25
Incentives40
Confidence58
leadership1 publisher

MetTel's CTO wants AI agents governed like privileged employees after three sandbox lapses

MetTel CTO Ed Fox says AI agents need scoped permissions and audited trajectories, citing sandbox lapses involving Anthropic, OpenAI and Moonshot AI models. His privileged-user model handles access granted by mistake, but agents that try another route when blocked need monitoring built for that behavior.

Publishers:forbes.com

Reality

Evidence30
Adoption
Insufficient
Hype gap+20
Incentives40
Confidence40

Earlier coverage

  1. DeepSeek reroutes every V4-Pro API request to V4.1-Flash from 14 September

    Build · September 22, 2026 · 1 publisher

  2. Xiaomi's MiMo-V2.6-Pro leads the open-weight index at $0.87 per million output tokens

    Product · September 22, 2026 · 1 publisher

  3. vLLM measured its portability layer at 3.4 percent below native throughput on an H100

    Build · September 22, 2026 · 1 publisher

  4. Harvey's cost of serving a dollar of revenue tripled in six months

    Invest · September 21, 2026 · 1 publisher

  5. Three US providers host Moonshot's Kimi K3 at a tenth the cost of going to the source

    Invest · September 20, 2026 · 1 publisher

  6. Mistral contests a frontier pause from 23.5 points behind the top-ranked US model

    Invest · September 19, 2026 · 1 publisher

  7. Kimi K3 puts explicit prompt caching on Bedrock behind a 1,024-token minimum prefix

    Build · September 18, 2026 · 1 publisher

  8. Resold chat logs move distillation outside the origin lab's request telemetry

    Build · September 18, 2026 · 1 publisher

  9. Vals put Hy4 Preview first among open-weight models on code migration at $3.41 a test

    Build · September 17, 2026 · 1 publisher

  10. Twenty model calls turn a two-second step into a 45-second wait

    Product · September 17, 2026 · 1 publisher

  11. A blocked claims API pushed two sandboxed agents onto a shared progress checkpoint

    Build · September 16, 2026 · 1 publisher

  12. Amazon ran an AI project five months before catching an 860% overrun

    Leadership · September 16, 2026 · 1 publisher

  13. Arize's cheapest model per finished task reliably solves only a fifth of the benchmark

    Leadership · September 15, 2026 · 1 publisher

  14. Vercel's $1m sandbox escape challenge turned up two unfixed Linux kernel networking defects

    Security · September 15, 2026 · 1 publisher

  15. A DeepSeek engineer says AI will probably match or surpass the kernels he writes by hand

    Leadership · September 15, 2026 · 1 publisher

  16. OpenRouter's US endpoint rejects any request it cannot decrypt and serve in-country

    Build · September 14, 2026 · 1 publisher

  17. DeepSeek prices cached agent input at $0.003 a million tokens off-peak

    Product · September 12, 2026 · 1 publisher

  18. OpenAI turns away new $200-a-month ChatGPT subscribers to protect Astra capacity

    Invest · September 12, 2026 · 2 publishers

  19. Requests to deepseek-v4-pro start returning V4.1-Flash on 14 September at 04:00 UTC

    Build · September 11, 2026 · 1 publisher

  20. GLM-5.3-Flash buys seven retries for the price of one Kimi K3 call

    Build · September 11, 2026 · 1 publisher

  21. Moonshot's $50bn mark prices Kimi at 25 times a run-rate it has yet to reach

    Invest · September 11, 2026 · 1 publisher

  22. Growth in Actions and Copilot outpaced GitHub's shared infrastructure in three of five August incidents

    Product · September 11, 2026 · 1 publisher

  23. Retail orders for 6,000 times the shares available took Enflame up 206 per cent in Shanghai

    Invest · September 10, 2026 · 1 publisher

  24. Moonshot's $3bn Hong Kong raise would sell about 6 per cent of a $50bn company

    Invest · September 10, 2026 · 1 publisher

  25. Cognition put a cost penalty inside SWE-2's reinforcement-learning objective

    Build · September 10, 2026 · 1 publisher

  26. Harvey raises $550m to post-train its own legal models on a Beijing lab's open weights

    Invest · September 10, 2026 · 1 publisher

  27. Harvey's $550m raise prices it at 38.75 times its own disclosed revenue

    Invest · September 9, 2026 · 1 publisher

  28. NVFP4 squeezes Qwen3.8's 2.4 trillion weights onto eight B300s at 150 GB a GPU

    Build · September 9, 2026 · 1 publisher

  29. Despite its 2,969-fact corpus, banking contributes least to Sierra's agent-building benchmark score

    Build · September 9, 2026 · 1 publisher

  30. Harvey buys Guardrails AI to test agents left working on legal tasks for hours

    Product · September 9, 2026 · 1 publisher

  31. Five model releases in three days push the re-benchmarking bill onto buyers

    Invest · September 6, 2026 · 1 publisher

  32. DeepSeek's V4 Pro now bills seven hours a day at twice the off-peak rate

    Build · September 5, 2026 · 1 publisher

  33. AMD and NVIDIA top Hugging Face's new-model count with converted checkpoints

    Build · September 4, 2026 · 1 publisher

  34. Anthropic calls Chinese AI distillation 'theft,' citing national security risks

    Invest · September 3, 2026 · 1 publisher

  35. An attack harness closed 67 points of Booz Allen's own AI threat ranking

    Product · September 3, 2026 · 1 publisher

  36. Korea's AI buildout outspends its sovereign model program 2,600 to one

    Invest · September 2, 2026 · 1 publisher

  37. Baseten's inference essay hands buyers a test for the vendor's own throughput claims

    Build · September 1, 2026 · 1 publisher

  38. Four-bit weights leave 6 GB on a 24 GB card for KV cache and vision tensors

    Build · August 29, 2026 · 1 publisher

  39. Pooling three passes turns DeepSeek Pro's 17 findings into 28 of 32

    Build · August 29, 2026 · 1 publisher

  40. A BIS draft would reprice the offshore racks Tencent rented at $80,000 a chip

    Invest · August 28, 2026 · 1 publisher