Fetching from the wire…
Public story · 2026-07-19 · high
It's also third on DeepSWE, the first open-weights model to reach frontier level there, per AINews.
Why now: AINews surfaced these throughput and benchmark numbers in its July 19 roundup on Latent Space, ahead of any independent throughput testing.
K3's Kimi Delta Attention architecture claims up to 6x cheaper throughput at a 1 million token context window, per AINews's roundup on Latent Space.
That number matters more than where K3 lands on any leaderboard. Long-context agent loops call a model over and over inside the same task. A throughput multiple compounds across a run in a way a single benchmark score doesn't. A 6x cut at 1M context is the difference between a long-context agent workflow that's affordable and one that blows a budget.
AINews also has K3 sitting third on DeepSWE, and calls it the first open-weights model to reach frontier level on that benchmark. Artificial Analysis separately scores K3 at 64% on DeepSWE and 84% on Terminal-Bench v2, per the same rundown.
AINews doesn't say what baseline the 6x figure is measured against, or what K3 costs per token in dollar terms. That leaves the throughput claim directional until someone runs it under real load.
K3 is a bet that throughput economics beat leaderboard position for agent-heavy workloads. If the 6x figure holds up outside AINews's report, it undercuts frontier models priced by the token for anything that runs long.
Each link below shares sources, entities, or timing with this story.
Shared entities / Same source / Shared topic
Both cover AINews, DeepSWE, Latent Space; cite the same source (Latent Space); overlapping topics (agent, ainew, deepswe).
Shared entities / Shared topic / Earlier coverage / Tension
Both cover Artificial Analysis, Bench, Terminal; overlapping topics (agent, analysi, artificial); earlier Artificial Analysis coverage from 2026-07-13.
Shared entities / Shared topic / Earlier coverage
Both cover Artificial Analysis, Bench, Terminal; overlapping topics (agent, analysi, artificial, benchmark); earlier Artificial Analysis coverage from 2026-07-13.
Both cover Artificial Analysis, Bench, Terminal; overlapping topics (agent, benchmark); earlier Artificial Analysis coverage from 2026-07-13.
Shared entities / Shared topic / Earlier coverage / Tension
Both cover Bench, Terminal; overlapping topics (agent, claim, throughput); earlier Bench coverage from 2026-06-10.
Shared entities / Same source domain / Shared topic / Earlier coverage
Both cover AINews, Latent Space; reported by the same outlet (latent.space); overlapping topics (claim, context).
Shared entities / Shared topic / Earlier coverage / Tension
Both cover Artificial Analysis, DeepSWE; overlapping topics (benchmark, deepswe); earlier Artificial Analysis coverage from 2026-07-09.
Both cover Bench, Terminal; overlapping topics (agent, benchmark); earlier Bench coverage from 2026-06-19.