Fetching from the wire…
Public story · 2026-03-19 · source-backed
Together.ai released Mamba-3 (Apache 2.0, ICLR 2026) — an SSM achieving ~4% better language modeling than the Transformer baseline while running up to 7x faster on long sequences. Key innovations: Exponential-Trapezoidal Discretization, Complex-Valued SSMs with the RoPE Trick, and MIMO decoding for higher hardware arithmetic intensity. The architecture is explicitly inference-first, targeting agentic workloads where inference — not training — is the bottleneck. Together.ai
Each link below shares sources, entities, or timing with this story.
Shared entities / Same source / Shared topic / Earlier coverage / Downstream implication
Both cover ICLR, Mamba, SSM; cite the same source (Together.ai); overlapping topics (architecture, beat, complex-valued, mamba-3).
Shared entities / Shared topic / What happened next
Both cover Apache, Runs; overlapping topics (achieving, agentic, apache); picks up the Apache thread on 2026-04-21.
Shared entities / Shared topic / Earlier coverage
Both cover Mamba, Transformer; overlapping topics (agentic, architecture); earlier Mamba coverage from 2026-02-25.
RMM uses Transformer / Shared entity: Transformer / What happened next
Linked by a graph relationship (RMM uses Transformer); both cover Transformer; picks up the Transformer thread on 2026-08-15.
Shared entity: Apache / Shared topic / What happened next / Tension
Both cover Apache; overlapping topics (agentic, apache); picks up the Apache thread on 2026-06-10.
Both cover Apache; overlapping topics (agentic, apache); picks up the Apache thread on 2026-04-01.
Shared entity: Apache / Shared topic / What happened next
Both cover Apache; overlapping topics (agentic, apache, faster); picks up the Apache thread on 2026-03-22.
Shared entity: Faster / Shared topic / Tension
Both cover Faster; overlapping topics (achieving, architecture, faster); pushes against this story (but).