Fetching from the wire…
Models2026-07-24 · source-backed
Announced alongside FLUX 3, it bolts a lightweight action decoder onto intermediate FLUX 3 features to make a video-action model, beating prior VLA models even with the backbone frozen. Deployed at Audi for kitting, component insertion and flexible-material manipulation, with failure recovery it was never explicitly shown. Generative video pretraining is being harvested as a robotics world model rather than trained separately. No public release date.
Each link below shares sources, entities, or timing with this story.
Posted to Show HN August 26 by Trigger Labs: run npx openspender connect or add the MCP server, and the agent mints its own card with per-request, daily and total caps, settling in self-custodial USDC on Base with an itemized ledger. Coverage spans 15,171 x402 endpoints and 14...
Black Forest Labs put FLUX 3 into early access July 23, trained jointly across modalities using their Self-Flow approach, a departure from the single-modality FLUX 2 line. BFL reports FLUX 3 video preferred over Runway Gen-4.5 in 77% of head-to-head comparisons and over Luma R...
Xiaomi-Robotics-1 (arXiv 2607.15330) reports 74.5% average success on RoboCasa, beating RLDX-1, Cosmos Policy, GR00T N1.6, Pi-0.5 and Pi-0-FAST, plus a new SOTA 57.6% on RoboCasa365 against a prior best of 46.6%. 100K hours of *real* manipulation data is the moat, not the arch...
Still tiny at ~320 stars, vargHQ/sdk treats video as composable JSX components on top of the Vercel AI SDK. It's an early, genuinely novel DX bet, declarative React-style authoring for multi-model video pipelines. Worth watching as code-first AI video tooling finds its shape....
arXiv 2608.13010 scores top-five retrieval candidates against ranks 6–20 of the same query to spot answer-anchor concentration, and separately compares documents to lexically distinct neighbors to catch coordinated density before any query arrives. Deployed jointly, attack suc...
DeepMind announced July 30 in three variants: the VLA, an ER 2 embodied-reasoning model for multi-step planning, and On-Device 2. The step past 1.5 is whole-body locomotion control, previously upper-body only, mapping vision and language to motor commands "from feet to fingert...
MindPattern daily
One email a day at 7 AM. Sources and a take on every story. Unsubscribe anytime.