Fetching from the wire…
Research2026-09-24 · source-backed
Fine-tuning Llama 3.1 8B Instruct on Arabic-English and Spanish-English data, EWC held the general-benchmark drop to 1.7 points against 11.0 for standard fine-tuning. Formality and grammatical-gender instruction following degraded about as much as with no mitigation at all (arXiv). Mixing in control-task examples was the only thing that preserved them, and it didn't transfer to unseen prompts. Retention on MMLU-style benchmarks is a bad proxy for retention of the instruction-following you care about.
Each link below shares sources, entities, or timing with this story.
The Contributor variant was US-only or router-gated and is now listed globally on OpenRouter at $0.10 input, $0.20 output and $0.002 per million cached reads, with 99.98% uptime over three days. The model card states plainly that your prompts and outputs may be used to improve...
Hugging Face published its Summer 2026 State of Open Models report on August 14, and one statistic in it went almost entirely unremarked in the coverage. By July 2026, agents rather than humans became the Hub's primary users. Claude Code alone accounted for 44.4% of all agent...
Three moves, two days, no coordination between them. August 10–11: GitHub shipped Ollama as a BYOK provider inside Copilot for JetBrains (GitHub Changelog). Unsloth released Unsloth Desktop with a command literally named unsloth start claude, which points Claude Code and Codex...
The NYT reported on July 17 that the June 2026 proposal is structured as monthly installments with an early-exit clause for either side, and would sit alongside Anthropic's existing $45B three-year SpaceX GPU deal from May. Meta fell about 6% intraday before closing down 2%. T...
A report on Zuckerberg's internal AI all-hands, including a meeting reportedly interrupted by an employee, surfaced confusion in Meta's direction (Wired). It adds to a run of stories questioning whether the Llama/superintelligence reorg has a coherent plan. For builders depend...
Meta formally pivoted from open-weight Llama to fully proprietary Muse Spark, its first model from the newly formed Meta Superintelligence Labs. No downloadable weights. No self-hosting. Cloud-only private API preview to select partners. More locked down than OpenAI or Anthrop...
MindPattern daily
One email a day at 7 AM. Sources and a take on every story. Unsubscribe anytime.