DispatchUlysses Sequence Parallelism: 12x Longer Sequences on 4x H100Hugging Face Blog·high signalXBlueskyLinkedInCopy linkSnowflake Arctic protocol. SP=4 reduces memory 3.3x, 3.7x throughput at 64K tokens. Integrated into HF Accelerate, Transformers, TRL.SourceSource pageHugging Face Blog↳ Follow the threadStack layer / ContrastCactus Needle 3 is a 121M-parameter tool-calling model that ships in 8-29MB slices and refuses to chatCactus Compute / Hugging Face (via r/LocalLLaMA, 183 upvotes)Stack layer / ContrastTernary Bonsai 2 27B ships a true 1.72-bit quantization of Qwen3.8-27B at 5.95 GB, and hit 405,000 downloads in under two daysHugging Face / Prism ML (via r/LocalLLaMA, 1,347 upvotes)Stack layer / Follow-up threadAutoArk's Edge0-35B-A3B Streams MoE Experts Off SSD to Run a 35B Model in ~3GiB on a Mac MiniarXiv 2609.18063 (via Hugging Face)Stack layer / ContrastClaude Code makes the v2 MCP client and 2026-07-28 negotiation the default on Bedrock, Vertex and FoundryGitHubStack layer / ContrastCline's compaction trigger was estimating tokens at 3 characters each and never firing on dense contentGitHubStack layer / Update threadActObs: supervising environment observations during SFT changes how agents explore under RL, +3.4pp pass@16 on Terminal-Bench 2.0arXiv / HuggingFace Daily PapersPolicy dependency / Stack layerLangChain ships a first-party integration that deliberately does not wrap the vendor's SDKGitHubStack layer / Threat patternPlugin4Shell: One SHA-Pinning Bug Gives Zero-Click RCE in Claude Code, Codex, Copilot and Gemini CLIHelp Net Security (corroborated by The Register)