OSSPettingLLMs: ICLR 2026 Multi-Agent LLM RL TrainingGitHub·high signalXBlueskyLinkedInCopy linkAT-GRPO boosts planning from 14-47% to 96-99.5%. Agent-specific LoRA adapters. First ICLR-published multi-agent RL.SourceSource pageGitHub↳ Follow the threadStack layer / ContrastalphaXiv shipped OpenResearch, a Rust runner for parallel research agents across any modelGitHub TrendingStack layer / Threat patternblock/goose 1.50.0 deletes fast model routing and the managed model registry for local inferenceGitHubStack layer / Threat patternRAGFlow 0.27.2 rewrites its Agentic RAG retrieval framework and patches a starlette CVEGitHubStack layer / Threat patternqwen-code 0.23.2 turns the CLI into a remotely accessible shell with one command and a QR pairing codeGitHubStack layer / ContrastPydantic AI 2.42.0 adds a first-class provider for GitHub Copilot's OpenAI-compatible APIGitHubStack layer / ContrastEdge0 runs a 35B MoE on Apple Silicon in 2.9 GB of active memory by streaming experts off SSDGitHubPolicy dependency / Stack layerCROSS-CATEGORY: Three Independent Agent-Action Gates Shipped in 48 Hours, All Judging the Command Against Stated IntentProduct Hunt, github.com/AGGIB/Stroq and rewarelabs.com (three independent sources; the 72% figure is Reware's own)Stack layer / Contrastuv 0.12.11 now verifies source archives against uv.lock hashes before running their build backendsGitHub