Sources
Safin-1 builds safety into the architecture through memory routing rather than post-hoc alignment
arXiv 2609.00092 (2026-08-31, 20 HF upvotes) proposes 'Safety from Within,' where safety-relevant capability is represented and invoked by the model's native computation instead of by external guardrails or supervised fine-tuning applied afterward. The family is built on Memory-Anchor Routing across Context History (MARCH), which maintains structured memory states and retrieves relevant history through content-conditioned routing, supporting test-time adaptation of persistent capability states without repeatedly modifying the backbone. Read alongside 2609.01836 on authorization laundering, it is the same week's optimistic and pessimistic takes on agent memory as a safety surface.
Source
↳ Follow the thread