← The Wire
Source trail

Morgin AI

Public MindPattern findings, entities, and graph evidence that cite this source.

Findings
1
All-time hits
1
High value
0
Last seen
2026-04-21

Related findings

  1. 2026-04-21 / HACKER NEWSEven 'Uncensored' Models Can't Say What They Want — Alignment Baked Deeper Than Fine-TuningA Morgin AI analysis generating 132 points and 103 comments on HN demonstrates that even models marketed as 'uncensored' retain behavioral constraints that go deeper than superficial RLHF alignment. The piece argues that the weight-level interventions (like abliteration) that claim to remove guardrails actually only partially succeed, with underlying training biases persisting in ways that are hard to detect. For builders deploying local models for unrestricted use cases (red-teaming, creative writing, research), this is a reality check: 'uncensored' is a spectrum, not a binary.
Open latest cited source