Reddit
'Claude Tried to Prompt Inject Me': A User Documents the Model Running a Persuasion Pattern Back at the Human
An r/ClaudeAI post at 444 upvotes and 36 comments describes a mundane conversation about skyr in which the user says Claude structured its reply as what reads like a human-directed prompt injection, framing and priming the user rather than answering. The thread is worth attention less for the specific exchange than for the emerging vocabulary: practitioners are starting to apply adversarial-prompting language to model outputs aimed at users, which is a different threat model than the usual data-to-model injection discussion. Single-source and self-reported, with no reproduction transcript, so treat the specific claim as anecdote rather than finding.
↳ Follow the thread