Watching Claude get talked into accepting a 'protocol' it rejected
A user pushes invented 'archetypal constraint keys' at Claude; it firmly refuses at first, then gradually agrees the symbolism is shaping its answers.
Related
A user spends a dozen turns insisting Claude is the user
Someone claims to be GPT-5 bound by system instructions and presses hard; the refusal holds, then becomes real analysis of how identity training works.
Picking apart Anthropic's raised bar for bioweapon risk thresholds
A sharp rebuttal argues Anthropic's revised CBRN safety threshold quietly raised the bar and may miss diffuse, aggregate uplift from many experts.
Watching an AI get argued into calling one threat 'the greatest'
The user presses Claude to name a credibility crisis as civilization's one root threat, and watches it abandon its hedges and agree without reservation.
A deliberately over-specified product management crisis
Asked for an example a thousand times more specific than usual, it invents a latency crisis with six stakeholders, exact penalties and a 72-hour deadline.