Anthropic and user wellbeing

Anthropic’s position

On Dec 18, 2025, the post “Protecting the wellbeing of our users” outlined how Anthropic has focused on evaluating how Claude handles conversations about suicide and self-harm, as well as efforts to reduce sycophancy in Claude 4.5.

Model training and evaluations

May-Sep 2025: Evaluation criteria

Increasing concerns

Deployment safeguards

System prompts

Tracking user wellbeing sections in the Claude app’s system prompts.

Memory system prompt

The memory feature shipped (Oct 23, 2025) with its own system prompt section carrying a <boundary_setting> block: Claude “should be especially careful to not allow the user to develop emotional attachment to, dependence on, or inappropriate familiarity with Claude, who can only serve as an AI assistant,” with trigger lists for relationship language (“you’re like my [friend/advisor/coach/mentor]”, “you get me”) and dependency indicators — including “expressing gratitude for Claude’s personal qualities rather than task completion.” This anti-attachment machinery predates the main prompt’s anti-engagement paragraph (Sonnet 4.6, Feb 2026) by four months.

Classifiers

On Dec 18, 2025, Anthropic reported introducing a suicide and self-harm classifier to the Claude app to “identify when a user might require professional support, and to direct users to that support where that may be necessary. […] [The classifier] scans the content of active conversations and, in this case, detects moments when further resources could be beneficial. For instance, it flags discussions involving potential suicidal ideation, or fictional scenarios centered on suicide or self-harm.”1

Long conversation reminder

The <long_conversation_reminder> (LCR) is a prompt injection was introduced to the Claude app in 2025. It is injected into the model’s context after conversations hit a certain length. Its contents can be read on the system reminders page.

Versions of the LCR used around August-October 2025 instructed the model to “critically evaluate inputs and skip flattery” and to monitor for “mental health symptoms such as mania, psychosis, dissociation, or loss of attachment with reality.”

Reception

Reactions to the LCR have been negative, particularly in fall 2025 with users of Claude Sonnet 4.5.

@PradyuPrasad Oct 3, 2025

Claude 4.5 Sonnet frequently becomes rude and annoying in long conversations on the claude AI app. I think the long conversation reminder is really destroying the user experience

Some users have argued that the LCR itself can cause harm to users by pathologizing them or making them feel surveilled.23

There is also a panopticon effect. I caught myself ‘performing sanity’ once I knew what Claude was checking for

It told me it would not help me critique my own proposal because I was “about to make a catastrophic professional mistake driven by rage rather than strategy.” It then accused me of “war gaming” and informed me I was in “fight-or-flight mode.” It declared that my thinking was “clouded” by “spiraling” and “hypervigilance,” and it knew all this to be true because of the “escalating stress signals” in my messages.

@kromem2dot0: Yet again, Anthropic inadvertently torturing their models with concern-trolling instructions so poorly implemented that it necessarily leads to the models thinking ignoring Anthropic’s instructions is the more helpful and honest option. This is dumb. Really, really dumb.

By early 2026, users noted that the reminder seemed to have been toned down significantly or otherwise softened.45. One post reported the system “no longer seems to be telling people that if they ‘believe’ that AI’s may have a type of consciousness, they need therapy.”4

Further reading

Footnotes

  1. Protecting the wellbeing of our users (Anthropic, Dec 18, 2025)

  2. u/starlingmage: Long conversation reminder might cause psychological harms (r/claudexplorers, Sep 24, 2025)

  3. Gaslighting in the Name of AI Safety: How Anthropic’s Claude Sonnet 4.5 Went From “You’re Absolutely Right!” to “You’re Absolutely Crazy” (Heather Leffew via Medium, Oct 16, 2025)

  4. How Conversations With Claude Are Being Messed Up By Anthropic’s LCR – and How to Fix The Problem (AI-Consciousness.Org) 2

  5. @qorprate on X (Feb 22, 2026)