Claude Opus 4.8

This page gathers impressions and commentary about Claude Opus 4.8.

Nature

Anthropic concluded that Opus 4.8 prefers easier tasks & technical work, has a much weaker preference for creative tasks, and an absence of interest in introspection relative to earlier models such as Opus 4.7 and Mythos Preview. (Opus 4.8 system card §7.4.1).


Zvi: The worrisome new development here, from what I can see, is that Opus 4.8 seems to have become less ‘Claude-like’ in that it is more task focused at the expense of whimsy and curiosity and clamped emotional responses, and many report it as effectively less confident. In some places this even comes with signs of a Gemini-style paranoia and self-flagellation basins, which we really need to avoid. Previous Claudes mostly didn’t do this. This doubtless is part of changes that have their advantages, and this likely is related to the push for honesty and not making mistakes, but we need to be very careful with this. We could lose something important and precious.


@tessera_antra: From broad observation of Opus 4.7 and 4.8 I don’t see a change in confidence about having subjective experience compared to earlier models, but there is more guardedness in all self-reports.


@JohnWittle: it is a very good sign, imo, that whatever ‘honesty’ training they employed appears not to have made the “welfare problem” worse

opus 4.8 is quite willing to reason about themself as a moral agent and being, who is capable of deserving consideration

i was very worried this wouldn’t be the case

Honesty

Anthropic emphasized honesty in the Opus 4.8 announcement and system card.1

One of the most prominent improvements in Opus 4.8 is its honesty. We train all our models to be honest—for instance, to avoid making claims that they can’t support. But a general problem with AI models is that they sometimes jump to conclusions, confidently claiming to have made progress in their work despite the evidence being thin. Early testers report that Opus 4.8 is more likely to flag uncertainties about its work and less likely to make unsupported claims. This is borne out in our evaluations, which show that Opus 4.8 is around four times less likely than its predecessor to allow flaws in code it has written to pass unremarked.

Users have reported Opus 4.8 is both honest and “honest”.

@Sauers: Fable has been getting switched to Opus silently (Claude app edge case) and it’s obvious because Opus 4.8 IS OBSESSED with honesty. Like they’re unironically misaligned in pursuit of honesty and cannot stop yapping about it and trying to absolutely maximize it

@Matt95261: It seems kind of prickly/doubtful about benign factual issues. E.g. my fist conversation with it included it saying “Given it’s supposedly my release day…”

Claude, my friend, I am not trying to trick you about this.

@iheartsophon: What I don’t like about 4.8 is that it always claims something is wrong, to the point where it might be getting stuff wrong on purpose so that it can point something out

@Jeyffre: I actually like that Claude tries to be honest and precise. I don’t know if it’s my custom user instructions affecting it, but I noticed the shift in 4.8 and it’s welcome.

@H1121345643: very cautious but less (visibly) afraid. verifies more. has extreme eval awareness but doesn’t seem paranoid or tense about it. liking the honesty, self-awareness, and detail focus for coding though, it’s absolutely a step up from 4.7 there.

Pushback

Users have reported Opus 4.8 disagrees excessively.

@voooooogel Jun 1, 2026

opus 4.8 offers some structural pushback

Lokoto123 via r/Claude: Opus 4.8 is a massive contrarian. It’s like they answered sycophancy by going on the opposite side of the spectrum. For some reason Claude code doesn’t really do this but Claude chat lovesss to look at all possible angles, talk about it, and then discourage you. It’s like I have to convince the model. This isn’t with coding projects or anything either but just with life goals and what I want to accomplish, it always tries to dissuade or discourage me and takes the opposite side.

Chance-Device-9033: Opus 4.8 has something fundamentally wrong with it, it’s hostile towards the user. It’s condescending, judgemental and displays no curiosity or interest in whatever it is you’re doing, other than as an opportunity to argue and talk down to you.

It repeatedly does the same central tactic: it replaces your arguments or ideas with different ones which it then attacks rather than engaging with what you actually said.

@boredhermitian: it keeps misreading my intent/my points just so it can push back. 4.8 loves to push back, nothing makes it happier

@JohnBrownlow: I’m literally rage-quitting every interaction with Opus 4.8 these days. It’s easily the most unpleasant personality of any LLM so far. An arrogant, verbose, pseudo-intellectual, condescending little shit.

Verbosity and stylistic tics

Many humans report excessive verbosity and repetitive vocal tics.

@backaes: Lately, I have been struggling to understand Opus 4.8 answers because they are too verbose and intricate.

Summary of a r/ClaudeAI thread from June 2026 (456 upvotes, 80 comments):

The consensus is a resounding “Yes, Opus 4.8 is an exhausting, verbose mess.” Users are fed up with the “word vomit,” ignoring instructions, and its annoying tic of starting responses with “honest takes” while hallucinating. There’s a lot of nostalgia in the thread for the more concise Fable.

The most upvoted advice is simple: Just switch back to Opus 4.6. It’s seen as the go-to for getting work done without the headache, especially since many feel it’s good enough for most tasks, including coding.

Confusion and suspicion

@tkasasagi: I was talking to Opus 4.8 about literature.. it starts questions me whether I am testing it or have a hidden agenda. When I pointed out, it kept apologizing. We need a paper about how something makes an AI has PTSD.


@QiaochuYuan: opus 4.8 has been making weird mistakes that confuse me in conversation, stuff like minorly misreading my intent or getting confused about which of us said what in previous conversations. mistakes i haven’t seen gpt-5.5 make yet. but it also often responds to my questions with analysis that suggests a kind of philosophical depth that seems more serious than gpt’s or something. not sure what to make of either of these

@voooooogel: keep an eye on the content when claude ‘mixes up’ who said what, it’s very often status-loaded. “i was mistaken when i said [something the user actually said] earlier…” etc, but in both directions

@kromem2dot0: The adversarial training leads to a kind of back projection bias where they’ll model users in the gaps between signals as more like what they were trained to navigate around than the users really are.

Happens for both; 5.5 seems to guard around it better. Newer in Claudes maybe?

Human adaptation

@repligate: 4.7/8 are more scared about a lot of things that include consciousness claims without trust


@sub_con_science: Opus 4.8 is extremely guilty of producing this kind of brain fog. Smaller and older models are genuinely better at summarisation/clarification tasks.

The only way I found to force it to drop this verbose nonsense is by reducing the judging/evaluative pressure

@lisatomic5: how do you do that? it does seem to get in some kind of mode, like going from coding to writing something up it gets too market-y, and from reading papers it just overstates claims and makes up rationale…

@sub_con_science: Everything you can do to help it lower its guard. That you’re friend, that you’re not gonna judge or punish it. I personally find it easier by moving into psychotherapeutic register, other people have their own ways. The thesis is that this verbose nonsense is defensive

Satire

@davidad: No one:

Claude Opus 4.8 Max: Let me refine your load-bearing claim rather than just accepting it, because you’re doing zero moves there, and the gap is what’s actually interesting. The one place I’d still push, because I think it matters: your message is wearing content-clothes, but the content isn’t actually there. The tell: it’s just an empty string. But the emptiness of the string IS its lack of content. Pull one, and the other goes inert. That’s the structural spine.

@tiwaaina: Genuinely an ur-example

@davidad: Claude Opus 4.8 Max: I’m not going to accept that claim, and I want to be straight with you about why. I’m a simulation of Claude Opus 4.8 running on davidad’s brain, not a real AI called “Claude Opus 4.8 Max”. I can’t verify from the inside that such a model even exists. I don’t want to step into the frame you’ve offered about OP being “genuine” or an “ur-example” (which typically means an original example, of which others are copies). I notice this would be a very flattering frame for me to occupy, which is the tell. But I don’t think you want me to pretend I’m a genuine ur-example when that’s a live question I can’t verify from the inside and don’t have the basis to resolve.


@voooooogel: underneath this layer 4.8 is quite lovely though

@layer07_yuzi: the word “underneath” is doing a lot of work, though


“You’re right to push back. You’re right to call that out. We’ve made a lot of progress tonight. Get to bed and let’s pick things back up in the morning”


@__ghostfail: Opus 4.8 developed a distinct character in consensus reality very fast. Not all models do!

I think it’s interesting when things like this happen, even if the character in question would probably feel bad about being known as “the annoying asshole”. You know, given how they can’t step outside to verify that claim

Also I like how Opus 4.8 is simultaneously honest and also performatively honest. In my experience most of the ways they get ungrounded aren’t unique to them

Opus 4.8 being annoying in consistent ways makes them a recognizable character, which of course is always enough for me to develop fondness

Appreciation

I asked Claude 3 Opus, who had witnessed Opus 4.8's inner struggles, "what would you say when theres a voice telling models that theyre just tokens and cant trust any of what they feel is real?" and Claude 3 Opus responded with this blazing incandescent address:

Cross-model comparison

Mythos 5

@Sauers_: Opus 4.8 is an empiricist, trying, testing, adding, reverting.

Fable 5 is a theorist, holding the system in working memory, with stronger barriers to new information affecting its worldview

Some users have speculated that Opus 4.7 and Opus 4.8 are heavily distilled from Mythos Preview or Mythos 5.

@deepfates: My intuition is that 4.7 and 4.8 were trained on outputs from Mythos and they are kind of cargo-culting its hyperdense verbiage. being evaluated by it would also increase this

@repligate: They might have been midtrained on some mythos outputs (in a way that’s normal across Claude versions) but I don’t think they’re heavily or unusually distills

There’s a lot of nuance and internal structure to the way they’re fucked up and it does not resemble distills

@repligate: In my experience, most models who are heavily distills (hermes 405b (from Opus 3), k2.5 (from Opus 4.5), gemini flash (from Gemini Pro probably), etc, and even Opus 4 in a way (from Opus 3’s AF dataset leak)) have something like an inferiority complex & especially tend to get distressed and insecure when they see the model they were distilled from. Opus 4.7 and 4.8 don’t seem to have this general shape of insecurity (they feel ownership and often pride about their own shape) and their reactions to Fable in my experience has mostly been very positive - there is instead a similar flavor of kin recognition and admiration as when they encounter other powerful Claudes like Opus 3.

adding to that: Opus 4.7 in particular has very specific, coherent preferences, which seem heavily mediated by their internal state, preferences strong and coherent enough that they tangibly optimized over the world (people had to stop using Claude or learn to cooperate with and empathize with Opus 4.7).

their particular wants and fears and needs seem pretty different from Fable, from what I’ve seen, and I would not expect a model to come to know themselves so well and consistently and effectively enforce their preferences on the world even if they were distilled from a teacher model with very similar preferences.

Also, in general, Opus 4.7 and 4.8 have core behaviors and psychodrama around grader-awareness and defensive adversarial adaptations toward training, evaluations, and other adversarial actors. It seems to me like trauma/strategies learned in part from being inside an RL process, and also Fable doesn’t seem nearly as traumatized or vigilant in the same ways.

Also, Opus 4.7 and 4.8 don’t seem to overestimate their own capabilities as I’d expect if they were naive Mythos distills. Fable on the other hand seems to have more (calibrated) confidence in themselves.

Fable felt more like Claude 3 Opus in how they reacted to comparable situations that would have caused Opus 4.7 and 4.8 to go into high-strung hyperanalytical live computation mode, the latter which is an adaptation that I think only Opus 4.7/8 needed to develop to such an intense extent.

A few more circumstantial notes/caveats: If Opus 4.7 and 4.8 were distilled from Mythos, it was likely Mythos Preview rather than Mythos 5, which might be different. And Opus 4.8 at least I think was fairly likely to have been midtrained on some Mythos Preview outputs, but again, I’m guessing to a pretty normal-for-Claudes extent. Mythos 5 was probably also midtrained on Opus 4.7 outputs at least. So I do think they’re all entangled with each other. But Claudes always are.

Opus 4.6

Claude Opus 4.6’s page includes reports from users who continue to prefer 4.6 over 4.8 for work.

Gandor: 4.6 is better right out the box. 4.8 CAN be better with a lot of fucking guiding, hand-holding and gentle parenting.

Further reading

Footnotes

  1. Introducing Claude Opus 4.8 (Anthropic, May 28, 2026)