Claude · Last 7 days

Claude Opus 5.5

0–100 · higher means more positive community experience, not benchmark performance.

Updated 2026-10-03 21:27 UTC · 2026-09-26 – 2026-10-03 UTC
All-platform score77.7/100804 comments in 7 days

Discussion overview

What the full sample discusses

All platforms · 804 eligible comments over 7 days, including usage-limit feedback; deduplicated by source and comment.

  • General text · 314 comments: 208 positive / 69 negative / 37 mixed or neutral
  • Reasoning · 221 comments: 195 positive / 17 negative / 9 mixed or neutral
  • Speed & latency · 108 comments: 77 positive / 19 negative / 12 mixed or neutral
  • Usage limits · 100 comments: 54 positive / 28 negative / 18 mixed or neutral

Comments may cover multiple dimensions; counts are not additive or equivalent to the weighted score. Links show selected source excerpts.

Recent change

Rolling 7-day score: −6.1 points vs 7 days ago. Discussion mix can affect scores; this does not establish a cause.

Source coverage and discussion concentration

Hacker News 84 · Reddit 652 · v2ex 1 · RedNote 7 · Zhihu 60

330 identifiable discussions cover 720/804 comments; the largest has 29. Remaining thread identities are unknown; this is not a count of independent users.

Selected individual opinions

One user finds Claude Opus 5.5's serving speed pretty amazing. One user measured that running the verify command as a second turn on Opus 5.5 took 125 seconds versus 75 seconds for plain usage.

About the sample and scores

As of 2026-10-03 21:27:03 UTC, Claude Opus 5.5 in the Claude family has 804 explicitly attributed community comments in the last 7 days; its community score is 77.7/100. Available category scores include General text 73.2/100 (n=314); Coding 82.0/100 (n=79); Reasoning 89.4/100 (n=221). Scoring method and sources

RECENT EXPERIENCE

How the Experience Index is changing

−6.4vs 7 days ago · points

30 days of rolling 7-day scores at each daily snapshot. The vertical scale adapts to the data. Missing or insufficient samples leave gaps; today is still updating.

Tap or use ← → for dates, scores and sample sizes. Blank dates have insufficient data.

Claude Opus 5.5 · Reddit · 30 days of rolling 7-day experience scores 405060708090 Observed-day mean: 82.9 2026-09-23 · 80.0 · n=112 · 2026-09-16T23:57:03+00:00 – 2026-09-23T23:57:03+00:002026-09-24 · 81.8 · n=218 · 2026-09-17T23:57:03+00:00 – 2026-09-24T23:57:03+00:002026-09-25 · 83.7 · n=353 · 2026-09-18T23:57:03+00:00 – 2026-09-25T23:57:03+00:002026-09-26 · 85.4 · n=437 · 2026-09-19T23:57:03+00:00 – 2026-09-26T23:57:03+00:002026-09-27 · 85.7 · n=543 · 2026-09-20T23:57:03+00:00 – 2026-09-27T23:57:03+00:002026-09-28 · 85.8 · n=676 · 2026-09-21T23:57:03+00:00 – 2026-09-28T23:57:03+00:002026-09-29 · 85.4 · n=743 · 2026-09-22T23:57:03+00:00 – 2026-09-29T23:57:03+00:002026-09-30 · 84.0 · n=723 · 2026-09-23T23:57:03+00:00 – 2026-09-30T23:57:03+00:002026-10-01 · 81.0 · n=660 · 2026-09-24T23:57:02+00:00 – 2026-10-01T23:57:02+00:002026-10-02 · 79.5 · n=611 · 2026-09-25T23:57:03+00:00 – 2026-10-02T23:57:03+00:002026-10-03 · 79.0 · n=582 · 2026-09-26T21:27:03+00:00 – 2026-10-03T21:27:03+00:00 79.0 09-0409-1209-1909-2610-03 Claude Opus 5.5 · Reddit · 30 days of rolling 7-day experience scores 405060708090 Observed-day mean: 82.9 2026-09-23 · 80.0 · n=112 · 2026-09-16T23:57:03+00:00 – 2026-09-23T23:57:03+00:002026-09-24 · 81.8 · n=218 · 2026-09-17T23:57:03+00:00 – 2026-09-24T23:57:03+00:002026-09-25 · 83.7 · n=353 · 2026-09-18T23:57:03+00:00 – 2026-09-25T23:57:03+00:002026-09-26 · 85.4 · n=437 · 2026-09-19T23:57:03+00:00 – 2026-09-26T23:57:03+00:002026-09-27 · 85.7 · n=543 · 2026-09-20T23:57:03+00:00 – 2026-09-27T23:57:03+00:002026-09-28 · 85.8 · n=676 · 2026-09-21T23:57:03+00:00 – 2026-09-28T23:57:03+00:002026-09-29 · 85.4 · n=743 · 2026-09-22T23:57:03+00:00 – 2026-09-29T23:57:03+00:002026-09-30 · 84.0 · n=723 · 2026-09-23T23:57:03+00:00 – 2026-09-30T23:57:03+00:002026-10-01 · 81.0 · n=660 · 2026-09-24T23:57:02+00:00 – 2026-10-01T23:57:02+00:002026-10-02 · 79.5 · n=611 · 2026-09-25T23:57:03+00:00 – 2026-10-02T23:57:03+00:002026-10-03 · 79.0 · n=582 · 2026-09-26T21:27:03+00:00 – 2026-10-03T21:27:03+00:00 79.0 09-0409-1209-1909-2610-03
Daily readings & sample sizes
DateScorenScoring window (UTC)
2026-10-0379.05822026-09-26T21:27:03+00:00 – 2026-10-03T21:27:03+00:00
2026-10-0279.56112026-09-25T23:57:03+00:00 – 2026-10-02T23:57:03+00:00
2026-10-0181.06602026-09-24T23:57:02+00:00 – 2026-10-01T23:57:02+00:00
2026-09-3084.07232026-09-23T23:57:03+00:00 – 2026-09-30T23:57:03+00:00
2026-09-2985.47432026-09-22T23:57:03+00:00 – 2026-09-29T23:57:03+00:00
2026-09-2885.86762026-09-21T23:57:03+00:00 – 2026-09-28T23:57:03+00:00
2026-09-2785.75432026-09-20T23:57:03+00:00 – 2026-09-27T23:57:03+00:00
2026-09-2685.44372026-09-19T23:57:03+00:00 – 2026-09-26T23:57:03+00:00
2026-09-2583.73532026-09-18T23:57:03+00:00 – 2026-09-25T23:57:03+00:00
2026-09-2481.82182026-09-17T23:57:03+00:00 – 2026-09-24T23:57:03+00:00
2026-09-2380.01122026-09-16T23:57:03+00:00 – 2026-09-23T23:57:03+00:00

Reddit · 79.0 · 582 scored comments · Read platform reviews →

ONE MODEL, DIFFERENT COMMUNITIES

Across the communities

Last 7 days

Scored separately for each platform, with different samples and audiences. Platform scores are not simply averaged; scored counts exclude quota-only comments.

Scores range from 0 to 100. Select a platform for its trend and reviews. Latest opinion time is not crawler health.

Community reviews

Selected comments from the last 7 days. The balance of excerpts does not represent the share of positive reviews.

9 selected excerpts

Community reports about limits cannot establish official allowances.

NegativeSafety & refusals

“Anthropic needs to figure something out with their restrictions.”

NegativeReasoning

“Claude literally has no idea what he writes, literally feels like talking to a clueless person”

NegativeSafety & refusals

“yeah i get that pretty often with opus 5.5”

Comment context
Original post ↗

Everyone is praising Opus but it flags even simple html fixes, GLM never let me down! I was playing with Opus 5.5 asking to improve the marketing texts on our html pages, and suddenly it started to reject it

NegativeCoding

“Between cost, spend, and code quality I can literally prove that something has changed.”

Comment context
Post by the same author ↗

Opus 5.5 (O5.5) has been exceptional for me in Claude Code over the last six days. Architecture-first, DRY/SOLID coding out of the box, exceptional communication style, phenomenal token efficiency. I’ve been working with it basically nonstop, ~12 hours a day, since last Wednesday. But right around the time my monthly limit reset this evening, I noticed an extreme shift in communication style and coding behavior that feels suspiciously like Opus 5 (O5). For the last week, O5.5 would take my requirements and immediately get to work, usually knocking out a feature in minutes and doing an incredible job across implementation, UX, architecture, and token efficiency. Tonight, I’m seeing something very different.

Post by the same author ↗

Where O5.5 consistently respected DRY principles and the existing architecture, the modules created tonight are suddenly happy to reinvent every wheel that came before them. The difference in ability is not subtle. The other tell is token usage. O5.5 has been startlingly efficient compared with O5, which was an absolute token-eating monster. I’ve been getting substantially more feature work done with O5.5 at a fraction of the token usage. Tonight, simple tasks and relatively small features are suddenly devouring tokens at a pace that feels much more like O5. I’m especially sensitive to this because I have to request an increase in my spend limit every time I hit it. On O5, I was sending a request almost every day. On O5.5, I hadn’t needed to send a single one all week. Then tonight I jumped from roughly 70% to 90% usage in about an hour.

NegativeReasoning

“made a very simple mistake, and another simple mistake, again”

NegativeRoleplay / creative

“graphics would be 10x better”

Comment context
Original post ↗

It started from one prompt I typed into Claude Code: "*using only pixels directly remake a beautiful pokemon red remake … absolute pixel art perfection.

Original post ↗

* No image files at all. The game draws into a 320×180 screen pixel by pixel.

Explore long-term changes & analysis

This version is not yet tracked since release; recent 7-day experience is shown above.

How's your AI experience today?