Claude · Last 7 days
Claude Opus 5.5
0–100 · higher means more positive community experience, not benchmark performance.
Updated 2026-10-02 21:57 UTC · 2026-09-25 – 2026-10-02 UTCDiscussion overview
What the full sample discusses
All platforms · 834 eligible comments over 7 days, including usage-limit feedback; deduplicated by source and comment.
- General text · 324 comments: 220 positive / 67 negative / 37 mixed or neutral
- Reasoning · 236 comments: 209 positive / 18 negative / 9 mixed or neutral
- Speed & latency · 112 comments: 81 positive / 19 negative / 12 mixed or neutral
- Usage limits · 104 comments: 57 positive / 27 negative / 20 mixed or neutral
Comments may cover multiple dimensions; counts are not additive or equivalent to the weighted score. Links show selected source excerpts.
Recent change
Rolling 7-day score: −3.3 points vs 7 days ago. Discussion mix can affect scores; this does not establish a cause.
Source coverage and discussion concentration
Hacker News 89 · Reddit 679 · v2ex 1 · RedNote 3 · Zhihu 62
331 identifiable discussions cover 745/834 comments; the largest has 30. Remaining thread identities are unknown; this is not a count of independent users.
Selected individual opinions
One user's Sonnet agent got stuck ~8 hours on a big task; switching to Opus 5.5 finished that same task in about 30 minutes, showing stronger task completion. One user feels Opus 5.5 is much more likely to be confidently incorrect than Sol 6.1, making its reasoning less trustworthy.
About the sample and scores
As of 2026-10-02 21:57:03 UTC, Claude Opus 5.5 in the Claude family has 834 explicitly attributed community comments in the last 7 days; its community score is 78.5/100. Available category scores include General text 74.7/100 (n=324); Coding 80.2/100 (n=79); Reasoning 89.6/100 (n=236). Scoring method and sources
RECENT EXPERIENCE
How the Experience Index is changing
30 days of rolling 7-day scores at each daily snapshot. The vertical scale adapts to the data. Missing or insufficient samples leave gaps; today is still updating.
Tap or use ← → for dates, scores and sample sizes. Blank dates have insufficient data.
Daily readings & sample sizes
| Date | Score | n | Scoring window (UTC) |
|---|---|---|---|
| 2026-10-02 | 64.1 | 84 | 2026-09-25T21:57:03+00:00 – 2026-10-02T21:57:03+00:00 |
| 2026-10-01 | 64.5 | 85 | 2026-09-24T23:57:02+00:00 – 2026-10-01T23:57:02+00:00 |
| 2026-09-30 | 64.2 | 87 | 2026-09-23T23:57:03+00:00 – 2026-09-30T23:57:03+00:00 |
| 2026-09-29 | 61.5 | 125 | 2026-09-22T23:57:03+00:00 – 2026-09-29T23:57:03+00:00 |
| 2026-09-28 | 55.9 | 101 | 2026-09-21T23:57:03+00:00 – 2026-09-28T23:57:03+00:00 |
| 2026-09-27 | 63.0 | 61 | 2026-09-20T23:57:03+00:00 – 2026-09-27T23:57:03+00:00 |
| 2026-09-26 | 62.6 | 58 | 2026-09-19T23:57:03+00:00 – 2026-09-26T23:57:03+00:00 |
| 2026-09-25 | 60.5 | 52 | 2026-09-18T23:57:03+00:00 – 2026-09-25T23:57:03+00:00 |
| 2026-09-24 | 59.7 | 51 | 2026-09-17T23:57:03+00:00 – 2026-09-24T23:57:03+00:00 |
| 2026-09-23 | 58.5 | 47 | 2026-09-16T23:57:03+00:00 – 2026-09-23T23:57:03+00:00 |
Hacker News · 64.1 · 84 scored comments · Read platform reviews →
ONE MODEL, DIFFERENT COMMUNITIES
Across the communities
Scored separately for each platform, with different samples and audiences. Platform scores are not simply averaged; scored counts exclude quota-only comments.
Scores range from 0 to 100. Select a platform for its trend and reviews. Latest opinion time is not crawler health.
Explore long-term changes & analysis
This version is not yet tracked since release; recent 7-day experience is shown above.
Community reviews
Selected comments from the last 7 days. The balance of excerpts does not represent the share of positive reviews.
0 selected excerpts
Community reports about limits cannot establish official allowances.
Comment context
I switched to sonnet for my implementation agents when it released and it seemed fine at first, then I let it go on a big task and when I got home at the end of the day a sonnet agent had been stuck for almost 8 hours trying to solve something. I canceled it, deleted the branch and restarted with opus 5.5 and it finished the entire task in about 30 minutes.
Comment context
I have some on going projects like extensive high level research, personal game development project working with UE5 and blender through MCPs, and improving personal profile and portfolio.
Comment context
And it's SUPER usage-efficient. It's so efficient that I've had to revive a few old lanes I'd paused because I had to budget carefully for Astra usage. Otherwise, I wouldn't be able to use up my two accounts, and they'd last me about five days running Astra at low to medium reasoning. Opus 5.5 is still the best, and just as usage-efficient.
Comment context
I've ran so many prompts on both of them. The reality is when Codex works, it's great, but Astra Low, same prompt as Opus 5.5 low, the Delta is nuts.
Comment context
It also has different strengths to Opus5.5 - I find Opus better technically at writing code and avoiding bloat, but Astra is substantially better at developing original ideas and exploring topics.
Comment context
What’s with GPT 6.1 and Opus 5.5 both refusing to type passwords for smoke tests? I’ve got a local environment setup for my testing and staging environment, I explain to both Claude and GPT it’s all local and a staging environment and both models refuse to make up passwords for test accounts I want them to walk through and create. They both keep stopping and asking me to do it.
Show original
路由的claude opus 5.5 没有争议,已经取关了。 什么,你说opus 5.5思考速度没那么快?
Comment context
For example if you have Opus 5 process a bunch of reviews from an subagent fanout that is reviewing code, then synthesizing those into a final report. It keeps more of the actual defects vs 5.5, by a small margin. But it also keeps more false positive and takes about twice as long and twice as many tokens as 5.5
Comment context
Everyone is praising Opus but it flags even simple html fixes, GLM never let me down! I was playing with Opus 5.5 asking to improve the marketing texts on our html pages, and suddenly it started to reject it
Comment context
Opus 5.5 (O5.5) has been exceptional for me in Claude Code over the last six days. Architecture-first, DRY/SOLID coding out of the box, exceptional communication style, phenomenal token efficiency. I’ve been working with it basically nonstop, ~12 hours a day, since last Wednesday. But right around the time my monthly limit reset this evening, I noticed an extreme shift in communication style and coding behavior that feels suspiciously like Opus 5 (O5). For the last week, O5.5 would take my requirements and immediately get to work, usually knocking out a feature in minutes and doing an incredible job across implementation, UX, architecture, and token efficiency. Tonight, I’m seeing something very different.
Where O5.5 consistently respected DRY principles and the existing architecture, the modules created tonight are suddenly happy to reinvent every wheel that came before them. The difference in ability is not subtle. The other tell is token usage. O5.5 has been startlingly efficient compared with O5, which was an absolute token-eating monster. I’ve been getting substantially more feature work done with O5.5 at a fraction of the token usage. Tonight, simple tasks and relatively small features are suddenly devouring tokens at a pace that feels much more like O5. I’m especially sensitive to this because I have to request an increase in my spend limit every time I hit it. On O5, I was sending a request almost every day. On O5.5, I hadn’t needed to send a single one all week. Then tonight I jumped from roughly 70% to 90% usage in about an hour.
Comment context
SOL 5.6 and 6.1 both are monsters for high level backend coding.
Show original
这个确实牛逼在, Claude 整出来的这个算法研究本身,看起来确实有些新东西
Comment context
It started from one prompt I typed into Claude Code: "*using only pixels directly remake a beautiful pokemon red remake … absolute pixel art perfection.
* No image files at all. The game draws into a 320×180 screen pixel by pixel.
Comment context
I gave Opus 5.5 my Minecraft mod file and told it to create a Trailer
No selected comments for this filter in the last 7 days.