Claude · Last 7 days

Claude Opus 5

0–100 · higher means more positive community experience, not benchmark performance.

Updated 2026-10-04 08:32 UTC · 2026-09-27 – 2026-10-04 UTC
All-platform score23.1/10066 comments in 7 days

Discussion overview

What the full sample discusses

All platforms · 66 eligible comments over 7 days, including usage-limit feedback; deduplicated by source and comment.

  • General text · 40 comments: 4 positive / 31 negative / 5 mixed or neutral
  • Reasoning · 15 comments: 5 positive / 10 negative / 0 mixed or neutral
  • Usage limits · 5 comments: 2 positive / 2 negative / 1 mixed or neutral
  • Speed & latency · 4 comments: 0 positive / 3 negative / 1 mixed or neutral

Comments may cover multiple dimensions; counts are not additive or equivalent to the weighted score. Links show selected source excerpts.

Recent change

Rolling 7-day score: −4.0 points vs 7 days ago. Discussion mix can affect scores; this does not establish a cause.

Source coverage and discussion concentration

Hacker News 7 · Reddit 52 · v2ex 1 · RedNote 1 · Zhihu 5

48 identifiable discussions cover 59/66 comments; the largest has 5. Remaining thread identities are unknown; this is not a count of independent users.

Selected individual opinions

One user fed both models a script meant to print the max of two numbers; Opus 5 deduced it correctly, while Opus 5.5 wrongly concluded it prints 1/0. One user replied that even after the nerf, Opus 5.5 is still much better than Opus 5.

About the sample and scores

As of 2026-10-04 08:32:03 UTC, Claude Opus 5 in the Claude family has 66 explicitly attributed community comments in the last 7 days; its community score is 23.1/100. Available category scores include General text 21.1/100 (n=40); Reasoning 38.3/100 (n=15); Usage limits 49.3/100 (n=5). Scoring method and sources

RECENT EXPERIENCE

How the Experience Index is changing

−4.0vs 7 days ago · points

30 days of rolling 7-day scores at each daily snapshot. The vertical scale adapts to the data. Missing or insufficient samples leave gaps; today is still updating.

Tap or use ← → for dates, scores and sample sizes. Blank dates have insufficient data.

Claude Opus 5 · All platforms · 30 days of rolling 7-day experience scores 2030405060 Observed-day mean: 32.1 2026-09-05 · 29.3 · n=185 · 2026-08-29T23:00:03+00:00 – 2026-09-05T23:00:03+00:002026-09-06 · 29.7 · n=185 · 2026-08-30T23:00:02+00:00 – 2026-09-06T23:00:02+00:002026-09-07 · 29.9 · n=197 · 2026-08-31T23:00:03+00:00 – 2026-09-07T23:00:03+00:002026-09-08 · 31.2 · n=183 · 2026-09-01T23:00:03+00:00 – 2026-09-08T23:00:03+00:002026-09-09 · 31.6 · n=180 · 2026-09-02T23:00:04+00:00 – 2026-09-09T23:00:04+00:002026-09-10 · 31.4 · n=175 · 2026-09-03T23:00:05+00:00 – 2026-09-10T23:00:05+00:002026-09-11 · 33.3 · n=149 · 2026-09-04T23:00:04+00:00 – 2026-09-11T23:00:04+00:002026-09-12 · 34.0 · n=158 · 2026-09-05T23:57:03+00:00 – 2026-09-12T23:57:03+00:002026-09-13 · 35.7 · n=160 · 2026-09-06T23:57:03+00:00 – 2026-09-13T23:57:03+00:002026-09-14 · 37.3 · n=161 · 2026-09-07T23:57:03+00:00 – 2026-09-14T23:57:03+00:002026-09-15 · 37.4 · n=175 · 2026-09-08T23:57:03+00:00 – 2026-09-15T23:57:03+00:002026-09-16 · 38.3 · n=194 · 2026-09-09T23:57:03+00:00 – 2026-09-16T23:57:03+00:002026-09-17 · 40.1 · n=198 · 2026-09-10T23:57:03+00:00 – 2026-09-17T23:57:03+00:002026-09-18 · 40.1 · n=212 · 2026-09-11T23:57:02+00:00 – 2026-09-18T23:57:02+00:002026-09-19 · 41.8 · n=192 · 2026-09-12T23:57:03+00:00 – 2026-09-19T23:57:03+00:002026-09-20 · 43.6 · n=190 · 2026-09-13T23:57:03+00:00 – 2026-09-20T23:57:03+00:002026-09-21 · 42.5 · n=219 · 2026-09-14T23:57:03+00:00 – 2026-09-21T23:57:03+00:002026-09-22 · 38.6 · n=242 · 2026-09-15T23:57:03+00:00 – 2026-09-22T23:57:03+00:002026-09-23 · 37.2 · n=254 · 2026-09-16T23:57:03+00:00 – 2026-09-23T23:57:03+00:002026-09-24 · 34.0 · n=239 · 2026-09-17T23:57:03+00:00 – 2026-09-24T23:57:03+00:002026-09-25 · 34.2 · n=219 · 2026-09-18T23:57:03+00:00 – 2026-09-25T23:57:03+00:002026-09-26 · 29.6 · n=209 · 2026-09-19T23:57:03+00:00 – 2026-09-26T23:57:03+00:002026-09-27 · 27.1 · n=200 · 2026-09-20T23:57:03+00:00 – 2026-09-27T23:57:03+00:002026-09-28 · 23.2 · n=156 · 2026-09-21T23:57:03+00:00 – 2026-09-28T23:57:03+00:002026-09-29 · 24.0 · n=100 · 2026-09-22T23:57:03+00:00 – 2026-09-29T23:57:03+00:002026-09-30 · 20.8 · n=80 · 2026-09-23T23:57:03+00:00 – 2026-09-30T23:57:03+00:002026-10-01 · 20.1 · n=76 · 2026-09-24T23:57:02+00:00 – 2026-10-01T23:57:02+00:002026-10-02 · 21.5 · n=72 · 2026-09-25T23:57:03+00:00 – 2026-10-02T23:57:03+00:002026-10-03 · 22.7 · n=70 · 2026-09-26T23:57:02+00:00 – 2026-10-03T23:57:02+00:002026-10-04 · 23.1 · n=66 · 2026-09-27T08:32:03+00:00 – 2026-10-04T08:32:03+00:00 23.1 09-0509-1309-2009-2710-04 Claude Opus 5 · All platforms · 30 days of rolling 7-day experience scores 2030405060 Observed-day mean: 32.1 2026-09-05 · 29.3 · n=185 · 2026-08-29T23:00:03+00:00 – 2026-09-05T23:00:03+00:002026-09-06 · 29.7 · n=185 · 2026-08-30T23:00:02+00:00 – 2026-09-06T23:00:02+00:002026-09-07 · 29.9 · n=197 · 2026-08-31T23:00:03+00:00 – 2026-09-07T23:00:03+00:002026-09-08 · 31.2 · n=183 · 2026-09-01T23:00:03+00:00 – 2026-09-08T23:00:03+00:002026-09-09 · 31.6 · n=180 · 2026-09-02T23:00:04+00:00 – 2026-09-09T23:00:04+00:002026-09-10 · 31.4 · n=175 · 2026-09-03T23:00:05+00:00 – 2026-09-10T23:00:05+00:002026-09-11 · 33.3 · n=149 · 2026-09-04T23:00:04+00:00 – 2026-09-11T23:00:04+00:002026-09-12 · 34.0 · n=158 · 2026-09-05T23:57:03+00:00 – 2026-09-12T23:57:03+00:002026-09-13 · 35.7 · n=160 · 2026-09-06T23:57:03+00:00 – 2026-09-13T23:57:03+00:002026-09-14 · 37.3 · n=161 · 2026-09-07T23:57:03+00:00 – 2026-09-14T23:57:03+00:002026-09-15 · 37.4 · n=175 · 2026-09-08T23:57:03+00:00 – 2026-09-15T23:57:03+00:002026-09-16 · 38.3 · n=194 · 2026-09-09T23:57:03+00:00 – 2026-09-16T23:57:03+00:002026-09-17 · 40.1 · n=198 · 2026-09-10T23:57:03+00:00 – 2026-09-17T23:57:03+00:002026-09-18 · 40.1 · n=212 · 2026-09-11T23:57:02+00:00 – 2026-09-18T23:57:02+00:002026-09-19 · 41.8 · n=192 · 2026-09-12T23:57:03+00:00 – 2026-09-19T23:57:03+00:002026-09-20 · 43.6 · n=190 · 2026-09-13T23:57:03+00:00 – 2026-09-20T23:57:03+00:002026-09-21 · 42.5 · n=219 · 2026-09-14T23:57:03+00:00 – 2026-09-21T23:57:03+00:002026-09-22 · 38.6 · n=242 · 2026-09-15T23:57:03+00:00 – 2026-09-22T23:57:03+00:002026-09-23 · 37.2 · n=254 · 2026-09-16T23:57:03+00:00 – 2026-09-23T23:57:03+00:002026-09-24 · 34.0 · n=239 · 2026-09-17T23:57:03+00:00 – 2026-09-24T23:57:03+00:002026-09-25 · 34.2 · n=219 · 2026-09-18T23:57:03+00:00 – 2026-09-25T23:57:03+00:002026-09-26 · 29.6 · n=209 · 2026-09-19T23:57:03+00:00 – 2026-09-26T23:57:03+00:002026-09-27 · 27.1 · n=200 · 2026-09-20T23:57:03+00:00 – 2026-09-27T23:57:03+00:002026-09-28 · 23.2 · n=156 · 2026-09-21T23:57:03+00:00 – 2026-09-28T23:57:03+00:002026-09-29 · 24.0 · n=100 · 2026-09-22T23:57:03+00:00 – 2026-09-29T23:57:03+00:002026-09-30 · 20.8 · n=80 · 2026-09-23T23:57:03+00:00 – 2026-09-30T23:57:03+00:002026-10-01 · 20.1 · n=76 · 2026-09-24T23:57:02+00:00 – 2026-10-01T23:57:02+00:002026-10-02 · 21.5 · n=72 · 2026-09-25T23:57:03+00:00 – 2026-10-02T23:57:03+00:002026-10-03 · 22.7 · n=70 · 2026-09-26T23:57:02+00:00 – 2026-10-03T23:57:02+00:002026-10-04 · 23.1 · n=66 · 2026-09-27T08:32:03+00:00 – 2026-10-04T08:32:03+00:00 23.1 09-0509-1309-2009-2710-04
Daily readings & sample sizes
DateScorenScoring window (UTC)
2026-10-0423.1662026-09-27T08:32:03+00:00 – 2026-10-04T08:32:03+00:00
2026-10-0322.7702026-09-26T23:57:02+00:00 – 2026-10-03T23:57:02+00:00
2026-10-0221.5722026-09-25T23:57:03+00:00 – 2026-10-02T23:57:03+00:00
2026-10-0120.1762026-09-24T23:57:02+00:00 – 2026-10-01T23:57:02+00:00
2026-09-3020.8802026-09-23T23:57:03+00:00 – 2026-09-30T23:57:03+00:00
2026-09-2924.01002026-09-22T23:57:03+00:00 – 2026-09-29T23:57:03+00:00
2026-09-2823.21562026-09-21T23:57:03+00:00 – 2026-09-28T23:57:03+00:00
2026-09-2727.12002026-09-20T23:57:03+00:00 – 2026-09-27T23:57:03+00:00
2026-09-2629.62092026-09-19T23:57:03+00:00 – 2026-09-26T23:57:03+00:00
2026-09-2534.22192026-09-18T23:57:03+00:00 – 2026-09-25T23:57:03+00:00
2026-09-2434.02392026-09-17T23:57:03+00:00 – 2026-09-24T23:57:03+00:00
2026-09-2337.22542026-09-16T23:57:03+00:00 – 2026-09-23T23:57:03+00:00
2026-09-2238.62422026-09-15T23:57:03+00:00 – 2026-09-22T23:57:03+00:00
2026-09-2142.52192026-09-14T23:57:03+00:00 – 2026-09-21T23:57:03+00:00
2026-09-2043.61902026-09-13T23:57:03+00:00 – 2026-09-20T23:57:03+00:00
2026-09-1941.81922026-09-12T23:57:03+00:00 – 2026-09-19T23:57:03+00:00
2026-09-1840.12122026-09-11T23:57:02+00:00 – 2026-09-18T23:57:02+00:00
2026-09-1740.11982026-09-10T23:57:03+00:00 – 2026-09-17T23:57:03+00:00
2026-09-1638.31942026-09-09T23:57:03+00:00 – 2026-09-16T23:57:03+00:00
2026-09-1537.41752026-09-08T23:57:03+00:00 – 2026-09-15T23:57:03+00:00
2026-09-1437.31612026-09-07T23:57:03+00:00 – 2026-09-14T23:57:03+00:00
2026-09-1335.71602026-09-06T23:57:03+00:00 – 2026-09-13T23:57:03+00:00
2026-09-1234.01582026-09-05T23:57:03+00:00 – 2026-09-12T23:57:03+00:00
2026-09-1133.31492026-09-04T23:00:04+00:00 – 2026-09-11T23:00:04+00:00
2026-09-1031.41752026-09-03T23:00:05+00:00 – 2026-09-10T23:00:05+00:00
2026-09-0931.61802026-09-02T23:00:04+00:00 – 2026-09-09T23:00:04+00:00
2026-09-0831.21832026-09-01T23:00:03+00:00 – 2026-09-08T23:00:03+00:00
2026-09-0729.91972026-08-31T23:00:03+00:00 – 2026-09-07T23:00:03+00:00
2026-09-0629.71852026-08-30T23:00:02+00:00 – 2026-09-06T23:00:02+00:00
2026-09-0529.31852026-08-29T23:00:03+00:00 – 2026-09-05T23:00:03+00:00

ONE MODEL, DIFFERENT COMMUNITIES

Across the communities

Last 7 days

Scored separately for each platform, with different samples and audiences. Platform scores are not simply averaged; scored counts exclude quota-only comments.

Scores range from 0 to 100. Select a platform for its trend and reviews. Latest opinion time is not crawler health.

Community reviews

Selected comments from the last 7 days. The balance of excerpts does not represent the share of positive reviews.

2 selected excerpts

Community reports about limits cannot establish official allowances.

NeutralGeneral text

“keeps more of the actual defects vs 5.5, by a small margin”

Comment context
From the same comment ↗

For example if you have Opus 5 process a bunch of reviews from an subagent fanout that is reviewing code, then synthesizing those into a final report. It keeps more of the actual defects vs 5.5, by a small margin. But it also keeps more false positive and takes about twice as long and twice as many tokens as 5.5

Explore long-term changes & analysis

Model lifecycle

Claude Opus 5

Released 2026-07-24 · Back to Claude

Exact-version signals observed 2026-09-21–2026-10-03 · n=131

Data window: 2026-09-06 00:00:00 to 2026-10-04 00:00:00 UTC · Scoring method: experience_score_v3.

Insufficient historical baseline The original baseline window lacks enough current-method evidence for a long-term comparison.

Lifecycle updated daily · latest complete UTC day 2026-10-03

The current classification method has insufficient evidence in the original baseline window. The window is preserved; no long-term improvement or decline is inferred, and old-method scores are not compared.

Fixed 14-day baseline
–
n=0
Latest 28 days
21.6
n=123
Change vs baseline
–
90-day slope
–
pts / 30 days

14-day rolling experience

Each daily point summarizes the previous 14 complete UTC days. The dashed line is the fixed release baseline; the verdict still uses independent non-overlapping 14-day segments.

14-day rolling experience20304050602026-09-13–2026-09-27 · 25.2 · n=682026-09-14–2026-09-28 · 23.5 · n=772026-09-15–2026-09-29 · 22.7 · n=932026-09-16–2026-09-30 · 23.0 · n=1052026-09-17–2026-10-01 · 21.6 · n=1112026-09-18–2026-10-02 · 22.3 · n=1182026-09-19–2026-10-03 · 21.6 · n=12108-1308-2208-3109-0909-1809-2710-03

0 of 4 independent 14-day segments ready for trend judgment.

What explains the change

Experience change and discussion-mix change are separated. These are observational contributions, not proof of cause.

No category has enough samples in both windows for a contribution conclusion.

Category Baseline → current Category change Weight share Experience contribution Discussion-mix contribution
Coding Baseline → currentnot enough datan=0 → 10 Category change– Weight share– Experience contribution– Discussion-mix contribution–
Image generation Baseline → currentnot enough datan=0 → 0 Category change– Weight share– Experience contribution– Discussion-mix contribution–
Image understanding Baseline → currentnot enough datan=0 → 0 Category change– Weight share– Experience contribution– Discussion-mix contribution–
Local deploy Baseline → currentnot enough datan=0 → 0 Category change– Weight share– Experience contribution– Discussion-mix contribution–
Reasoning Baseline → currentnot enough datan=0 → 29 Category change– Weight share– Experience contribution– Discussion-mix contribution–
Roleplay / creative Baseline → currentnot enough datan=0 → 0 Category change– Weight share– Experience contribution– Discussion-mix contribution–
Safety & refusals Baseline → currentnot enough datan=0 → 6 Category change– Weight share– Experience contribution– Discussion-mix contribution–
Speed & latency Baseline → currentnot enough datan=0 → 9 Category change– Weight share– Experience contribution– Discussion-mix contribution–
General text Baseline → currentnot enough datan=0 → 70 Category change– Weight share– Experience contribution– Discussion-mix contribution–
Video generation Baseline → currentnot enough datan=0 → 0 Category change– Weight share– Experience contribution– Discussion-mix contribution–

Lifecycle scores use the same experience-signal weights as the main index, without sample shrinkage after n=30. They measure public user perception, not model capability or backend causes.

How's your AI experience today?