Claude · Last 7 days
Claude Sonnet 5
0–100 · higher means more positive community experience, not benchmark performance.
Updated 2026-10-02 14:32 UTC · 2026-09-25 – 2026-10-02 UTCRECENT EXPERIENCE
How the Experience Index is changing
30 days of rolling 7-day scores at each daily snapshot. The vertical scale adapts to the data. Missing or insufficient samples leave gaps; today is still updating.
Daily readings & sample sizes
| Date | Score | n | Scoring window (UTC) |
|---|---|---|---|
| 2026-10-02 | 30.7 | 8 | 2026-09-25T14:32:03+00:00 – 2026-10-02T14:32:03+00:00 |
| 2026-10-01 | 30.7 | 8 | 2026-09-24T23:57:02+00:00 – 2026-10-01T23:57:02+00:00 |
| 2026-09-30 | 30.7 | 8 | 2026-09-23T23:57:03+00:00 – 2026-09-30T23:57:03+00:00 |
| 2026-09-29 | 31.7 | 9 | 2026-09-22T23:57:03+00:00 – 2026-09-29T23:57:03+00:00 |
| 2026-09-28 | 31.8 | 9 | 2026-09-21T23:57:03+00:00 – 2026-09-28T23:57:03+00:00 |
| 2026-09-27 | — | 3 | 2026-09-20T23:57:03+00:00 – 2026-09-27T23:57:03+00:00 |
| 2026-09-26 | — | 2 | 2026-09-19T23:57:03+00:00 – 2026-09-26T23:57:03+00:00 |
| 2026-09-25 | — | 2 | 2026-09-18T23:57:03+00:00 – 2026-09-25T23:57:03+00:00 |
| 2026-09-24 | — | 2 | 2026-09-17T23:57:03+00:00 – 2026-09-24T23:57:03+00:00 |
| 2026-09-23 | — | 1 | 2026-09-16T23:57:03+00:00 – 2026-09-23T23:57:03+00:00 |
| 2026-09-22 | — | 2 | 2026-09-15T23:57:03+00:00 – 2026-09-22T23:57:03+00:00 |
| 2026-09-21 | — | 2 | 2026-09-14T23:57:03+00:00 – 2026-09-21T23:57:03+00:00 |
| 2026-09-20 | — | 2 | 2026-09-13T23:57:03+00:00 – 2026-09-20T23:57:03+00:00 |
| 2026-09-19 | — | 2 | 2026-09-12T23:57:03+00:00 – 2026-09-19T23:57:03+00:00 |
| 2026-09-18 | — | 2 | 2026-09-11T23:57:02+00:00 – 2026-09-18T23:57:02+00:00 |
| 2026-09-17 | — | 3 | 2026-09-10T23:57:03+00:00 – 2026-09-17T23:57:03+00:00 |
| 2026-09-16 | — | 3 | 2026-09-09T23:57:03+00:00 – 2026-09-16T23:57:03+00:00 |
| 2026-09-15 | — | 2 | 2026-09-08T23:57:03+00:00 – 2026-09-15T23:57:03+00:00 |
| 2026-09-14 | — | 1 | 2026-09-07T23:57:03+00:00 – 2026-09-14T23:57:03+00:00 |
| 2026-09-13 | — | 1 | 2026-09-06T23:57:03+00:00 – 2026-09-13T23:57:03+00:00 |
| 2026-09-12 | — | 3 | 2026-09-05T23:57:03+00:00 – 2026-09-12T23:57:03+00:00 |
| 2026-09-11 | — | 3 | 2026-09-04T23:00:04+00:00 – 2026-09-11T23:00:04+00:00 |
| 2026-09-10 | — | 4 | 2026-09-03T23:00:05+00:00 – 2026-09-10T23:00:05+00:00 |
| 2026-09-09 | 58.5 | 6 | 2026-09-02T23:00:04+00:00 – 2026-09-09T23:00:04+00:00 |
| 2026-09-08 | 50.7 | 8 | 2026-09-01T23:00:03+00:00 – 2026-09-08T23:00:03+00:00 |
| 2026-09-07 | 50.7 | 8 | 2026-08-31T23:00:03+00:00 – 2026-09-07T23:00:03+00:00 |
| 2026-09-06 | 46.9 | 7 | 2026-08-30T23:00:02+00:00 – 2026-09-06T23:00:02+00:00 |
| 2026-09-05 | 42.6 | 6 | 2026-08-29T23:00:03+00:00 – 2026-09-05T23:00:03+00:00 |
| 2026-09-04 | 38.4 | 5 | 2026-08-28T23:00:03+00:00 – 2026-09-04T23:00:03+00:00 |
| 2026-09-03 | — | 4 | 2026-08-27T23:00:03+00:00 – 2026-09-03T23:00:03+00:00 |
Hacker News · 30.7 · 8 scored comments · Read platform reviews →
ONE MODEL, DIFFERENT COMMUNITIES
Across the communities
Scored separately for each platform, with different samples and audiences. Platform scores are not simply averaged; scored counts exclude quota-only comments.
Scores range from 0 to 100. Select a platform for its trend and reviews. Latest opinion time is not crawler health.
What people think
One user finds that Opus 5.5 takes shortcuts on small side tasks, while Sonnet and Fable would have completed them properly. One user found their month-long experience with codex/GPT 5.6 noticeably better than Claude Sonnet 5 overall.
About the sample and scores
As of 2026-10-02 14:32:03 UTC, Claude Sonnet 5 in the Claude family has 30 explicitly attributed community comments in the last 7 days; its community score is 32.5/100. Available category scores include General text 38.4/100 (n=19); Speed & latency 34.9/100 (n=5). Scoring method and sources
Community reviews
Selected comments from the last 7 days. The balance of excerpts does not represent the share of positive reviews.
4 selected excerpts
Community reports about limits cannot establish official allowances.
Comment context
But after a few tasks I realised it was omitting huge details of complex tasks that Fable just would not have missed. It instead creates a version that passes its own tests and asserts that the job is done (which to be fair, all LLMs do). But in comparison to 5.1, its like it takes intentional shortcuts to get the job 'done'. It looks like it's been tuned to trim "non-essential" aspects of a task to get to the end faster. And I don't think this is isolated to large tasks that have room for interpretation. I've caught doing the same with much smaller side tasks that both Sonnet or Fable would have nailed every time.
Show original
啊?我还觉得sonnet5挺好用的,居然扑街了?
Comment context
I built an adversarial esoteric programming language to benchmark LLM models and just ran it on Sonnet 5.5 It does worse than Sonnet 5.
Comment context
At work we have copilot and even with 750$ credits I use opus very sparingly and I even replaced it with sonnet 5 recently because I wasn't very impressed with the results, can't wait we enable opus 5.5 next week.
Comment context
I asked it to find online prices for some industrial automation items.
No selected comments for this filter in the last 7 days.