Claude · Last 7 days
Claude Sonnet 5.5
0–100 · higher means more positive community experience, not benchmark performance.
Updated 2026-10-01 11:22 UTC · 2026-09-24 – 2026-10-01 UTCExperience score65.4/100113 comments in 7 days
What people think
One user added a mini router that dispatches harder coding slices to Sonnet 5.5, and the setup improved first-pass implementation quality. One user states they wouldn't trust Sonnet to provide any architecture or design advice.
About the sample and scores
As of 2026-10-01 11:27:03 UTC, Claude Sonnet 5.5 in the Claude family has 113 explicitly attributed community comments in the last 7 days; its community score is 65.4/100. Available category scores include General text 67.2/100 (n=55); Coding 75.0/100 (n=8); Reasoning 63.6/100 (n=23). Scoring method and sources
Community reviews
Selected comments from the last 7 days. The balance of excerpts does not represent the share of positive reviews.
1 selected excerpts
Community reports about limits cannot establish official allowances.
Comment context
I learned through my testing that sonnet 5.5 sometimes spawns opus 5.5 subagents. Might be one of the main reasons why sonnet 5.5 feels a lot more expensive for people.
And, this is not about the API price, but the percentage weekly usage of the subscription.
Comment context
I actually added a mini router in front of it this week, and the harder slices get dispatched to sonnet 5.5 instead, pretty effective so far, it improved first-pass implementation quality.
Show original
这玩意儿哪里都不如Opus 5.5,也是给A畜学到国产小参数做题家模型雷霆大思考刷分的精髓了
Comment context
So I don't know what's your process but I would suggest models wise to just use Opus 5.5 high or medium as the orchestrator to supervise a Sonnet 5.5 high implementer and use 6.1 sol xhigh (or medium if you notice xhigh is not sustainable for you) as a reviewer.
Show original
如果你是用API,那就是路边一条,思考等级定低了就是傻狗,定高了成本就朝Opus那边去了
Comment context
It really depends on the task, but it's often either sonnet 5.5 high or opus 5.5 med.
Show original
丫抠字儿太快了 找她吹牛逼容易跟不上思路
Show original
依靠无限制拉长思考来换取所谓的更高智商,感觉可能不是正确方向
Show original
sonnet5.5审美确实一眼高级
Show original
追问,发现有幻觉
Comment context
I discussed a JEV-related question and found that Company A's models now have a certain characteristic.
Machine translatedShow original
我拿一个跟JEV相关的问题讨论了一下,发现A家的模型现在有个特点。
Show original
在智能体命令行编程测试 Terminal-Bench 4.0 中,它以 70.6% 的高分拿下 全球第一 ,相比 Sonnet 5 的 10.3% 属于跨越式暴涨
Show original
输出速度快了30%
Comment context
The price is the same as the previous generation Sonnet5, 50% cheaper than Opus5.5, output speed increased by 30%, and the cost per task is up to 30% cheaper than the previous generation. If I make a table it will be more obvious.
Machine translatedShow original
价格方面与上一代sonnet5一样,比opus5.5便宜了50%,输出速度快了30%,每项任务的成本比上一代最多便宜30%,我做一个表格就能更明显看出来了。
Show original
A社的账号"娇贵"啊,隔三岔五给你大陆ip全干挺了
Show original
才发现Sonnet5.5还挺耐用啊
Show original
思考用更多token导致更慢
Show original
免费用户也能用几下sonnet5.5,反观隔壁免费用户只能用luna
Show original
同思考档位的sonnet 5.5 比 opus 消耗token更多,效率更低
Show original
单价只有Opus一半,额度还给得更多,还学会了O社时不时发重置次数
Comment context
I've never once hit a safeguard refusal. Dunno what you guys are doing
Comment context
If Astra is that great, how come its code introduces so many bugs? Give Astra a complicated task on Ultra, and then after give the prompt "review" - you will get about 5 problems.
Comment context
Tokens | cost to serve/run the same agentic task | performance across multiple benchmarks
No selected comments for this filter in the last 7 days.