GPT · Last 7 days

GPT-5.6 Sol

0–100 · higher means more positive community experience, not benchmark performance.

Updated 2026-10-03 09:17 UTC · 2026-09-26 – 2026-10-03 UTC
All-platform score46.4/100100 comments in 7 days

Discussion overview

What the full sample discusses

All platforms · 100 eligible comments over 7 days, including usage-limit feedback; deduplicated by source and comment.

  • General text · 48 comments: 22 positive / 24 negative / 2 mixed or neutral
  • Reasoning · 20 comments: 12 positive / 8 negative / 0 mixed or neutral
  • Usage limits · 16 comments: 3 positive / 12 negative / 1 mixed or neutral
  • Coding · 8 comments: 3 positive / 4 negative / 1 mixed or neutral

Comments may cover multiple dimensions; counts are not additive or equivalent to the weighted score. Links show selected source excerpts.

Recent change

Rolling 7-day score: −5.5 points vs 7 days ago. Discussion mix can affect scores; this does not establish a cause.

Source coverage and discussion concentration

Hacker News 8 · Reddit 75 · RedNote 2 · Zhihu 15

50 identifiable discussions cover 92/100 comments; the largest has 10. Remaining thread identities are unknown; this is not a count of independent users.

Selected individual opinions

One user still uses GPT-5.6 Sol in their custom Codex coding harness One user tested GPT-5.6 Sol on the "Three-color Garden" problem, but it could not solve it, whereas GPT-6 Astra solved it in about 30 minutes.

About the sample and scores

As of 2026-10-03 09:17:02 UTC, GPT-5.6 Sol in the GPT family has 100 explicitly attributed community comments in the last 7 days; its community score is 46.4/100. Available category scores include General text 49.1/100 (n=48); Coding 44.5/100 (n=8); Reasoning 54.0/100 (n=20). Scoring method and sources

RECENT EXPERIENCE

How the Experience Index is changing

+3.4vs 7 days ago · points

30 days of rolling 7-day scores at each daily snapshot. The vertical scale adapts to the data. Missing or insufficient samples leave gaps; today is still updating.

Tap or use ← → for dates, scores and sample sizes. Blank dates have insufficient data.

GPT-5.6 Sol · Hacker News · 30 days of rolling 7-day experience scores 3040506070 Observed-day mean: 47.7 2026-09-04 · 63.8 · n=11 · 2026-08-28T23:00:03+00:00 – 2026-09-04T23:00:03+00:002026-09-05 · 62.2 · n=10 · 2026-08-29T23:00:03+00:00 – 2026-09-05T23:00:03+00:002026-09-06 · 62.2 · n=10 · 2026-08-30T23:00:02+00:00 – 2026-09-06T23:00:02+00:002026-09-07 · 62.2 · n=10 · 2026-08-31T23:00:03+00:00 – 2026-09-07T23:00:03+00:002026-09-08 · 55.5 · n=8 · 2026-09-01T23:00:03+00:00 – 2026-09-08T23:00:03+00:002026-09-09 · 48.2 · n=10 · 2026-09-02T23:00:04+00:00 – 2026-09-09T23:00:04+00:002026-09-10 · 36.8 · n=8 · 2026-09-03T23:00:05+00:00 – 2026-09-10T23:00:05+00:002026-09-11 · 39.4 · n=7 · 2026-09-04T23:00:04+00:00 – 2026-09-11T23:00:04+00:002026-09-12 · 39.4 · n=7 · 2026-09-05T23:57:03+00:00 – 2026-09-12T23:57:03+00:002026-09-13 · 41.9 · n=9 · 2026-09-06T23:57:03+00:00 – 2026-09-13T23:57:03+00:002026-09-14 · 41.9 · n=9 · 2026-09-07T23:57:03+00:00 – 2026-09-14T23:57:03+00:002026-09-15 · 44.5 · n=9 · 2026-09-08T23:57:03+00:00 – 2026-09-15T23:57:03+00:002026-09-16 · 44.9 · n=7 · 2026-09-09T23:57:03+00:00 – 2026-09-16T23:57:03+00:002026-09-17 · 43.4 · n=6 · 2026-09-10T23:57:03+00:00 – 2026-09-17T23:57:03+00:002026-09-18 · 43.4 · n=6 · 2026-09-11T23:57:02+00:00 – 2026-09-18T23:57:02+00:002026-09-19 · 39.9 · n=7 · 2026-09-12T23:57:03+00:00 – 2026-09-19T23:57:03+00:002026-09-20 · 37.3 · n=5 · 2026-09-13T23:57:03+00:00 – 2026-09-20T23:57:03+00:002026-09-21 · 32.8 · n=7 · 2026-09-14T23:57:03+00:00 – 2026-09-21T23:57:03+00:002026-09-22 · 35.3 · n=8 · 2026-09-15T23:57:03+00:00 – 2026-09-22T23:57:03+00:002026-09-23 · 38.1 · n=7 · 2026-09-16T23:57:03+00:00 – 2026-09-23T23:57:03+00:002026-09-24 · 43.3 · n=5 · 2026-09-17T23:57:03+00:00 – 2026-09-24T23:57:03+00:002026-09-25 · 47.6 · n=6 · 2026-09-18T23:57:03+00:00 – 2026-09-25T23:57:03+00:002026-09-26 · 50.3 · n=6 · 2026-09-19T23:57:03+00:00 – 2026-09-26T23:57:03+00:002026-09-27 · 50.3 · n=6 · 2026-09-20T23:57:03+00:00 – 2026-09-27T23:57:03+00:002026-09-28 · 52.5 · n=5 · 2026-09-21T23:57:03+00:00 – 2026-09-28T23:57:03+00:002026-09-29 · 56.0 · n=7 · 2026-09-22T23:57:03+00:00 – 2026-09-29T23:57:03+00:002026-09-30 · 55.0 · n=9 · 2026-09-23T23:57:03+00:00 – 2026-09-30T23:57:03+00:002026-10-01 · 55.0 · n=9 · 2026-09-24T23:57:02+00:00 – 2026-10-01T23:57:02+00:002026-10-02 · 53.8 · n=7 · 2026-09-25T23:57:03+00:00 – 2026-10-02T23:57:03+00:002026-10-03 · 53.8 · n=7 · 2026-09-26T09:17:02+00:00 – 2026-10-03T09:17:02+00:00 53.8 09-0409-1209-1909-2610-03 GPT-5.6 Sol · Hacker News · 30 days of rolling 7-day experience scores 3040506070 Observed-day mean: 47.7 2026-09-04 · 63.8 · n=11 · 2026-08-28T23:00:03+00:00 – 2026-09-04T23:00:03+00:002026-09-05 · 62.2 · n=10 · 2026-08-29T23:00:03+00:00 – 2026-09-05T23:00:03+00:002026-09-06 · 62.2 · n=10 · 2026-08-30T23:00:02+00:00 – 2026-09-06T23:00:02+00:002026-09-07 · 62.2 · n=10 · 2026-08-31T23:00:03+00:00 – 2026-09-07T23:00:03+00:002026-09-08 · 55.5 · n=8 · 2026-09-01T23:00:03+00:00 – 2026-09-08T23:00:03+00:002026-09-09 · 48.2 · n=10 · 2026-09-02T23:00:04+00:00 – 2026-09-09T23:00:04+00:002026-09-10 · 36.8 · n=8 · 2026-09-03T23:00:05+00:00 – 2026-09-10T23:00:05+00:002026-09-11 · 39.4 · n=7 · 2026-09-04T23:00:04+00:00 – 2026-09-11T23:00:04+00:002026-09-12 · 39.4 · n=7 · 2026-09-05T23:57:03+00:00 – 2026-09-12T23:57:03+00:002026-09-13 · 41.9 · n=9 · 2026-09-06T23:57:03+00:00 – 2026-09-13T23:57:03+00:002026-09-14 · 41.9 · n=9 · 2026-09-07T23:57:03+00:00 – 2026-09-14T23:57:03+00:002026-09-15 · 44.5 · n=9 · 2026-09-08T23:57:03+00:00 – 2026-09-15T23:57:03+00:002026-09-16 · 44.9 · n=7 · 2026-09-09T23:57:03+00:00 – 2026-09-16T23:57:03+00:002026-09-17 · 43.4 · n=6 · 2026-09-10T23:57:03+00:00 – 2026-09-17T23:57:03+00:002026-09-18 · 43.4 · n=6 · 2026-09-11T23:57:02+00:00 – 2026-09-18T23:57:02+00:002026-09-19 · 39.9 · n=7 · 2026-09-12T23:57:03+00:00 – 2026-09-19T23:57:03+00:002026-09-20 · 37.3 · n=5 · 2026-09-13T23:57:03+00:00 – 2026-09-20T23:57:03+00:002026-09-21 · 32.8 · n=7 · 2026-09-14T23:57:03+00:00 – 2026-09-21T23:57:03+00:002026-09-22 · 35.3 · n=8 · 2026-09-15T23:57:03+00:00 – 2026-09-22T23:57:03+00:002026-09-23 · 38.1 · n=7 · 2026-09-16T23:57:03+00:00 – 2026-09-23T23:57:03+00:002026-09-24 · 43.3 · n=5 · 2026-09-17T23:57:03+00:00 – 2026-09-24T23:57:03+00:002026-09-25 · 47.6 · n=6 · 2026-09-18T23:57:03+00:00 – 2026-09-25T23:57:03+00:002026-09-26 · 50.3 · n=6 · 2026-09-19T23:57:03+00:00 – 2026-09-26T23:57:03+00:002026-09-27 · 50.3 · n=6 · 2026-09-20T23:57:03+00:00 – 2026-09-27T23:57:03+00:002026-09-28 · 52.5 · n=5 · 2026-09-21T23:57:03+00:00 – 2026-09-28T23:57:03+00:002026-09-29 · 56.0 · n=7 · 2026-09-22T23:57:03+00:00 – 2026-09-29T23:57:03+00:002026-09-30 · 55.0 · n=9 · 2026-09-23T23:57:03+00:00 – 2026-09-30T23:57:03+00:002026-10-01 · 55.0 · n=9 · 2026-09-24T23:57:02+00:00 – 2026-10-01T23:57:02+00:002026-10-02 · 53.8 · n=7 · 2026-09-25T23:57:03+00:00 – 2026-10-02T23:57:03+00:002026-10-03 · 53.8 · n=7 · 2026-09-26T09:17:02+00:00 – 2026-10-03T09:17:02+00:00 53.8 09-0409-1209-1909-2610-03
Daily readings & sample sizes
DateScorenScoring window (UTC)
2026-10-0353.872026-09-26T09:17:02+00:00 – 2026-10-03T09:17:02+00:00
2026-10-0253.872026-09-25T23:57:03+00:00 – 2026-10-02T23:57:03+00:00
2026-10-0155.092026-09-24T23:57:02+00:00 – 2026-10-01T23:57:02+00:00
2026-09-3055.092026-09-23T23:57:03+00:00 – 2026-09-30T23:57:03+00:00
2026-09-2956.072026-09-22T23:57:03+00:00 – 2026-09-29T23:57:03+00:00
2026-09-2852.552026-09-21T23:57:03+00:00 – 2026-09-28T23:57:03+00:00
2026-09-2750.362026-09-20T23:57:03+00:00 – 2026-09-27T23:57:03+00:00
2026-09-2650.362026-09-19T23:57:03+00:00 – 2026-09-26T23:57:03+00:00
2026-09-2547.662026-09-18T23:57:03+00:00 – 2026-09-25T23:57:03+00:00
2026-09-2443.352026-09-17T23:57:03+00:00 – 2026-09-24T23:57:03+00:00
2026-09-2338.172026-09-16T23:57:03+00:00 – 2026-09-23T23:57:03+00:00
2026-09-2235.382026-09-15T23:57:03+00:00 – 2026-09-22T23:57:03+00:00
2026-09-2132.872026-09-14T23:57:03+00:00 – 2026-09-21T23:57:03+00:00
2026-09-2037.352026-09-13T23:57:03+00:00 – 2026-09-20T23:57:03+00:00
2026-09-1939.972026-09-12T23:57:03+00:00 – 2026-09-19T23:57:03+00:00
2026-09-1843.462026-09-11T23:57:02+00:00 – 2026-09-18T23:57:02+00:00
2026-09-1743.462026-09-10T23:57:03+00:00 – 2026-09-17T23:57:03+00:00
2026-09-1644.972026-09-09T23:57:03+00:00 – 2026-09-16T23:57:03+00:00
2026-09-1544.592026-09-08T23:57:03+00:00 – 2026-09-15T23:57:03+00:00
2026-09-1441.992026-09-07T23:57:03+00:00 – 2026-09-14T23:57:03+00:00
2026-09-1341.992026-09-06T23:57:03+00:00 – 2026-09-13T23:57:03+00:00
2026-09-1239.472026-09-05T23:57:03+00:00 – 2026-09-12T23:57:03+00:00
2026-09-1139.472026-09-04T23:00:04+00:00 – 2026-09-11T23:00:04+00:00
2026-09-1036.882026-09-03T23:00:05+00:00 – 2026-09-10T23:00:05+00:00
2026-09-0948.2102026-09-02T23:00:04+00:00 – 2026-09-09T23:00:04+00:00
2026-09-0855.582026-09-01T23:00:03+00:00 – 2026-09-08T23:00:03+00:00
2026-09-0762.2102026-08-31T23:00:03+00:00 – 2026-09-07T23:00:03+00:00
2026-09-0662.2102026-08-30T23:00:02+00:00 – 2026-09-06T23:00:02+00:00
2026-09-0562.2102026-08-29T23:00:03+00:00 – 2026-09-05T23:00:03+00:00
2026-09-0463.8112026-08-28T23:00:03+00:00 – 2026-09-04T23:00:03+00:00

Hacker News · 53.8 · 7 scored comments · Read platform reviews →

ONE MODEL, DIFFERENT COMMUNITIES

Across the communities

Last 7 days

Scored separately for each platform, with different samples and audiences. Platform scores are not simply averaged; scored counts exclude quota-only comments.

Scores range from 0 to 100. Select a platform for its trend and reviews. Latest opinion time is not crawler health.

Community reviews

Selected comments from the last 7 days. The balance of excerpts does not represent the share of positive reviews.

2 selected excerpts

Community reports about limits cannot establish official allowances.

NegativeUsage limits

“cutting my usage in half”

NegativeCoding

“GPT 5.6 Sol on medium effort hallucinated not one but two correctness issues with the code”

Explore long-term changes & analysis

Model lifecycle

GPT-5.6 Sol

Released 2026-07-10 · Back to GPT

Exact-version signals observed 2026-09-23–2026-10-02 · n=168

Data window: 2026-09-05 00:00:00 to 2026-10-03 00:00:00 UTC · Scoring method: experience_score_v3.

Insufficient historical baseline The original baseline window lacks enough current-method evidence for a long-term comparison.

Lifecycle updated daily · latest complete UTC day 2026-10-02

The current classification method has insufficient evidence in the original baseline window. The window is preserved; no long-term improvement or decline is inferred, and old-method scores are not compared.

Fixed 14-day baseline
–
n=0
Latest 28 days
52.1
n=142
Change vs baseline
–
90-day slope
–
pts / 30 days

14-day rolling experience

Each daily point summarizes the previous 14 complete UTC days. The dashed line is the fixed release baseline; the verdict still uses independent non-overlapping 14-day segments.

14-day rolling experience4050602026-09-16–2026-09-30 · 51.3 · n=1262026-09-17–2026-10-01 · 51.0 · n=1312026-09-18–2026-10-02 · 51.6 · n=1382026-09-19–2026-10-03 · 52.1 · n=14208-1408-2309-0109-1009-1909-2810-03

0 of 4 independent 14-day segments ready for trend judgment.

What explains the change

Experience change and discussion-mix change are separated. These are observational contributions, not proof of cause.

No category has enough samples in both windows for a contribution conclusion.

Category Baseline → current Category change Weight share Experience contribution Discussion-mix contribution
Coding Baseline → currentnot enough datan=0 → 12 Category change– Weight share– Experience contribution– Discussion-mix contribution–
Image generation Baseline → currentnot enough datan=0 → 0 Category change– Weight share– Experience contribution– Discussion-mix contribution–
Image understanding Baseline → currentnot enough datan=0 → 0 Category change– Weight share– Experience contribution– Discussion-mix contribution–
Local deploy Baseline → currentnot enough datan=0 → 0 Category change– Weight share– Experience contribution– Discussion-mix contribution–
Reasoning Baseline → currentnot enough datan=0 → 34 Category change– Weight share– Experience contribution– Discussion-mix contribution–
Roleplay / creative Baseline → currentnot enough datan=0 → 3 Category change– Weight share– Experience contribution– Discussion-mix contribution–
Safety & refusals Baseline → currentnot enough datan=0 → 4 Category change– Weight share– Experience contribution– Discussion-mix contribution–
Speed & latency Baseline → currentnot enough datan=0 → 12 Category change– Weight share– Experience contribution– Discussion-mix contribution–
General text Baseline → currentnot enough datan=0 → 81 Category change– Weight share– Experience contribution– Discussion-mix contribution–
Video generation Baseline → currentnot enough datan=0 → 0 Category change– Weight share– Experience contribution– Discussion-mix contribution–

Lifecycle scores use the same experience-signal weights as the main index, without sample shrinkage after n=30. They measure public user perception, not model capability or backend causes.

How's your AI experience today?