← All models

GPT · Last 7 days

GPT-6 Luna

0–100 · higher means more positive community experience, not benchmark performance.

Updated 2026-10-01 11:07 UTC · 2026-09-24 – 2026-10-01 UTC
Experience score35.1/100112 comments in 7 days

What people think

One user notes that on the DeepSWE v1.1 software engineering test, GPT-6 Luna scored 66.6%, near Opus 5/Fable 5 medium effort, with per-task cost ~93%/96% lower. One user asked GPT-6 Luna to use computer use to solve a problem, but it stubbornly refused despite repeated reminders

About the sample and scores

As of 2026-10-01 11:27:03 UTC, GPT-6 Luna in the GPT family has 112 explicitly attributed community comments in the last 7 days; its community score is 35.1/100. Available category scores include General text 37.4/100 (n=51); Coding 34.9/100 (n=14); Reasoning 36.6/100 (n=15). Scoring method and sources

Community reviews

Selected comments from the last 7 days. The balance of excerpts does not represent the share of positive reviews.

1 selected excerpts

Community reports about limits cannot establish official allowances.

NegativeReasoning

“It's slightly worse, but I think as a sub-agent it doesn't matter—just modify the main agent's prompt to have it break down complex tasks into finer pieces.”

Machine translated
Show original

稍差一点,但是我觉得当子智能体的话不耽误啊,改一下主智能体的提示词,让它把复杂任务拆细就是了。

Comment context
Comment being replied to ↗

I haven't used it much recently. Is 6 luna a regression compared to 5.6 luna? I previously set the sub-agent to force 6 luna max; do I need to change it back to 5.6 luna max? [crying laughing]

Machine translated
Show original

我最近没怎么用,6 luna跟5.6luna比难道是退步的么?我之前是设置了子代理强制用6luna max,需要改回5.6 luna max么[笑哭]

How's your AI experience today?