gpt-6-luna real-task evaluations
Every report keeps its original task, method, finding, limitation, evaluator identity and verification date. A community report is evidence about that run, not a universal score.
GPT-6 Astra, Sol and Luna on one SVG animation prompt
By 雪瑜 (@xueyu1125). Programmer and AI-tools builder on X who posted the timings and quota readings Published 2026-09-23; retrieved 2026-09-24; verified 2026-09-24.
- Task
- Generate an SVG animation of a pelican riding a bicycle, shown in H5
- Method
- Same prompt on GPT-6 Astra, GPT-6 Sol, GPT-6 Luna and GPT-5.6 Sol; the author recorded time taken and the share of a five-hour subscription quota each used.
- Finding
- GPT-6 Luna was the quickest and lightest of the four at 2 min 18 s and 1% of the five-hour quota, but the author dismissed its result as a flop (“拉完了”), ranking it alongside MiMo-V2.6-Flash.
- Limitation
- One prompt judged by the author without a score; the quota shares are a subscription meter, not API cost. X's automatic translation renders the verdict as its opposite, so read the original Chinese.