gpt-6-luna real-task evaluations

Every report keeps its original task, method, finding, limitation, evaluator identity and verification date. A community report is evidence about that run, not a universal score.

  • GPT-6 Astra, Sol and Luna on one SVG animation prompt

    By 雪瑜 (@xueyu1125). Programmer and AI-tools builder on X who posted the timings and quota readings Published 2026-09-23; retrieved 2026-09-24; verified 2026-09-24.

    Task
    Generate an SVG animation of a pelican riding a bicycle, shown in H5
    Method
    Same prompt on GPT-6 Astra, GPT-6 Sol, GPT-6 Luna and GPT-5.6 Sol; the author recorded time taken and the share of a five-hour subscription quota each used.
    Finding
    GPT-6 Luna was the quickest and lightest of the four at 2 min 18 s and 1% of the five-hour quota, but the author dismissed its result as a flop (“拉完了”), ranking it alongside MiMo-V2.6-Flash.
    Limitation
    One prompt judged by the author without a score; the quota shares are a subscription meter, not API cost. X's automatic translation renders the verdict as its opposite, so read the original Chinese.