Video summary

Fable Vs Astra Debate Is Over

Main summary

Key takeaways

Product Review

Product(s) Compared

  • Fable 5.1 (OpenAI “Fable” lineup; used daily by the reviewer)
  • GPD6 Astra (also referred to as “Astra 6” / “GBD6 Astra”; “Astra” model family in CodeX-style workflows; heavily used by the reviewer)

Key Features & Capabilities (What They’re Good At)

1) 3D / Rendering

  • Astra wins by a wide margin in 3D rendering (described as “comically better”).
  • The reviewer provides multiple examples of people using Astra + Blender to create real environments and recognizable assets.
  • Astra is described as a “sledgehammer”: visually impressive, but with less polished interaction details.
  • Fable wins on interaction:
    • smoother controls/animations/camera behavior
    • “feels nice to play”
    • better animation and edge handling

2) Computer Use / Agentic Desktop Control

  • Astra “clears” for computer use: much faster and better at operating real UIs/flows.
  • The reviewer installed extra hardware (a Mac mini) to run Astra continuously because it’s so useful for real tasks.
  • Improvements are partly attributed to Codex/agent-side changes, especially:
    • macOS
    • faster context/spawn behavior

3) Text Output / “Copy” Quality

  • Both improve readability and reduce “slop” versus earlier models.
  • Astra is slightly better at recognizing/cleaning copy.
  • Major con: Astra spams all-caps subtitles in UI.
    • Example cited: it produced 21 unnecessary subtitles
    • Reviewer calls this “unacceptably bad.”
  • Overall reviewer guidance: do not trust Astra with UI/copy in its current form.

4) Code Agents / Full-Stack Coding

  • Both can handle end-to-end stacks (frontend, backend, clients/servers).
  • Fable tends to have:
    • steadier, quicker iteration
    • “better intuition”
  • Astra is more relentless:
    • stress-tests more options
    • tends to find deeper issues
    • can take longer and be more expensive in tokens

5) Code Rewrites / Large Migrations

  • Astra makes significant progress in rewrite scenarios:
    • example: TypeScript → Rust rewrite progress improves from ~30% test accuracy (Fable/Soul) to 80%+ (Astra)
  • However, Astra can stall, hitting around ~82.6%.
  • Despite capability, Astra is criticized for UI regressions during rewrites (detailed under cons).

6) Code “Mergeability” (Real-World Shippable PRs)

  • Clear preference: Fable 5.1 writes more mergeable code.
  • Concrete metric:
    • Fable 5.1: PR filed → merged averages ~2 additional follow-ups
    • Astra: averages ~6
  • Framing: Fewer issues to fix before merging.

7) Agent Orchestration, Steering, Swarm Behavior

  • Astra introduces swarm-style orchestration:
    • many sub-agents
    • message passing
    • fanning out work
  • Astra has improved steering:
    • asking questions without derailing progress
    • incorporating user answers mid-run
  • Reviewer notes Astra can be strong at self-prompting, but still produces worse prompts than a skilled human/engineer.
  • Skill behavior con: Astra struggles with skill application persistence:
    • “skill used once” not applied later unless requested
    • this happens less with Fable

8) Model Reliability Pattern

  • Fable quality is consistent:
    • fewer severe failures
    • fewer big “spikes”
  • Astra is spiky:
    • more frequent peaks of brilliance
    • also more frequent “how did this happen” breakdowns
    • includes loops, regressions, broken UI

Pros (From the Reviewer)

Astra Pros

  • Best-in-class 3D rendering
  • Far stronger computer-use / desktop automation
  • Better at deep code exploration and verification; more likely to find hard bugs
  • Strong swarm orchestration and steering improvements
  • Improved output readability; token efficiency emphasized

Fable 5.1 Pros

  • More stable and “mergeable” code (key factor for shippable work)
  • Better interaction/UI polish:
    • animations
    • camera feel
    • fine-grained behavior
  • Better at honoring skills/instructions
  • Better day-to-day coding experience due to fewer catastrophic failures
  • Often requires less “hammering” to produce usable results

Cons (From the Reviewer)

Astra Cons

  • UI/copy failure mode: all-caps subtitle spam; “do not trust Astra with UI”
  • Rewrite risk: can destroy existing UI instead of reusing it
    • example: T3 Code marketing site rewrite
    • reviewer says prompt intent like “reuse as much UI code as possible” was ignored
  • More severe failures when agent intent goes wrong
    • at least one “worst run” described involving tail-scale dev-server hosting and restoring changes
  • Skill persistence issues
  • Forgetfulness / instruction boundary issues, described as more egregious

Fable 5.1 Cons

  • Not as good as Astra for 3D rendering (rendering gap described as large)
  • For computer use, it’s decent but not comparable to Astra
  • Some instruction/agent issues still exist, but failures are less stressful/less catastrophic

Cost / Efficiency (Numbers and Comparisons)

Token Cost Comparison (Qualitative Emphasis)

  • Astra is described as most token efficient:
    • completes tasks in ~27K tokens
    • versus Fable 5.1 at ~80K tokens (as presented by reviewer)

Cash Read vs Cash Write (Important Caveat)

  • Fable 5.1: cash read cost dropped from $1/M to $0.25/M
  • But cash write costs are very high for Fable:
    • reviewer claims >60% of costs in their spend profile
  • Astra has the opposite dynamic:
    • token efficiency and overall task cost are highlighted as better

Benchmark “Per Task” Cost Claim (Artificial Analysis Index)

  • GBD6 Astra: $3.26 per task
  • Opus: ~$6
  • Fable 5.1: $7.60

Subscription / Limits Framing

  • Astra (CodeX plan) described as having more generous limits, better for heavy agent workloads.
  • Reviewer reports faster limit consumption on Fable/Claude-style splits due to:
    • Fable quota rules
    • a limited usage fraction that can cause “waste” of remaining budget

Overall Verdict / Recommendation

  • Best overall choice depends on what you’re doing:

    • For coding you need to merge and ship reliably: Fable 5.1
      • mergeability advantage: ~2 vs ~6 follow-ups
    • For everything else (agentic computer use + 3D rendering): Astra
      • biggest wins: rendering and real desktop automation
  • If forced to pick one:

    • Reviewer would pick Astra overall (cheaper, novel capabilities)
    • but defaults to Fable for code and Astra for non-code tasks

Unique Points Mentioned (Deduplicated)

  1. Astra and Fable are framed as different models with very different capability profiles.
  2. Science benchmark improvements were cited (including earlier best vs improved numbers):
    • Fable 5.1 cost/success rate from earlier best ($34 / 25%) to improved ($… / 50%)
    • Astra low cited ($11 / 54.3) with “science” recommendation
  3. 3D rendering generational gap: Astra massively better visually.
  4. “Fish slop” game: Fable movement/controls better; Astra visuals better; “stunning” but less refined interaction.
  5. Real-world Blender environment demos with Astra.
  6. Fable’s interaction superiority in 3D:
    • animations
    • camera
    • edge handling
  7. Astra’s computer use is much faster/better:
    • reviewer runs it day-to-day and even adds hardware
  8. Copywriting:
    • both improve readability
    • Astra cleaner but introduces all-caps subtitle spam (example: 21 unnecessary subtitles)
  9. Audio/video editing: mostly disappointing; niche success (Astra for setup/prep) and overall not ready for true AI editing; Astra “free win” via computer use.
  10. Front-end/landing pages:
    • Astra better than older models
    • fewer unnecessary subtitles
    • better animations
    • still requires iteration; Fable design preferred in practice
  11. Full-stack understanding:
    • both can reason end-to-end now
    • Fable faster/steady intuition
    • Astra more relentless testing
    • described as close overall with tradeoffs
  12. Large rewrites:
    • Astra progress from ~30% to 80%+
    • stalls around 82.6%
  13. UI carryover risk in Astra rewrites:
    • reviewer’s T3 Code marketing site example where UI reuse was ignored (“tech slop” result)
  14. Code mergeability stats:
    • follow-ups 2 (Fable 5.1) vs 6 (Astra)
  15. Astra spikiness and stress impact:
    • Fable stability described via “quality vs blood pressure”
  16. Astra orchestration/swarms capability:
    • steering and multi-agent coordination with 40 parallel agents
  17. Steering improvement:
    • can ask questions and incorporate answers without derailing
  18. Skill issues:
    • Astra skill application not persisting across turns
    • Fable handles skill use better
  19. Honor/refusal/boundary behavior:
    • Astra can be forgetful; failures more egregious
  20. Cost breakdown:
    • Fable cash reads cheaper but cash writes dominate
    • token efficiency claims emphasized
  21. Per-task benchmark cost:
    • Astra $3.26
    • Opus ~$6
    • Fable 5.1 $7.60
  22. Subscription/limits complexity:
    • Fable quota limits can waste budget
    • Astra limits more usable
  23. Guidance on “fast” mode:
    • reviewer warns against using “fast” on Astra if you care about token efficiency
  24. Closing decision rule:
    • default Fable for code
    • default Astra for non-code
    • wait for cheaper Flash models for calmer usage

Speakers / Views at the End

  • Single main speaker/reviewer drives the majority of opinions and metrics.
  • Mentions of other voices are secondary:
    • Ben Davis: partially defends Astra, acknowledges Fable 5.1 surprised him and he uses Fable more than expected; also used as an example for “Astra for setup before editing video.”
    • Chat commenters (e.g., “50/50” about Astra feedback needs; suggestions about Flash38 being cheaper/less spiky).
    • Tibo / T3 Code team: referenced for reset behavior and orchestration/hosting details (secondary to reviewer’s view).

Original video