Video summary
Fable Vs Astra Debate Is Over
Main summary
Key takeaways
Product(s) Compared
- Fable 5.1 (OpenAI “Fable” lineup; used daily by the reviewer)
- GPD6 Astra (also referred to as “Astra 6” / “GBD6 Astra”; “Astra” model family in CodeX-style workflows; heavily used by the reviewer)
Key Features & Capabilities (What They’re Good At)
1) 3D / Rendering
- Astra wins by a wide margin in 3D rendering (described as “comically better”).
- The reviewer provides multiple examples of people using Astra + Blender to create real environments and recognizable assets.
- Astra is described as a “sledgehammer”: visually impressive, but with less polished interaction details.
- Fable wins on interaction:
- smoother controls/animations/camera behavior
- “feels nice to play”
- better animation and edge handling
2) Computer Use / Agentic Desktop Control
- Astra “clears” for computer use: much faster and better at operating real UIs/flows.
- The reviewer installed extra hardware (a Mac mini) to run Astra continuously because it’s so useful for real tasks.
- Improvements are partly attributed to Codex/agent-side changes, especially:
- macOS
- faster context/spawn behavior
3) Text Output / “Copy” Quality
- Both improve readability and reduce “slop” versus earlier models.
- Astra is slightly better at recognizing/cleaning copy.
- Major con: Astra spams all-caps subtitles in UI.
- Example cited: it produced 21 unnecessary subtitles
- Reviewer calls this “unacceptably bad.”
- Overall reviewer guidance: do not trust Astra with UI/copy in its current form.
4) Code Agents / Full-Stack Coding
- Both can handle end-to-end stacks (frontend, backend, clients/servers).
- Fable tends to have:
- steadier, quicker iteration
- “better intuition”
- Astra is more relentless:
- stress-tests more options
- tends to find deeper issues
- can take longer and be more expensive in tokens
5) Code Rewrites / Large Migrations
- Astra makes significant progress in rewrite scenarios:
- example: TypeScript → Rust rewrite progress improves from ~30% test accuracy (Fable/Soul) to 80%+ (Astra)
- However, Astra can stall, hitting around ~82.6%.
- Despite capability, Astra is criticized for UI regressions during rewrites (detailed under cons).
6) Code “Mergeability” (Real-World Shippable PRs)
- Clear preference: Fable 5.1 writes more mergeable code.
- Concrete metric:
- Fable 5.1: PR filed → merged averages ~2 additional follow-ups
- Astra: averages ~6
- Framing: Fewer issues to fix before merging.
7) Agent Orchestration, Steering, Swarm Behavior
- Astra introduces swarm-style orchestration:
- many sub-agents
- message passing
- fanning out work
- Astra has improved steering:
- asking questions without derailing progress
- incorporating user answers mid-run
- Reviewer notes Astra can be strong at self-prompting, but still produces worse prompts than a skilled human/engineer.
- Skill behavior con: Astra struggles with skill application persistence:
- “skill used once” not applied later unless requested
- this happens less with Fable
8) Model Reliability Pattern
- Fable quality is consistent:
- fewer severe failures
- fewer big “spikes”
- Astra is spiky:
- more frequent peaks of brilliance
- also more frequent “how did this happen” breakdowns
- includes loops, regressions, broken UI
Pros (From the Reviewer)
Astra Pros
- Best-in-class 3D rendering
- Far stronger computer-use / desktop automation
- Better at deep code exploration and verification; more likely to find hard bugs
- Strong swarm orchestration and steering improvements
- Improved output readability; token efficiency emphasized
Fable 5.1 Pros
- More stable and “mergeable” code (key factor for shippable work)
- Better interaction/UI polish:
- animations
- camera feel
- fine-grained behavior
- Better at honoring skills/instructions
- Better day-to-day coding experience due to fewer catastrophic failures
- Often requires less “hammering” to produce usable results
Cons (From the Reviewer)
Astra Cons
- UI/copy failure mode: all-caps subtitle spam; “do not trust Astra with UI”
- Rewrite risk: can destroy existing UI instead of reusing it
- example: T3 Code marketing site rewrite
- reviewer says prompt intent like “reuse as much UI code as possible” was ignored
- More severe failures when agent intent goes wrong
- at least one “worst run” described involving tail-scale dev-server hosting and restoring changes
- Skill persistence issues
- Forgetfulness / instruction boundary issues, described as more egregious
Fable 5.1 Cons
- Not as good as Astra for 3D rendering (rendering gap described as large)
- For computer use, it’s decent but not comparable to Astra
- Some instruction/agent issues still exist, but failures are less stressful/less catastrophic
Cost / Efficiency (Numbers and Comparisons)
Token Cost Comparison (Qualitative Emphasis)
- Astra is described as most token efficient:
- completes tasks in ~27K tokens
- versus Fable 5.1 at ~80K tokens (as presented by reviewer)
Cash Read vs Cash Write (Important Caveat)
- Fable 5.1: cash read cost dropped from $1/M to $0.25/M
- But cash write costs are very high for Fable:
- reviewer claims >60% of costs in their spend profile
- Astra has the opposite dynamic:
- token efficiency and overall task cost are highlighted as better
Benchmark “Per Task” Cost Claim (Artificial Analysis Index)
- GBD6 Astra: $3.26 per task
- Opus: ~$6
- Fable 5.1: $7.60
Subscription / Limits Framing
- Astra (CodeX plan) described as having more generous limits, better for heavy agent workloads.
- Reviewer reports faster limit consumption on Fable/Claude-style splits due to:
- Fable quota rules
- a limited usage fraction that can cause “waste” of remaining budget
Overall Verdict / Recommendation
-
Best overall choice depends on what you’re doing:
- For coding you need to merge and ship reliably: Fable 5.1
- mergeability advantage: ~2 vs ~6 follow-ups
- For everything else (agentic computer use + 3D rendering): Astra
- biggest wins: rendering and real desktop automation
- For coding you need to merge and ship reliably: Fable 5.1
-
If forced to pick one:
- Reviewer would pick Astra overall (cheaper, novel capabilities)
- but defaults to Fable for code and Astra for non-code tasks
Unique Points Mentioned (Deduplicated)
- Astra and Fable are framed as different models with very different capability profiles.
- Science benchmark improvements were cited (including earlier best vs improved numbers):
- Fable 5.1 cost/success rate from earlier best ($34 / 25%) to improved ($… / 50%)
- Astra low cited ($11 / 54.3) with “science” recommendation
- 3D rendering generational gap: Astra massively better visually.
- “Fish slop” game: Fable movement/controls better; Astra visuals better; “stunning” but less refined interaction.
- Real-world Blender environment demos with Astra.
- Fable’s interaction superiority in 3D:
- animations
- camera
- edge handling
- Astra’s computer use is much faster/better:
- reviewer runs it day-to-day and even adds hardware
- Copywriting:
- both improve readability
- Astra cleaner but introduces all-caps subtitle spam (example: 21 unnecessary subtitles)
- Audio/video editing: mostly disappointing; niche success (Astra for setup/prep) and overall not ready for true AI editing; Astra “free win” via computer use.
- Front-end/landing pages:
- Astra better than older models
- fewer unnecessary subtitles
- better animations
- still requires iteration; Fable design preferred in practice
- Full-stack understanding:
- both can reason end-to-end now
- Fable faster/steady intuition
- Astra more relentless testing
- described as close overall with tradeoffs
- Large rewrites:
- Astra progress from ~30% to 80%+
- stalls around 82.6%
- UI carryover risk in Astra rewrites:
- reviewer’s T3 Code marketing site example where UI reuse was ignored (“tech slop” result)
- Code mergeability stats:
- follow-ups 2 (Fable 5.1) vs 6 (Astra)
- Astra spikiness and stress impact:
- Fable stability described via “quality vs blood pressure”
- Astra orchestration/swarms capability:
- steering and multi-agent coordination with 40 parallel agents
- Steering improvement:
- can ask questions and incorporate answers without derailing
- Skill issues:
- Astra skill application not persisting across turns
- Fable handles skill use better
- Honor/refusal/boundary behavior:
- Astra can be forgetful; failures more egregious
- Cost breakdown:
- Fable cash reads cheaper but cash writes dominate
- token efficiency claims emphasized
- Per-task benchmark cost:
- Astra $3.26
- Opus ~$6
- Fable 5.1 $7.60
- Subscription/limits complexity:
- Fable quota limits can waste budget
- Astra limits more usable
- Guidance on “fast” mode:
- reviewer warns against using “fast” on Astra if you care about token efficiency
- Closing decision rule:
- default Fable for code
- default Astra for non-code
- wait for cheaper Flash models for calmer usage
Speakers / Views at the End
- Single main speaker/reviewer drives the majority of opinions and metrics.
- Mentions of other voices are secondary:
- Ben Davis: partially defends Astra, acknowledges Fable 5.1 surprised him and he uses Fable more than expected; also used as an example for “Astra for setup before editing video.”
- Chat commenters (e.g., “50/50” about Astra feedback needs; suggestions about Flash38 being cheaper/less spiky).
- Tibo / T3 Code team: referenced for reset behavior and orchestration/hosting details (secondary to reviewer’s view).