Video summary
Claude Code (Free Plan) + YouTube = $77,000/Month
Main summary
Key takeaways
High-level concept of the video
The video claims you can create “stick man” / doodle-style YouTube videos (static images stitched together) using a fully free-ish, end-to-end AI workflow, covering:
- ideation → script → voiceover
- timestamp-aligned scene prompts → image generation
- editing sync → viral metadata + thumbnails → upload
The core argument is that these videos succeed due to rhythm, not graphics or animation.
Key analysis / product-feature claims (technological ideas)
1. Main differentiation: “Voice-over first, scenes second.”
The creator argues most workflows generate scenes/images first and then create the voiceover, which causes a “collision” that breaks viewer rhythm.
Correct workflow:
- Generate voiceover first
- Build scenes around natural pauses in the audio
- Each scene is born from a pause
- Editing becomes timestamp-driven rather than guessed
2. Single prompt to Claude (free plan) for ideation + scripts
- Uses the Claude AI free plan with one prompt (as stated in a pinned comment).
- Claude returns 5 viral topic ideas.
- After selecting one topic, Claude generates a long-form script (claimed to be ~5–15 minutes).
- Claude also outputs a downloadable text file containing the script.
3. Free text-to-speech via 11Labs free plan
- Uploads the script to 11Labs using the free plan.
- Selects a voice named “Raunak”.
- Generates the full voiceover audio, which can be downloaded.
4. Timestamp extraction from the voiceover using faziscribe.ai
- Uploads the voiceover to faziscribe.ai.
- Uses an “accuracy” mode to produce:
- transcripts with exact timestamps for pauses
- Claims it can auto-detect languages (e.g., Hindi/English/Urdu).
- Produces a timestamp script, which is fed back into Claude.
5. Claude converts timestamped lines into image prompts (scene-by-scene)
- Claude receives the timestamp script.
- Generates detailed text-to-image prompts for every timestamped line.
- Prompts are produced in batches, continued by typing “next.”
- Claude can assemble everything into a downloadable prompt file (one prompt per scene/timestamp).
6. Bulk image generation using a Chrome extension for Google Flow
- Uses flow.google with “agent mode” turned off.
- Configures image generation:
- aspect ratio: 16:9
- output per prompt: 1
- image model: “Nano Banana 2”
- Uses a Zappy Flow Chrome extension to generate in bulk.
- Prompts are supplied by:
- pasting into the extension (prompts separated by one line break), or
- uploading the downloaded prompt text file
- Mentions important toggles:
- developer name should be zappywala.ai (to avoid fake extensions)
- “include serial number in file name” should be toggled off
- random delay is configurable
- auto-save settings adjusted (first option turned off)
- Generation runs in the background; you can switch tabs/minimize, but must not close the tab/browser/extension.
7. Editing sync approach based on filename timestamps
- In the editor, place the voiceover on the timeline.
- Import scene images and arrange them using their timestamps embedded in filenames.
- Method:
- move playhead to the next scene start time
- trim the scene image clip to match
- iterate for each timestamp to preserve rhythm
8. Viral metadata and thumbnail generation using Claude
- Claude generates YouTube metadata:
- title, description, tags
- For thumbnails:
- requests 5 high CTR thumbnail prompt ideas
- requires prompts in a copyable code block
- Then uses the same Zappy extension workflow to generate 5 thumbnails with one click.
- User selects the best thumbnail and publishes.
Practical workflow checklist (tutorial-style steps)
-
Claude ideation + script
- Open Claude AI free plan
- Paste the single “master prompt”
- Get 5 viral topics
- Select one → generate script (downloadable text)
-
Voiceover (11Labs)
- Use 11Labs free plan
- Select voice “Raunak”
- Generate and download audio
-
Transcription + pause timestamps (faziscribe.ai)
- Upload voiceover to faziscribe.ai
- Transcribe with exact pause timestamps
- Copy/download the timestamp script
-
Image prompts (Claude)
- Paste timestamp script into Claude
- Generate text-to-image prompts per timestamp
- Collect/download prompt file (batch with “next”)
-
Bulk image generation (Chrome extension + flow.google)
- Use Zappy Flow Chrome extension + flow.google
- Settings:
- agent mode off
- image model: Nano Banana 2
- aspect ratio: 16:9
- Bulk-generate all scene images automatically
-
Edit and sync
- Place voiceover on timeline
- Add scene images in order
- Trim each image to its timestamp (from filename)
- Preview and export
-
Metadata + thumbnails
- Use Claude-generated metadata (title/description/tags)
- Use Claude thumbnail prompt ideas → generate 5 thumbnails
- Pick one → upload
Notable “supporting evidence” used
- The video opens with an example “stick man” channel:
- claims include:
- 7.5 million views on one video
- 14 videos
- 137,000 subscribers
- 14.5 million total views
- first upload 2 months ago
- claims include:
- It emphasizes that success comes from rhythm created by voiceover pauses rather than visual complexity.
Main speakers / sources
- Speaker: the video creator (narrator) presenting and claiming they personally built the workflow.
- Tools/Services referenced:
- Claude AI (Anthropic)
- 11Labs (text-to-speech)
- faziscribe.ai (transcription with pause timestamps)
- flow.google (image generation workflow)
- Zappy Flow Chrome extension (automation for Google Flow)
- YouTube (metadata + thumbnail + upload workflow)
- Named voice: Raunak (used in 11Labs voice selection).