Skip to content

letmehappycode /auto-video-agent

Subtitle file in, illustrated slideshow video out, run by Claude Code
ComfyUI workflowNo licenseChinese README

About auto-video-agent

Give it an SRT subtitle file and Claude Code splits it into 2–6 second shots, writes an illustration prompt for each, draws them with Seedream and stitches the frames into a video. A local web page lets you browse and redo shots. It's built for narrated explainer clips, so the result is still pictures on a timeline, not moving AI video. There is no LICENSE file, so reuse rights are unclear.

Best for: Creators of narrated story or explainer videos who want each line illustrated without drawing by hand.

From the auto-video-agent README on GitHub. Tap an image to see it full size.

What it does

  • Splits an SRT file into shots of 2–6 seconds by meaning, with start and end times
  • Writes a structured illustration prompt per shot and checks it against ten rules
  • Draws every shot through Volcano Engine's Seedream text-to-image API
  • Stitches the images into a silent video that follows the subtitle timing
  • Includes a local web UI to browse shots, preview images and edit shot data
  • Handles long scripts by cutting them into chunks first, then merging the shots

Quickstart

Rewritten from the README. Check the repo for the latest steps.

  1. 1
    Clone the repo and copy .env.example to .env, then add your Volcano Engine Seedream key.
  2. 2
    Open the folder in Claude Code.
  3. 3
    Give Claude the path to your SRT file and the art style you want.
  4. 4
    Claude runs the agents and workflows, then scripts/generate_images.py and scripts/compose_video.py.
  5. 5
    Open the web UI in web/ to review shots and redo any image.

Models and languages

Models it works with

  • Seedream
  • Claude

Interface and docs

  • Chinese

Alternatives

Other comfyui workflow projects people compare with auto-video-agent.

SkyNotSilent

insightcut-jianying-image-video

Topic or script in, image video and editable CapCut draft out

Editing & audio
  • Agnes Image 2.1 Flash
  • Doubao TTS
  • Xiaomi MiMo TTS
  • +1
122 GitHub stars+1/30d6 days agoMIT

gnipbao

story-to-handdrawn-video

Agent skill that makes a hand-drawn 3:4 video from a story or images

Agent skills
  • OpenAI
2.1k GitHub stars+26/30d14 days agoMIT

harry0703

MoneyPrinterTurbo

Writes a narrated short video from a topic: script, voiceover, footage, subtitles

Video framework
  • MiniMax H3
  • Seedance
  • Wan
  • +7
129k GitHub stars+814/30d2 days agoMIT

NikoDemon80

ComfyUI-H3-Motion-Context

Longer MiniMax H3 sequences with motion and sound carried across cuts

ComfyUI workflow
  • MiniMax H3
1.2k GitHub stars+148/30d29 days agoGPL-3.0

Featured on OpenMicroDrama

Do you maintain auto-video-agent? Add this badge to your README so English speakers can find our write-up.

Featured on OpenMicroDrama
markdown
[![Featured on OpenMicroDrama](https://openmicrodrama.com/badges/featured.svg)](https://openmicrodrama.com/projects/letmehappycode-auto-video-agent)

The badge links to this page. Want your description changed? Email support@openmicrodrama.com.

Reviewed Oct 6, 2026. Approved from the discovery queue 2026-10-06. No LICENSE file. We write these descriptions ourselves. The repo's own docs are the final word.