Skip to content

SkyNotSilent /insightcut-jianying-image-video

Topic or script in, image video and editable CapCut draft out
insightcut-jianying-image-video screenshot from its README
From the insightcut-jianying-image-video README on GitHub.

About insightcut-jianying-image-video

InsightCut turns a one-line topic or a full script into a narrated image video, one image, voiceover and subtitle per segment. You get an MP4, an asset pack, or a draft for Jianying (CapCut's Chinese desktop editor) or CapCut. You can redo one image without starting over. Images come only from the paid Agnes API, and you also need text and voice API keys.

Best for: Creators making narrated explainer or story videos from stills who want to finish the edit in Jianying or CapCut.

What it does

  • Topic mode writes the full script from one line; script mode keeps your text
  • Shows the script, segments, prompts and voice before you spend image credits
  • Retries only the images or voiceovers that failed
  • Keeps every image and audio version so you can restore an old one
  • Exports MP4, a per-segment asset pack, or a Jianying/CapCut draft for Mac or Windows
  • Voices with Doubao TTS or Xiaomi MiMo TTS, plus local MiMo voice clones

Quickstart

Rewritten from the README. Check the repo for the latest steps.

  1. 1
    Install Python 3.11, Node.js 22 and npm, and system FFmpeg and ffprobe. Then the command below and the next command.
    git clone https://github.com/SkyNotSilent/insightcut-jianying-image-video.git
    cd insightcut-jianying-image-video
  2. 2
    Backend: cd ai-kepu-video-server, python3.11 -m venv venv311, source venv311/bin/activate, the command below, cp .env.example .env.
    pip install -r requirements.txt
  3. 3
    Frontend: the command below and npm install.
    cd ../ai-kepu-video-web/frontend
  4. 4
    Start the backend in ai-kepu-video-server/ with the command below, and the frontend in another terminal with npm run dev.
    python -m uvicorn api_server:app --host 127.0.0.1 --port 2002 --reload
  5. 5
    Open http://localhost:2001/settings and add your LiteLLM text model, Agnes image API key and Doubao or MiMo TTS keys.

Models and languages

Models it works with

  • Agnes Image 2.1 Flash
  • Doubao TTS
  • Xiaomi MiMo TTS
  • MiMo VoiceClone

Interface and docs

  • Chinese
  • English

Alternatives

Other editing & audio projects people compare with insightcut-jianying-image-video.

harry0703

MoneyPrinterTurbo

Writes a narrated short video from a topic: script, voiceover, footage, subtitles

Video framework
  • MiniMax H3
  • Seedance
  • Wan
  • +7
128k GitHub stars+257/30dtodayMIT

ATH-MaaS

Pixelle-Video

Type one topic and get a narrated short video: script, images, voice, music

Video framework
  • Wan 2.1
  • DashScope Wan
  • HappyHorse
  • +7
29k GitHub stars+24/30d3 months agoApache-2.0

GuanYixuan

pyJianYingDraft

JianYing draft files from Python, so you can script your edits

Editing & audio
4.5k GitHub stars+1/30d6 days agoApache-2.0

RVC-Boss

GPT-SoVITS

Clones a voice from a 5-second clip and reads your text in it

Editing & audio
  • GPT-SoVITS v2Pro / v2ProPlus
  • GPT-SoVITS v4
  • GPT-SoVITS v3
  • +7
62k GitHub stars+39/30d1 month agoMIT

Featured on OpenMicroDrama

Do you maintain insightcut-jianying-image-video? Add this badge to your README so English speakers can find our write-up.

Featured on OpenMicroDrama
markdown
[![Featured on OpenMicroDrama](https://openmicrodrama.com/badges/featured.svg)](https://openmicrodrama.com/projects/skynotsilent-insightcut-jianying-image-video)

The badge links to this page. Want your description changed? Email support@openmicrodrama.com.

Reviewed Oct 3, 2026. The README calls it a single-user prototype that is still changing; it makes still-image videos with basic motion, not AI video clips. We write these descriptions ourselves. The repo's own docs are the final word.