Skip to content

mrbizarro /Phosphene

Video, image and music models on your Mac, no cloud or API keys
Model toolingMITEnglish README
Stars
242
30 days
New
Last push
today

About Phosphene

Phosphene runs video, image and music models on your Mac, with no cloud or API keys. It's Apple Silicon only, built on MLX instead of PyTorch or CUDA. You get two video engines, stills, music and character LoRA training in one panel. Full character training needs 64 GB of memory; smaller Macs get cut-down recipes.

Best for: Mac users who want to render video with synced audio and train their own character LoRAs on their own Mac.

What it does

  • Renders video with synced audio (lip-sync, footsteps, ambience) in one pass
  • Switches between LTX-Video 2.5 and Hailuo H3 engines from the header
  • Trains character LoRAs from 15–500 images, with local Gemma 3 12B captions
  • Edits stills with Qwen-Image-Edit and up to three reference images
  • Copies motion, camera moves and poses from a clip with Motion Control
  • Queues renders and training over a local HTTP API on 127.0.0.1:8198

Quickstart

Rewritten from the README. Check the repo for the latest steps.

  1. 1
    Install Pinokio, then install Phosphene from it. The base install fetches LTX-Video 2.5 (about 27.5 GB).
  2. 2
    Add Hailuo H3 (about 75 GB) from the Pinokio sidebar if you want dialogue shots.
  3. 3
    Install the YuE2 music engine (about 11 GB) from Audio → Compose.
  4. 4
    Launch the panel and pick your engine with the segmented control in the header.
  5. 5
    Queue a render over HTTP: the command below
    curl -s -X POST http://127.0.0.1:8198/queue/add --data-urlencode "mode=t2v" --data-urlencode "prompt=..."

Models and languages

Models it works with

  • LTX-Video 2.5
  • MiniMax H3
  • Qwen-Image-Edit-2509
  • Gemma 3 12B
  • YuE2
  • mflux

Interface and docs

  • English

Alternatives

Other model tooling projects people compare with Phosphene.

james-see

ltx-video-mac

Mac app that renders AI video with sound using LTX and MiniMax

Model tooling
  • LTX-2
  • LTX-2.3
  • LTX-2.5
  • +4
423 GitHub stars23 days agoMIT

Trains and runs image, video and audio diffusion models on your own GPU

Model tooling
  • MiniMax H3
  • LTX-2.5
  • Wan-Animate-2
  • +7
13k GitHub starsyesterdayApache-2.0

ModelTC

LightX2V

Loads open image and video models onto your own GPU

Model tooling
  • MiniMax H3
  • Wan 2.2
  • Wan 2.1
  • +7
2.9k GitHub starstodayApache-2.0

Wan-Video

Wan2.2

Open video models for text, image, speech and character animation

Model tooling
  • Wan 2.2-T2V-A14B
  • Wan 2.2-I2V-A14B
  • Wan 2.2-TI2V-5B
  • +5
18k GitHub stars10 days agoApache-2.0

Featured on OpenMicroDrama

Do you maintain Phosphene? Add this badge to your README so English speakers can find our write-up.

Featured on OpenMicroDrama
markdown
[![Featured on OpenMicroDrama](https://openmicrodrama.com/badges/featured.svg)](https://openmicrodrama.com/projects/mrbizarro-phosphene)

The badge links to this page. Want your description changed? Email support@openmicrodrama.com.

Reviewed Oct 1, 2026. Bundles the LTX-Video MLX port and vanch007's YuE2 MLX port; not affiliated with Lightricks or MiniMax. We write these descriptions ourselves. The repo's own docs are the final word.