Skip to content

mrbizarro /Phosphene

Video, image and music models on your Mac, no cloud or API keys
Model toolingMITEnglish README
Stars
243
30 days
+1
Last push
yesterday
Phosphene screenshot from its README
From the Phosphene README on GitHub.

About Phosphene

Phosphene runs video, image and music models on your Mac, with no cloud or API keys. It's Apple Silicon only, built on MLX instead of PyTorch or CUDA. You get two video engines, stills, music and character LoRA training in one panel. Full character training needs 64 GB of memory; smaller Macs get cut-down recipes.

Best for: Mac users who want to render video with synced audio and train their own character LoRAs on their own Mac.

What it does

  • Renders video with synced audio (lip-sync, footsteps, ambience) in one pass
  • Switches between LTX-Video 2.5 and Hailuo H3 engines from the header
  • Trains character LoRAs from 15–500 images, with local Gemma 3 12B captions
  • Edits stills with Qwen-Image-Edit and up to three reference images
  • Copies motion, camera moves and poses from a clip with Motion Control
  • Queues renders and training over a local HTTP API on 127.0.0.1:8198

Quickstart

Rewritten from the README. Check the repo for the latest steps.

  1. 1
    Install Pinokio, then install Phosphene from it. The base install fetches LTX-Video 2.5 (about 27.5 GB).
  2. 2
    Add Hailuo H3 (about 75 GB) from the Pinokio sidebar if you want dialogue shots.
  3. 3
    Install the YuE2 music engine (about 11 GB) from Audio → Compose.
  4. 4
    Launch the panel and pick your engine with the segmented control in the header.
  5. 5
    Queue a render over HTTP: the command below
    curl -s -X POST http://127.0.0.1:8198/queue/add --data-urlencode "mode=t2v" --data-urlencode "prompt=..."

Models and languages

Models it works with

  • LTX-Video 2.5
  • MiniMax H3
  • Qwen-Image-Edit-2509
  • Gemma 3 12B
  • YuE2
  • mflux

Interface and docs

  • English

Alternatives

Other model tooling projects people compare with Phosphene.

james-see

ltx-video-mac

Mac app that renders AI video with sound using LTX and MiniMax

Model tooling
  • LTX-2
  • LTX-2.3
  • LTX-2.5
  • +4
424 GitHub stars+1/30d24 days agoMIT

modelscope

DiffSynth-Studio

Trains and runs image, video and audio diffusion models on your own GPU

Model tooling
  • MiniMax H3
  • LTX-2.5
  • Wan-Animate-2
  • +7
13k GitHub stars-1/30d2 days agoApache-2.0

ModelTC

LightX2V

Loads open image and video models onto your own GPU

Model tooling
  • MiniMax H3
  • Wan 2.2
  • Wan 2.1
  • +7
2.9k GitHub stars-2/30dyesterdayApache-2.0

deepbeepmeep

Wan2GP

Desktop app that runs open video, image and audio models on your own PC

Model tooling
  • Wan 2.1/2.2
  • MiniMax H3
  • LTX-2/2.3/2.5
  • +7
9.8k GitHub stars+32/30dyesterdayCustom license

Featured on OpenMicroDrama

Do you maintain Phosphene? Add this badge to your README so English speakers can find our write-up.

Featured on OpenMicroDrama
markdown
[![Featured on OpenMicroDrama](https://openmicrodrama.com/badges/featured.svg)](https://openmicrodrama.com/projects/mrbizarro-phosphene)

The badge links to this page. Want your description changed? Email support@openmicrodrama.com.

Reviewed Oct 1, 2026. Bundles the LTX-Video MLX port and vanch007's YuE2 MLX port; not affiliated with Lightricks or MiniMax. We write these descriptions ourselves. The repo's own docs are the final word.