Open-source local pipeline generates full radio dramas from one click
fflluuxxuuss · reddit · 2026-09-20
An open-source ComfyUI node pack produces a complete radio drama from one queue press: a local LLM writes the script, voice-casting locks each character to one neural voice, a music model scores it, image/video models render scenes, and a finished .mp4 lands in your folder — no account, API key or paid service needed.
- Reusable plumbing: unified TTS adapters (Kokoro, Bark, Chatterbox, Dia, IndexTTS2 cloning, plus cloud options), video lanes (LTX 2.5 with joint audio-video foley, Wan 2.2, AnimateDiff, Blender), images (Z-Image-Turbo, Flux.1-dev, SD 1.5), a per-episode JSON ledger, and -14 LUFS mastering.
- Hardware: default graph runs on 8 GB GPUs and Apple Silicon; heavy video lanes want 16 GB. Verified on laptops, a Mac mini and a rented pod.
- License: MIT, commercial use explicitly welcome (underlying models like Flux dev and MusicGen carry their own non-commercial terms).
More from coding & agent
- AI designs a 4-layer PCB with 135 parts and 514 pads, human barely touched it — Paimaamu · 2026-09-20
- "AIs that won't use developers will be replaced by AIs that will," quips X user — ashishllm · 2026-09-20
- Agent product NoSpoon going private this month, creator says consumer AI market too early — Kyrannio · 2026-09-20
- Da7em Bench: independent AI benchmark scores models on 200 real client tasks across 12 areas — airesearch12 · 2026-09-20
- Using Codex + GPT-6 Astra to plan art installations: rebuild the wall in Blender, skip recalculation — perilli · 2026-09-20
- What does an AI engineer's day actually look like? A learner asks Reddit — tech_kie · 2026-09-20