Spira Maxima turns plain text into finished social videos at $0.09/sec
iamfakhrealam · x · 2026-10-07
Spira Lab launched Spira Maxima 1.0, billed as the most complete text-to-social-video model:
- Generates ready-to-post videos from plain text: presenter, trending B-roll, motion, captions, and music
- Post-trained on trending videos and performance data so outputs follow what's currently working on feeds
- Priced at $0.09/second; hit #1 on launch day
- Promises to open-source when the announcement post reaches 2M views or the account hits 5K followers
The author argues this automates the finishing work (captions, cuts, music) that video editors get paid for, inviting pushback from editors.
Related event: Spira Maxima 1.0: Text-to-Finished Social Video Model Released(4 posts)→
More from Multimodal
- Multi-Image Edit Arena launches: gpt-image-2.5 tops 44 models on 9.2M votes — arena · 2026-10-07
- Gemini 3 Pro image bills output at 60x input: one user's $253 lesson — Atm1n9 · 2026-10-07
- S2PD: serial computation in high-noise diffusion makes video models follow physics — elliottszwu · 2026-10-07
- Minimax ref-to-video output looks fake: thickened lines and smoothed textures — maxiedaniels · 2026-10-07
- fal launches Claude Connector and GPT plugin to generate media inside Claude and Codex — gorkem · 2026-10-07
- MiniMax H3 demo video "Bar Rescue - Ghost" shared on Reddit — Certain_Potato_4509 · 2026-10-07