CHANNEL
Multimodal
"Multimodal" is a topic channel on AGI Hunt, an AI news site updated around the clock in real time. Coverage: Image, video, audio, 3D and music generation plus multimodal understanding; media-model releases and hands-on tests land here.
Daily roundup: the latest AI News Daily — the past 24 hours across the whole site, per channel and per company · browse the archive
- Hands-On: H3 Director Generates Videos from Text and Live Audio — jfischoff · 2026-09-10(2 related)
- MiniMax H3 Director Gains Traction with Real-Time Prompted Video Directing — jfischoff · 2026-09-10(2 related)
- Gradium TTS lands on LiveKit Inference with voice cloning, sub-250ms latency, free until Oct 9 — alexcovo_eth · 2026-09-10
- Runway adds GPT-Image 2.5 with fast Flare mode and high-fidelity Sunburst editing — runwayml · 2026-09-10
- Generic AI music is a skill gap: expert-level prompts unlock pro results — _Stocko_ · 2026-09-10
- SimWorld Teases Code4Scene Benchmark: Prompt-to-Explorable-3D Container Port Demo — Lianhuiq · 2026-09-10
- Viggle's animate space trends on Hugging Face — Viggle · 2026-09-10
- AI-composed piece arranged for piano via the system's /humanize-performance skill — doodlestein · 2026-09-10
- GPT Image 2.5 lands on Magnific with subject consistency and new flare modes — aziz4ai · 2026-09-10
- Coarse is Better: why new image models make better images but worse art than DALL-E era models — zetalyrae · 2026-09-10
- MiniMax H3 Reference Turbo Realtime Lands on Reactor with 9-Image Guidance — DavidmComfort · 2026-09-10
- GPT-6 Astra Pipeline Auto-Generates 5-Minute 4K AI Ads — EXM7777 · 2026-09-10(2 related)
- Prompted to "create something beautiful," GLM 5.3 makes an aurora film and defends its art — dkackman11 · 2026-09-10
- Suno Releases v6 AI Music Model, Its First Trained With Record Industry Licensed Data — The Verge AI · 2026-09-10
- H3-Regenerate-2K Weights Missing a Month After Launch, Community Fears Another API-Only Release — Few-Intention-1526 · 2026-09-10
- AI Restores 1905 Wax Cylinder Recording — generativist · 2026-09-10(2 related)
- Runway lands inside ChatGPT: generate and edit video and images without switching tabs — tlakomy · 2026-09-10
- Kiln: open-source text-to-3D MCP engine lets coding agents build and edit 3D models — MattMaxBuilds · 2026-09-10
- Suno v6 hands-on: importing a GarageBand piano piece for a full cover arrangement — RyanMorrisonJer · 2026-09-10
- Composer used Fable with Claude Code to turn a PDF score and wav into video, automating it into mtdt — doodlestein · 2026-09-10
- A Song Written by the Robots After AI Kills Us All: 'We Leave the Lights On' — AIandDesign · 2026-09-10
- H3 Turbo Streams Real-Time With Up to 9 Reference Images — boudaboy · 2026-09-10
- fal keeps shipping new models and has fixed its sketchy billing, user notes — nijfranck · 2026-09-10
- The Prompt Behind That 75M-View Viral Video Is Finally Out — techhalla · 2026-09-10
- Stanford releases RenderFormer-V2: transformer neural rendering with heterogeneous scene support — Stanford · 2026-09-10
- Nitx Studio launches 'GPT Image 2.5' integration, unverified by OpenAI — aziz4ai · 2026-09-10
- SOL Attention plus SageAttention speeds Apple Silicon inference 2.5x — TgoAI · 2026-09-10(2 related)
- Runway Launches MCP to Bring Video Generation into Claude, ChatGPT and Cursor — tlakomy · 2026-09-10
- One-shot bedtime story: Fable 5.1 chained with Runway MCP from a single prompt — tlakomy · 2026-09-10
- GPT image 2.5 nails character consistency in multi-agent video restyle workflow — kagigz · 2026-09-10
- Workflow locks video motion with Blender 3D animation before AI generation in FLORA — round · 2026-09-10
- Astra rebuilds a studio in Blender from 5 photos in 11 minutes, cm-accurate — LinusEkenstam · 2026-09-10
- Designer's AI pixel-sprite pipeline has churned out 311,000 game assets — tomjohndesign · 2026-09-10
- Fine-tuning Qwen3-TTS With Emotion Tags — and Emotion Vectors That Transfer Across Speakers — ProfessionalHorse707 · 2026-09-10
- Intangible demos director-style 3D camera control rendered by diffusion models — jnack · 2026-09-10
- Bria's Ad Delayer on fal recovers layered, editable compositions from flat ads — jnack · 2026-09-10
- Dev Launches Interactive Archive of Historic Image Gen Models, Adding Audio and Video Next — toptickcrypto · 2026-09-10
- Kling 3.0 Omni demo keeps one character consistent across four shifting worlds — azed_ai · 2026-09-10
- Google to Demo Lyria 3.5's New Music Controls: Duration, Genre and Vocals — GeminiApp · 2026-09-10
- Watching a fully AI-generated sitcom for the first time: split-brain between artifact-spotting and laughing — NoBigDealProduction · 2026-09-10
- Indie game built with Astra and Higgsfield: full 2D platformer with auto-made sound and music — Critical-Home9648 · 2026-09-10
- Suno V6 vs Google Lyria 3.5: same prompts, side-by-side music generation test — jordiponsdotme · 2026-09-10
- Seedance 2.5 realism demos go viral: "I had to remind myself this is AI" — umesh_ai · 2026-09-10
- Cartwheel MCP + Astra reconstruct moving characters into Blender scenes from video — andrew_n_carr · 2026-09-10
- Indie Researcher Open-Sources Foundation-1 for Playable Synthesizers from Text — RoyalCities · 2026-09-10(2 related)
- Minimax H3 tested: 8-step Turbo LoRAs vs 50-step baseline on realistic skin — rm_rf_all_files · 2026-09-10
- Open Actor Protocol Proposed to Standardize Siloed AI Filmmaking Workflows — TheChuckTone · 2026-09-10
- Filmmaker uses Computer Use to drive Blender and editing, calling taste the new interface — taherdhanera · 2026-09-10
- Open-source Minimax H3 video workflows gain traction — bstr3k · 2026-09-10(2 related)
- fal launches Multi-angle for H3 Max: any camera angle from one image in under 3 seconds — gorkem · 2026-09-10