Forge Neo gets an Image2Prompt tab for turning images into prompts locally
adeliogentile · reddit · 2026-07-23
Image2Prompt is a new Forge Neo extension that adds an Image2Prompt tab for turning images into prompts inside the UI.
Workflow:
- Upload or paste an image
- Pick a vision-language model
- Generate a prompt in the chosen style
- Send it directly to txt2img or img2img
Supported models include Qwen2-VL 2B, Qwen2.5-VL 3B, Qwen2-VL 7B, and Microsoft Florence-2 base/large. The Qwen models are automatically downloaded from Hugging Face on first use, with the author listing approximate VRAM requirements from about 1–3 GB for Florence-2 to about 16 GB for Qwen2-VL 7B.
More from Multimodal
- Midjourney style code share: --sref 2912175708 — tisch_eins · 2026-09-11
- Astra storyboards plus Minimax H3 per-shot generation boost video success rates — Hailuo_AI · 2026-09-11
- MiniMax H3 MAX nails cooking anime clips: 15-second curry demo with prompts shared — Hailuo_AI · 2026-09-11
- MiniMax Music Production Toolkit 2.5 for ComfyUI adds full mastering chain — Vivid_Promise1700 · 2026-09-11
- New Node Finder for ComfyUI ranks fresh nodes by star velocity and recency — Luke2642 · 2026-09-11
- Using a finisher move on one mosquito with MiniMax H3 MAX — the bug survives — Hailuo_AI · 2026-09-11