Forge Neo gets an Image2Prompt tab for turning images into prompts locally
adeliogentile · reddit · 2026-07-23
Image2Prompt is a new Forge Neo extension that adds an Image2Prompt tab for turning images into prompts inside the UI.
Workflow:
- Upload or paste an image
- Pick a vision-language model
- Generate a prompt in the chosen style
- Send it directly to txt2img or img2img
Supported models include Qwen2-VL 2B, Qwen2.5-VL 3B, Qwen2-VL 7B, and Microsoft Florence-2 base/large. The Qwen models are automatically downloaded from Hugging Face on first use, with the author listing approximate VRAM requirements from about 1–3 GB for Florence-2 to about 16 GB for Qwen2-VL 7B.
More from Multimodal
- Meta says SAM 3 and DINOv3 cut a lab imaging workflow from one month to 15 minutes — imjustnewatai · 2026-07-23
- A terse French reply says the identity-law backlash is intentional — IgorCarron · 2026-07-23
- Facetnoir-style image generations shared with a reusable sref code — OVolosin82152 · 2026-07-23
- Midjourney images show fashion portraits and a stylized Ferrari render — Salmaaboukarr · 2026-07-23
- WAN 2.2 user asks whether video motion can be limited to one masked region — TekeshiX · 2026-07-23
- User says image models ignore reference photos unless outfits are spelled out — Ok-Star-6755 · 2026-07-23