DeepSeek 智能体结合视觉模型,耗时 1 小时迭代生成动画
PandaBearFred · reddit · 2026-08-13
A developer shared an experiment using a PI Agent where the DeepSeek model wrote HTML animations and used the Muse-Glimmer vision model for evaluation.
- Workflow: After DeepSeek generated code, it called Chrome to take a screenshot and sent it to the vision model for review, then iteratively modified the code based on feedback until the vision model was satisfied.
- Results: This visual feedback loop took about 30-60 minutes. Compared to the DeepSeek-only version, the version with vision feedback better matched the prompt details, though the standalone DeepSeek version spontaneously added a fade-in intro, showing different creative flair.
「编程与Agent」频道最新
- Obsidian 变身无头 CMS:VaultCMS 助你用 Markdown 驱动 Astro 网站 — tom_doerr · 2026-08-13
- AI Agent 预订酒店难在哪?开发者热议集成痛点 — RouteStack · 2026-08-13
- Ito 工具实测:每次 PR 自动运行应用以抓取运行时 Bug — Shruti_0810 · 2026-08-13
- CyberScraper 2077:用LLM智能抓取网页,支持OpenAI/Gemini/Ollama — tom_doerr · 2026-08-13
- 观点:90% 的“Agentic AI”只是加了推理层的 RPA — alex_verem · 2026-08-13
- 独立开发者实测 Fugu Ultra:163 次调用后的工程经验与反思 — Future-Cook-6365 · 2026-08-13