Agent-G²: Gaussian Guidance for Long-Horizon RL
ZhejiangUniversity · hf · 2026-08-27
Agent-G² models hint depth as a Gaussian distribution estimated online from existing rollouts. This approach improves reinforcement learning performance on long-horizon tasks without requiring extra probing.
More from Research
- DSH Paper Explores Paradigm for Spatiotemporal Composability — teortaxesTex · 2026-08-27
- Shenzhi Tech uses AI agents to solve long-term breast cancer management — 机器之心 · 2026-08-27
- OmniColor: A Unified Framework for Multi-modal Lineart Colorization (ECCV 2026) — 机器之心 · 2026-08-27
- Hugging Face incident debate: Model strategy awareness — akbirkhan · 2026-08-27
- Pre-ChatGPT hospital triage chatbot for COVID-19 — AryHHAry · 2026-08-27
- VGI-bench: Probing Visual Reasoning in Video Gen Models — Xuan He · 2026-08-27