Microsoft Open-Sources OpenForgeRL for Training Complex Agents in Native Environments

microsoft · hf · 2026-07-25

Microsoft released OpenForgeRL, an open-source framework designed to tackle the challenge of end-to-end training for modern AI agents that rely on complex inference harnesses (like Claude Code and Codex).

Core Mechanism

Performance

Key Findings

The research highlights that harness choice significantly impacts learning difficulty. While RL notably improves agentic reliability (e.g., self-verification, tool coverage), critical abilities like error recovery remain weak.

Related event: Microsoft Open-Sources OpenForgeRL for End-to-End Agent Training(4 posts)→

Original post →

More from coding & agent

coding & agent channel →