NVIDIA and LangChain Tune Deep Agents: Nemotron 3 Ultra Nears Opus

NVIDIA and LangChain have partnered to tune the open-source Deep Agents harness. This collaboration highlights a deep integration in agent frameworks and engineering stacks, demonstrating the high cost-efficiency potential of open-source models through concrete benchmark data.

Key Details and Performance Comparison

According to LangChain's internal deep-agents benchmark, NVIDIA's Nemotron 3 Ultra achieved a composite score of 0.86 (86%). In comparison, Claude Opus 4.8 scored 0.87, meaning the two models are separated by just 1 percentage point in performance. However, on the cost side, Nemotron 3 Ultra's inference cost was only $4.48. Several users, such as @BraceSproul, emphasized that with Deep Agents support, the model's usage cost has been reduced to a fraction of the original, enabling faster AI workflows and stronger decision-making.

Reactions and Industry Impact

In his repost, @hwchase17 pointed out that AI progress is not dictated solely by the release cadence of cutting-edge closed-source models; open-source models and their accompanying systems are improving rapidly as well. NVIDIA's official account also confirmed that this tuning work is now public. This collaboration and the subsequent evaluation results visually demonstrate that open-source agent stacks can now approach the performance of top-tier closed-source models with significantly better cost-efficiency.

2026-07-08 ~ 2026-07-10 · 6 related posts

Full story(2 episodes)→