Trafy
Infrastructure

NVIDIA Nemotron Achieves Benchmark-Leading Performance With LangChain Deep Agents Harness

NVIDIA Nemotron 3 Ultra is offering leading performance at lower cost than top closed models with the largest and most widely adopted AI agent orchestration platform. LangChain tuned its Deep Agents harness for NVIDIA Nemotron 3 Ultra, achieving the highest accuracy among open models, while completing more tasks at higher throughput and running at 10x […]

Adel El Hallak·Jul 8, 2026·4 min read·Original source ↗
NVIDIA Nemotron Achieves Benchmark-Leading Performance With LangChain Deep Agents Harness

NVIDIA Nemotron 3 Ultra is offering leading performance at lower cost than top closed models with the largest and most widely adopted AI agent orchestration platform.  LangChain tuned its Deep Agents harness for NVIDIA Nemotron 3 Ultra, achieving the highest accuracy among open models, while completing more tasks at higher throughput and running at 10x lower inference cost per run than leading closed models.  Measured against LangChain’s Deep Agents benchmark, Nemotron 3 Ultra also achieved business task parity with the highest-scoring closed models. No model retraining was required. Every gain came from engineering the environment around the model, not the model itself.  At a tenth of the cost, teams harnessing NVIDIA Nemotron 3 Ultra can run evaluations continuously, experiment faster and build specialized agents across more of their business.  LangChain’s agent engineering platform has more than 200 million monthly downloads. By tuning its Deep Agents harness specifically for NVIDIA Nemotron 3 Ultra, it allows for high-performing agents that complete more tasks, run faster and give enterprises a fully open stack they can customize, own and run anywhere. “The way to build better agents is to keep improving the system around the model,” said Harrison Chase, cofounder and CEO of LangChain. “Memory, tool use, evaluation and model behavior compound when teams can tune them together. Our work with NVIDIA shows that enterprises can get strong performance from an open stack while keeping control over the agent systems they are building.” Abridge, Amdocs and Box are embedding specialized agents directly into their platforms and global systems integrator EY is expanding its NVIDIA implementation capabilities around NVIDIA NemoClaw blueprints for LangChain Deep Agents, helping clients customize, evaluate and govern specialized agents across high-value workflows.  NVIDIA founder and CEO Jensen Huang recently sat down with Chase to discuss why the last six months have seen a leap in useful AI for enterprises.

Harness Engineering, Not Fine-Tuning LangChain’s team ran Nemotron 3 Ultra against its public Deep Agents benchmark suite, then analyzed the deep agent’s execution traces to find exactly where it lost points. Instead of retraining the model, the team tuned the harness around it — adjusting system prompts, tool descriptions and middleware. Every developer using LangChain Deep Agents with Nemotron 3 Ultra can put this to work today — the tuned profile is available directly through LangChain. An Open Stack Built to Own NVIDIA NemoClaw for LangChain Deep Agents is the open reference blueprint that packages this work for enterprises building their own specialized AI — systems of models, tools and runtime — tuned for their own workflows. It combines LangChain Deep Agents Code, tuned for Nemotron 3 Ultra, with the NVIDIA OpenShell secure runtime for executing agent actions safely. An open model, an open harness and an open secure runtime means enterprises own the full stack, end to end. They can customize it around the expertise that sets their business apart, keep improving it and run it anywhere — their own infrastructure, their own cloud, their own governance.  That distinction matters more as agents take on higher-stakes work. The shift from AI assistants that answer questions to agents that take action inside core systems changes what businesses get from their AI.  NemoClaw for LangChain Deep Agents and the tuned Nemotron 3 Ultra model profile are available now. Developers can pull the tuned Deep Agents harness directly from LangChain, or use the NemoClaw for LangChain Deep Agents blueprint as a starting point for building specialized agents from scratch.  How to Get Started LangChain developers can access Nemotron 3 Ultra on Baseten, Crusoe Cloud, DeepInfra, Fireworks, Nebius and Together AI  platforms, giving them a direct, hosted path to the tuned harness in production. EY can help enterprises start building their own specialized agents today, using this open software stack.   Learn more about NVIDIA NemoClaw for LangChain Deep Agents and NVIDIA Nemotron.  Stay up to date on agentic AI, NVIDIA Nemotron and more by subscribing to NVIDIA news, joining the community, and following NVIDIA AI on LinkedIn, Instagram, X and Facebook.   Explore self-paced video tutorials and livestreams.  

Categories: Deep Learning

Tags: Nemotron

Related News

Research

How Open Models Are Driving AI Research

Jul 6, 2026

AI

NVIDIA and Partners Build in America, for America

Jul 1, 2026

AI

NVIDIA BioNeMo Agent Toolkit Brings Accelerated AI to Life Sciences Researchers in Claude Science

Jun 30, 2026

Robotics

Into the Omniverse: Three Workflows for Improving Vision AI Agent Accuracy With Synthetic Data and Fine-Tuning

Jun 30, 2026

Related

Looking back on Microsoft’s FY26: From AI experimentation to Frontier Transformation - The Official Microsoft BlogInfrastructure

Looking back on Microsoft’s FY26: From AI experimentation to Frontier Transformation - The Official Microsoft Blog

Throughout this past fiscal year, customers across every industry and segment moved from AI experimentation to deploying AI for real-world business outcomes. They unlocked innovation and created new opportunities for growth. We saw the emergence of Frontier Firms as they moved beyond efficiency gains to focus on human ambition and embed AI at the core...

Microsoft AI · Jul 28, 2026
12 min
Rethinking security for the age of AI - The Official Microsoft BlogInfrastructure

Rethinking security for the age of AI - The Official Microsoft Blog

Why security needs a new cyber stack — Introducing Project Perception The physics of cybersecurity are changing. Autonomous systems can now reason, adapt and operate continuously. At the same time, the cost of offense is falling, while the volume, velocity and complexity of what must be secured continues to grow. Attackers can generate exploits faster,...

Microsoft AI · Jul 27, 2026
7 min
Powering America's Genesis Mission: Microsoft's commitment to scientific discovery - The Official Microsoft BlogInfrastructure

Powering America's Genesis Mission: Microsoft's commitment to scientific discovery - The Official Microsoft Blog

Today, we’re excited to share a long-term commitment to the Department of Energy’s (DOE) Genesis Mission, backed by a $60 million investment designed to accelerate AI for science and the breakthroughs it can deliver for the country. This deepened commitment includes Microsoft’s new Scientific Partnership Advancing Research & Knowledge coordination hub and program office, otherwise...

Microsoft AI · Jul 22, 2026
10 min