Agent Lightning v1.0: A 3,500-Line Lightweight Agentic RL Framework for Training Agents with Real Harnesses

·Microsoft Research··

At a glance Harnessed Agentic RL: Microsoft Research Asia introduces a training paradigm in which the same agent harness used in deployment participates directly in reinforcement learning, removing the need to reimplement the agent inside the training framework. Lightweight by design: Agent Lightning v1.0 delivers a complete agent RL control plane in roughly 3,500 lines of code. Native Kubernetes support: agents run as standard Kubernetes jobs on self-managed clusters, cloud Kubernetes, or local...

Read full article →

Related Articles

Study: Claude, ChatGPT Offer Different Shopping Prices Based on Wealth
sbulaev · Hacker News · 1h ago
OpenAI’s o1 correctly diagnosed 67% of ER patients vs. 50-55% by triage doctors
donsupreme · Hacker News · 5mo ago
Accelerating Gemma 4: faster inference with multi-token prediction drafters
amrrs · Hacker News · 5mo ago
A couple million lines of Haskell: Production engineering at Mercury
unignorant · Hacker News · 5mo ago
OpenAI just dropped 700 preprints of mathematical proofs and counterexamples
tootie · Hacker News · 17h ago