← All concepts

agentic reinforcement learning

Training approach that lets LLM-based agents learn multi-step tool use and planning via reinforcement learning, often with dense or per-tool rewards.

Papers
26
Released code
4
First seen
Mar 2026
Latest
Sept 2026

18 papers in the last two months, against 7 in the two before.

Who is working on it

Zhejiang University 3Alibaba Group 2Tsinghua University 2

The papers

Most central to this idea first, not most recent.

Concepts are extracted from each paper and reused across the corpus, so this page grows on its own as the desk reads.