Connect the Dots: Training LLMs for Long-Lifecycle Agents with Cross-Domain Generalization Via Reinforcement Learning

arXiv:2606.20002 · cs.LG, cs.AI, cs.CL · Submitted 2026-06-18 · Read on arXiv

cs.LG, cs.AI, cs.CL

Submitted: 2026-06-18

Updated: 2026-09-20

Comments: arXiv v2 updates: scale up experiments to larger Qwen3.6/3.8 models and new domains; support multi-teacher on-policy distillation; add theoretical study of learning dynamics

Code: https://github.com/agentscope-ai/Trinity-RFT

Project page: https://lifelongagent.github.io/iclr26.html

License: http://arxiv.org/licenses/nonexclusive-distrib/1.0/

Terminology

Sources

Related papers