On-policy reinforcement learning for multi-agent language systems
Project page
View project page and paper
PettingLLMs applies on-policy reinforcement learning to coordinate multiple language agents. Built on rLLM, it explores how RL training can improve multi-agent collaboration and communication.
GitHub repository
View source code
Was this page helpful?
⌘I
Assistant
Responses are generated using AI and may contain mistakes.