A reinforcement learning bridge for LLM-based agent applications, making it simple and flexible to train and optimize agent behaviors.
5.6k upvotes
Visit GitHub →
On mobile tap Share for WeChat / RED (Xiaohongshu) / X; on desktop use Copy text and paste into the app.