LlamaGym is a GitHub library that helps you improve language model agents by letting them learn from experience. It applies online reinforcement learning, so an agent can be fine-tuned as it interacts with an environment.
It lowers the effort needed to experiment with this style of training. It is free and open source, best suited to machine learning researchers, students and developers curious about training agents rather than just prompting them.