Last released Aug 4, 2025
Inverse Reinforcement Learning for LLMs. Infer rewards and refine models.
Supported by