Last released Nov 12, 2025
A reinforcement learning environment for sheep herding simulation with PPO, SAC, and TD3 algorithms
Supported by