README.md
August 10, 2026 · View on GitHub
Enterprise-Grade Reinforcement Learning for Large-Scale Model Training
High-Performance Rollout • Low Precision Training • Production Stability
Quick Start | Advanced Features | Developer Guide | Documentation
Miles is a high-performance, enterprise-ready reinforcement learning (RL) framework specifically optimized for Large-Scale model Post-Training. Miles bridges the gap between research-grade RL and production-grade reliability by integrating SGLang for high-throughput rollout and Megatron-LM for scalable training.
"A journey of a thousand miles begins with a single rollout."