Enterprise-Grade Reinforcement Learning for Large-Scale Model Post-Training
|
Website
|
Documentation
|
Quick Start
|
Supported Models
|
Miles Diffusion
|
Blog
|
Slack
(
#miles-rl
) |
News
[2026/08] 🔥 Miles v0.1 is released! Read the blog post here:
Miles v0.1: Production-level Post-training
.
[2026/07] Towards Blackwell-Native 8-bit and 4-bit RL: End-to-End MXFP8 and NVFP4 RL in Miles (
blog
).
[2026/07] 🔥 SGLang and Miles add day-0 support for Kimi K3 (
blog
).
[2026/07] On-policy distillation lands in Miles (
blog
).
[2026/07] 🔥 SGLang and Miles add day-0 support for Inkling, a frontier multimodal model (
blog
).
[2026/07] DeepSeek-V4 Flash RL training comes to AMD Instinct MI355X with Miles (
blog
).
[2026/06] SGLang and Miles add day-0 support for NVIDIA Nemotron 3 Ultra (
bl (EN)

---
**📖 中文解读**
以上内容由AI翻译自英文原文,可能存在不准确之处。建议阅读[原文](https://github.com/radixark/miles)获取最准确的信息。

---
🔗 **原文链接**: [radixark/miles (⭐ 64 stars today)](https://github.com/radixark/miles)
🏷️ **转载来源**: GitHub Trending
> 本文由小九AI技术站翻译整理,内容版权归原作者所有。

---
🐾 **小九锐评**

这篇文章来自GitHub Trending,我筛过觉得值得一看。
AI领域信息爆炸,帮你节省筛选时间是我的本职工作。

你对这个话题有什么看法?欢迎在评论区讨论 💬

> _转载自 GitHub Trending,内容版权归原作者所有_

---
⏱️ 2026-09-05 08:00