跳到正文
Together AI·· 2026-04-24AI 评分44

Together AI 发布 distribution-aware speculative decoding 框架,将 RL 训练 rollout 速度提升最高 50%

Accelerate RL rollouts by up to 50% with distribution-aware speculative decoding

AI 导读

Together AI 推出的 distribution-aware speculative decoding (DAS) 框架将 RL 后训练中的 rollout 阶段速度提升了最高 50%,且不影响模型输出质量。该框架包含自适应后缀树草稿器和长度感知调度策略,在 DeepSeek-R1-Distill-Qwen-7B 的数学推理任务中实现了超过 50% 的 rollout 时间缩减。

来源:Together AI · together.ai