vibehacker
News
DAPO ·

ByteDance Seed and Tsinghua open-source DAPO for scalable LLM RL

ByteDance Seed and Tsinghua AIR open-sourced DAPO (Decoupled Clip and Dynamic Sampling Policy Optimization): algorithm, verl training stack, DAPO-Math-17k, and Qwen2.5-32B weights. Trained from scratch, DAPO-Qwen-32B hit 50% on AIME 2024 with about half the steps of DeepSeek-R1-Zero-Qwen-32B.

More news

View all

Amika: persistent VMs so coding agents keep state between PRs

Amika’s Sept 20 write up pushes “Rigs”—reusable cloud VMs that keep repos, services, caches, and agents (Claude Code, Codex, OpenCode) across related tasks instead of one disposable sandbox per PR. Rigs can stay persistent, clone, or go ephemeral; secret and network controls are still labeled coming soon…

RuntimeWire

Spotted something we missed? Start a thread.