arxiv
PublishedSeptember 29, 2026 at 4:00 AM
—neutral
Rufus-Air: An Open LLM Post-Training Recipe
Publisher summary· verbatim
arXiv:2609.29421v2 Announce Type: replace-cross Abstract: Rufus-Air is an open and reproducible post-training recipe on GLM-4.5-Air-Base (106B-A12B), organized as a serial pipeline of eight stages: SFT, Reasoning RL, Coding RL, Instruction-Following RL, General Agent, Coding Agent, Search Agent, and
Stay posted· Newsletter
A 5-min weekly brief — top movers, price watch, story of the week.
Discussion
No replies yet. Be first.
Related coverage
More from ARXIV
arxivReasoning Externalization for Faithful Large Language Model Narratives of Stock Return Predictions12harxivConsistent Plan-Act for Long-Horizon Agentic Tasks12harxivPredictive Self-Supervised Learning Provably Identifies Stochastic Signals under Nuisance12harxivBoosting Adversarial Robustness and Generalization with Dictionary Structure12hThe Bubble Brief
WEEKLYRead AI insights every Tuesday — top movers, new releases, story of the week.
Originally published on arxiv ↗