Reinforcement Learning · Michigan State University
Rethinking Muon Explained: Where Muon Collapses and Pion's High-Pass Fix
Muon's uniform whitening wins pretraining but collapses under RL with verifiable rewards and lags on robot policies; Pion keeps dominant directions, suppresses the noisy tail, and beats both Muon and AdamW.