Machine Learning Papers

Last 7 Days (September 09 – September 15, 2026)

← Previous Week

🏆 Top Papers This Week

#1 TOP PAPER (Score: 86)
Ivan Moshkov, Stephen Ge, George Armstrong ... · NVIDIA · arXiv
We study how model post-training and test-time inference design affect natural-language proof generation for hard olympiad mathematics. Starting from Nemotron 3 Ultra, we train two specialist checkpoints using supervised fine-tuning and reinforcement learning, and evaluate checkp...
#2 TOP PAPER (Score: 81)
Jianman Lin, Shailesh Shailesh, Zhongyi Luo ... · National University of Singapore (MagicLab) +1 · arXiv
Robot foundation models achieve strong in-distribution performance but often degrade under visual distribution shifts. When learning to generate actions from pretrained visual representations, models may exploit task-irrelevant visual cues that correlate with demonstrated actions...
#3 TOP PAPER (Score: 80)
Junyao Yang, Yucheng Shi, Zhongzhi Li ... · Alibaba Group +1 · arXiv
Agent usage is shifting toward long-horizon tasks such as coding and scientific discovery, among which terminal tasks are especially important. We introduce T1, a Mixture-of-Experts model of 122B total trained with reinforcement learning, operating a real shell in a cloud sandbox...