Machine Learning Papers

Last 7 Days (August 21 – August 27, 2026)

← Previous Week

🏆 Top Papers This Week

#1 TOP PAPER (Score: 89)
Itay Safran · Weizmann Institute of Science · arXiv (likely intended for a major theoretical venue like NeurIPS/ICML/COLT given the citation of "Safran2026" and "Safran-Reichman-Valiant-2025", which suggests a very recent or forthcoming submission; however, based strictly on provided text, venue is unknown/arXiv)
We prove a depth hierarchy for ReLU neural networks in which every additional ReLU layer can save exponentially many neurons. For every $\ell\geq 3$, a globally $[0,1]$-valued, $1$-Lipschitz function is realized by a depth-$\ell$ network of width $\mathcal{O}(d^4)$, whereas every...
#2 TOP PAPER (Score: 88)
Qinglin Ye, Zhiyuan Gu, Jingjie Xia ... · University of Chinese Academy of Sciences +8 · arXiv
Search-augmented reasoning remains difficult for small language models. On-policy distillation (OPD) from trained teachers offers a promising direction, but suffers from two issues: (1) high-quality multi-turn search trajectories depend on dynamic retriever responses, making SFT ...
#3 TOP PAPER (Score: 88)
Jiaming Zhou, Qihang Zhang, Gangwei Xu ... · Ant Group (Robby Ant Research) · arXiv
Zero-shot cross-task generalization, where a policy must execute manipulation tasks never seen during training, remains a central challenge in robot learning. In large language models, a novel task can be performed simply by specifying it in the context, without any parameter upd...