Machine Learning Papers

Last 7 Days (August 20 – August 26, 2026)

← Previous Week

🏆 Top Papers This Week

#1 TOP PAPER (Score: 89)
Itay Safran · Weizmann Institute of Science · arXiv (likely intended for a major theoretical venue like NeurIPS/ICML/COLT given the citation of "Safran2026" and "Safran-Reichman-Valiant-2025", which suggests a very recent or forthcoming submission; however, based strictly on provided text, venue is unknown/arXiv)
We prove a depth hierarchy for ReLU neural networks in which every additional ReLU layer can save exponentially many neurons. For every $\ell\geq 3$, a globally $[0,1]$-valued, $1$-Lipschitz function is realized by a depth-$\ell$ network of width $\mathcal{O}(d^4)$, whereas every...
#2 TOP PAPER (Score: 88)
Qinglin Ye, Zhiyuan Gu, Jingjie Xia ... · University of Chinese Academy of Sciences +8 · arXiv
Search-augmented reasoning remains difficult for small language models. On-policy distillation (OPD) from trained teachers offers a promising direction, but suffers from two issues: (1) high-quality multi-turn search trajectories depend on dynamic retriever responses, making SFT ...
#3 TOP PAPER (Score: 85)
Yixin Tao, Weiqiang Zheng · Shanghai University of Finance and Economics +1 · arXiv (preprint)
We settle the minimax-optimal alternating regret, a regret notion motivated by alternating learning dynamics in games, for both online linear optimization (OLO) and online convex optimization (OCO). For OLO over the probability simplex $Δ_d$, we give an algorithm with $O(\log d...