Machine Learning Papers

Last 7 Days (September 04 – September 10, 2026)

← Previous Week

🏆 Top Papers This Week

#1 TOP PAPER (Score: 84)
Zeyang Li, Yunan Wang, Paolo Giaretta ... · Massachusetts Institute of Technology · arXiv
We develop Newton Matching, a unified framework for fine-tuning and sampling in generative modeling. The target is $π\proptoμe^{τr}$, where $r$ is the reward, $τ>0$ the inverse temperature, and $μ$ denotes the pretrained model's terminal density for fine-tuning or the constant $1...
#2 TOP PAPER (Score: 82)
Zili Wang, Zhaopeng Qiu, Yuekai Zhang ... · NVIDIA · MLSys 2025
Speculative decoding accelerates rollout generation, which dominates the cost of reinforcement learning (RL) post-training. Online co-training can further increase the draft's accuracy, yielding greater speedups. However, scaling this approach to co-training on large models with ...
#3 TOP PAPER (Score: 81)
Aashiq Muhamed, Virginia Smith · Carnegie Mellon University · arXiv
Model misalignment, prompt injection, or operator misuse could lead AI agents operating frontier-lab accounts to exfiltrate model weights, poison training data, or weaken release gates. Existing benchmarks do not test whether defenders can detect this activity among routine work ...