Machine Learning Papers

Last 7 Days (October 02 – October 08, 2026)

← Previous Week

🏆 Top Papers This Week

#1 TOP PAPER (Score: 85)
Jakub Macina, Manu Kapur, Mrinmaya Sachan · ETH Zurich (Inferred from "eth-lre" in GitHub URL and Swiss AI Initiative funding) +2 · ICLR 2027 (Inferred from "iclr2027_conference" in text)
Large language models (LLMs) trained to answer questions are natively poor at teaching. Reinforcement Learning (RL) against a simulated student is a promising approach to improve their pedagogy, but existing RL-trained tutors reward the student's success on the tutored problem wi...
#2 TOP PAPER (Score: 83)
Shikhar Srivastava, Christopher Kanan · University at Buffalo (inferred from Empire AI Consortium and NSF grants) +6 · ICLR 2027 (Preprint/Under Review)
As data propagates through a Transformer, the norm of its hidden states grows by orders of magnitude with depth, a phenomenon framed as 'curse of depth' and nearly universally treated as a pathology to be suppressed. We take the opposite view. Across 16 pre-trained LLMs from 9 fa...
#3 TOP PAPER (Score: 83)
Songyuan Zhang, Oswin So, Eric Yang Yu ... · Massachusetts Institute of Technology · NeurIPS 2026
While offline reinforcement learning (RL) enables policy optimization from static datasets without costly online interaction, it remains bottlenecked by the risk of executing out-of-distribution (OOD) actions. Recent approaches mitigate this by learning a behavior-cloning policy ...