Recent Videos

Generalization in AI Agents: Lessons from Linear-Quadratic Control

Nadav Cohen
May 12, 2026

Lightning Talks

Lightning Talks
May 11, 2026

Provably Efficient Learning in Nonlinear Dynamical Systems 

Elad Hazan
May 11, 2026

Return of the Reward Function

Drew Bagnell
May 11, 2026

Understanding the foundation model pipeline through coverage

Dylan Foster
May 11, 2026

Learning to Answer from Correct Demonstrations

Nathan Srebro
May 11, 2026

Fisher Random Walk: Automatic Preference Inference for Language Models

Junwei Lu
April 24, 2026

PPO Fine-Tuning of Diffusion Models: Provable Convergence across Interpolated Trajectories

Yingbin Liang
April 24, 2026

Stochastic Zeroth-Order Policy Optimization for RLHF

Lei Ying
April 24, 2026