Publications
Reasoning agents Model architecture Interactive decision making Safety & privacy Information theory
2026
- The Price of Hidden Curvature: Improved Lower Bounds for Bandit Convex Optimizationpreprint
- Learning to Reason with Curriculum II: Compositional Generalizationpreprint
- Select and Improve: Understanding the Mechanics of Post-Training for Reasoningpreprint
- Learning to Reason with Curriculum I: Provable Benefits of AutocurriculumCOLT 2026
- Interactive Learning of Single-Index Models via Stochastic Gradient DescentICLR 2026
- From Markov to Laplace: How Mamba In-context Learns Markov chainsICLR 2026 (oral)
2025
- Scaling Test-Time Compute Without Verification or RL is SuboptimalICML 2025 (spotlight)
- SPECS: Faster Test-Time Scaling through Speculative DraftsNeurIPS 2027; ICML 2025 Workshop on Efficient Systems for Foundation Models (ES-FOMO III)
- Computational Intractability of Strategizing against Online LearnersCOLT 2025
- The Space Complexity of Learning Unlearning AlgorithmsCOLT 2025
2024
- An Analysis of Tokenization: Transformers under Markov DataNeurIPS 2024 (spotlight)
- Transformers on Markov data: Constant depth sufficesNeurips 2024; ICML 2024 Workshop on Mechanistic Interpretability 2024
2023
- Statistical complexity and optimal algorithms for non-linear ridge banditsAnnals of Statistics
- Greedy pruning with group lasso provably generalizes for matrix sensingNeurIPS 2023
2022
- Sample efficient deep reinforcement learning via local planningpreprint
-
- Semi-supervised active linear regressionNeurIPS 2022
-
2021
- On the value of interaction and function approximation in imitation learningNeurIPS 2021
- Not just age but age and quality of informationIEEE Journal on Selected Areas in Communications 39 (5), 1325–1338
- Provably breaking the quadratic error compounding barrier in imitation learning, optimallypreprint
2020
- Towards the fundamental limits of imitation learningNeurIPS 2020
- FastSecAgg: Scalable Secure Aggregation for Privacy-preserving Federated LearningICML Workshop on FL for User Privacy and Data Confidentiality (2020); CCS Workshop on Privacy-Preserving Machine Learning in Practice (2020); ISIT (2021)
- Missing Mass of Markov ChainsISIT 2020
2019
- Robust correlation clusteringAPPROX/RANDOM 2019
- Convergence of Chao unseen species estimatorISIT 2019