People / Sumeet Motwani

Sumeet Motwani
PhD student, co-supervised
University of Oxford
Co-supervised with Prof. Philip Torr
Education
DPhil, University of Oxford, ongoing
Publications with the lab
Full publication list on Google Scholar.
- 2026LongCoT: Benchmarking Long-Horizon Chain-of-Thought ReasoningS. R. Motwani, D. Nichols, Charlie London, P. Li, and othersICML 2026Paper
- 2026AI Models Can Provably Hide Arbitrary CapabilitiesA. Draguns, S. R. Motwani, R. Douglas, C. Schroeder de WittPreprint
- 2026h1: Bootstrapping LLMs to Reason over Longer Horizons via Reinforcement LearningA. Ivanova*, S. R. Motwani*, Z. Cai, P. Torr, R. Islam, S. Shah, C. Schroeder de Witt†, Charlie London (* equal contribution, † joint supervision)ICML 2026Paper
- 2026Rubric Curriculum RL: Exploiting the Generation-Verification Gap in Non-Verifiable DomainsT. Krishnan*, S. R. Motwani*, C. London, S. M. Bhat, H. Jiao, P. Torr, R. Islam, C. Summerfield, C. Schroeder de Witt, Q. Gu, S. Shah (* equal contribution)ICML 2026
- 2025MALT: Improving Reasoning with Multi-Agent LLM TrainingS. R. Motwani, C. Smith, R. J. Das, R. Rafailov, I. Laptev, P. H. S. Torr, F. Pizzati, R. Clark, C. Schroeder de WittCOLM 2025Paper
- 2025REAL: Benchmarking Autonomous Agents on Deterministic Simulations of Real WebsitesD. Garg, D. Caples, A. Draguns, N. Ravi, P. Putta, N. Garg, P. Hebbar, and others, C. Schroeder de Witt, S. MotwaniNeurIPS 2025Paper
- 2024Secret Collusion among AI Agents: Multi-Agent Deception via SteganographyS. R. Motwani, M. Baranchuk, M. Strohmeier, V. Bolina, P. H. S. Torr, L. Hammond, C. Schroeder de WittNeurIPS 2024Paper
- 2024Unelicitable Backdoors in Language Models via Cryptographic Transformer CircuitsA. Draguns*, A. Gritsevskiy*, S. R. Motwani, C. Rogers-Smith, J. Ladish, C. Schroeder de Witt (* equal contribution)NeurIPS 2024Paper