Home
I’m a PhD student at the Max Planck Institute for Intelligent Systems, primarily advised by Prof. Moritz Hardt and also co-advised by Prof. Tim Rocktäschel. My research is broadly in the domain of LLM post-training and AI evals.
I’m currently interning at Mistral AI in their post-training team in Paris, working on improving agents in hard-to-verify domains.
Research interests:
- automated data curation and RSI (eg: hillclimbing PostTrainBench)
- building open-ended long-horizon evals
- parallel inference-time compute (multi-agent), counterfactual simulations, unsupervised env design
- pushing model's capabilities in hard exploration problems (where pass@k is ~0).
News
- [July'26] FutureSim received the Best Paper Award at the AI Forecasting Workshop at ICML'26!
- [July'26] Presenting OpenForecaster and FutureSim at ICML'26 in Seoul!
- [June'26] Joined Mistral AI as an intern in their post-training team. Based in Paris from June to September. Hmu if you would like to chat!
- [May'26] FutureSim is out! A benchmark that replays real-world events to evaluate how agents adapt their predictions as new information arrives over time with a horizon of 3 months. [Paper] [Website]
- [Dec'25] We released OpenForecaster! We scale open-ended reasoning to predict the future, training an 8B model that matches much larger proprietary models at judgemental forecasting. Models, code, and data are all open-sourced. [Paper] [Website]
Old
- [June'25] I will be attending the YC AI Startup School from June 15-21. If you are in SF and would like to chat anything about startups or AI research (evals, scalable oversight, forecasting, and more), do reach out to me!
- [Sept 2024] Started PhD in Tübingen, Germany!
- From Feb. to May 2024, I was in Toronto working with Prof. Gillian Hadfield at the Vector Institute exploring the scope of normative alignment in RL-based agents.
- [Feb 22nd] Our work on fair sequential decision making won the Outstanding Paper Award at AAAI'24!
- From 21st June'22, I will be in Vienna, Austria, attending SoCS'22 and IJCAI'22.
- In May 2022, I began my research internship at LAMSADE, Université Paris Dauphine - PSL under Dr. Jérôme Lang and Dominik Peters. I am working at the intersection of computational social choice and automated decision-making, focusing on long-term fairness in the paradigm of virtual democracy.
Collaborate?
If you are working on a project and feel I could be worth collaborating or a challenging problem in which I might be interested, you can reach out to me here
