Home

I’m a PhD student at the Max Planck Institute for Intelligent Systems, primarily advised by Prof. Moritz Hardt and also co-advised by Prof. Tim Rocktäschel. My research is broadly in the domain of LLM post-training and AI evals.

I’m currently interning at Mistral AI in their post-training team in Paris, working on improving agents in hard-to-verify domains.

Research interests:

  • automated data curation and RSI (eg: hillclimbing PostTrainBench)
  • building open-ended long-horizon evals
  • parallel inference-time compute (multi-agent), counterfactual simulations, unsupervised env design
  • pushing model's capabilities in hard exploration problems (where pass@k is ~0).

News

  • [July'26] FutureSim received the Best Paper Award at the AI Forecasting Workshop at ICML'26!
  • [July'26] Presenting OpenForecaster and FutureSim at ICML'26 in Seoul!
  • [June'26] Joined Mistral AI as an intern in their post-training team. Based in Paris from June to September. Hmu if you would like to chat!
  • [May'26] FutureSim is out! A benchmark that replays real-world events to evaluate how agents adapt their predictions as new information arrives over time with a horizon of 3 months. [Paper] [Website]
  • [Dec'25] We released OpenForecaster! We scale open-ended reasoning to predict the future, training an 8B model that matches much larger proprietary models at judgemental forecasting. Models, code, and data are all open-sourced. [Paper] [Website]
Old

Collaborate?

If you are working on a project and feel I could be worth collaborating or a challenging problem in which I might be interested, you can reach out to me here