Profile 👋
I’m a machine learning researcher working on reinforcement learning and language models.
I am graduating from ENPC and the MVA program at ENS Paris-Saclay, with interests spanning empirical and theoretical approaches to learning algorithms.
My recent research has focused on:
- RL / Formal reasoning (Leanstral)
- Overthinking in reasoning language models (Terminator at ICLR’26)
- Transformer architecture (FOG at NIPS’25)
- Pre-training (Apertus at ACL’26) and Distillation (Falcon) at scale
Recently, I joined Mistral AI’s science team to work on long-horizon reinforcement learning, supervised by Albert Jiang.
Between early 2025 and early 2026, I was a visiting student at MLO lab at EPFL, supervised by Prof. Martin Jaggi. I was also part of the Swiss AI Initiative core LLM team, an open-source initiative between EPFL, ETH Zurich, and CSCS.
Previously, I did an internship at the theory/frontier team of TII (UAE) and was a member of the Falcon LLM team.
🛎️ News 🛎️:
- [July 2026] Weights + tech report + SafeVerify + FLTEval are now public, following Leanstral 1.5 release.
- [March 2026] Our paper on tackling overthinking in RLMs is finally out.
- [Nov 2025] Excited to present our poster on FOG at NeurIPS in Paris on November 25th.
- [Sept 2025] Apertus family of fully open LLMs is released!
- [Dec 2024] Falcon3 family of open models is out: distillation FTW!
