My current research interests focus on:
- the improvement of obvious shortcomings in the RL recipes used to train LLMs.
- the search for the new paradigms that will supersede next-token prediction and verifiable RL. More broadly
I’ve also spent a fair amount of time recently thinking about the intersection between AI and personal computing, and between personal computing and offline real life.

