My current research interests focus on:
- the improvement over obvious shortcomings in the RL recipes used to train LLMs.
- the search for the new paradigms that will supersede next-token prediction and verifiable RL.
I am also thinking about the interplay between AI and personal computing, and between personal computing and offline real life.

