Hi, I’m Matteo Peluso. I think mostly about world modelling and curriculum, and the ideas behind them rather than any particular method. What does it take for a system to build a useful internal model of the world? What experiences, in what order?
I’m drawn to novel framings over incremental progress. This is where I share what I’m working on.
One step generation by efficiently pre-assigning gaussian latents to samples.
Beating AlphaZero policy networks with search-free self-play PPO, diverse resets and sparse terminal rewards.
Making a goal seeking policy with video pretraining only
A self-training autoencoder. Trained without reconstruction losses, pixel or otherwise, and no image-space supervision of any kind.
Simultaneous latent action and world model learning from passive video for controllable prediction.
Improving transformer memory efficiency by 2.7x
Using LLM Program Search to Build Robot Controllers Without RL Optimisers
This is a short fun blog in which I showcase an experiment in creating endless learnable content.
Exploratory post on curriculum as the missing ingredient required for true intelligence and creativity.
Microbiome World Modelling.
Cellular Automata Inspired Self Organising Text
Oniris: Autoregressive and Sample-Efficient Next-Gen Video Diffusion