Papers
Research papers, followed by shorter technical notes written with the team at Base Labs and Baseten.
Papers & preprints
5
6
10
11
Steering semantic search with interpretable features from sparse autoencoders
14
Technical notes · Base Labs and Baseten Research
Apr 2026
Post-training frontier legal agents with Baseten Research
with Mudith Jayasekara, Matthew Blau, Aaron Ellis-Bloor, Niko Grupen and Gabe Pereyra
Mar 2026
Towards infinite context windows: neural KV cache compaction
with Alex Sandomirsky and Harry Partridge
Mar 2026
Dense, on-policy, or both?
with Max Kirkby
Feb 2026
Distillation without the dark
with Max Kirkby and Mudith Jayasekara
Jan 2026
Oct 2025
Oct 2025
Lumina: building self-improving evaluation through customer-in-the-loop refinement
with Harry Partridge, Max Kirkby, Jonathon Liu, Paras Stefanopoulos and Mudith Jayasekara
Oct 2025
Upweight the strategy, not the tokens: faster training with explicit reasoning through RGT
with Harry Partridge and Mudith Jayasekara
Oct 2025
Attention-based attribution: what your model is actually looking at
with Jonathon Liu, Kimbrian Canavan, Max Kirkby and Mudith Jayasekara
Oct 2025
Oct 2025
Robust, sample-efficient SFT with prompt mutations
with Harry Partridge
Oct 2025
Iterative SFT: dense reward learning
with Jonathon Liu, Harry Partridge, Max Kirkby and Mudith Jayasekara
Oct 2025
Write small, learn forever: rank-1 LoRA for continual learning
with Max Kirkby, Harry Partridge and Jonathon Liu
Sep 2025
Practical LoRA research
with Max Kirkby
Feb 2025
Do transformers notice their own mistakes? Finding a linear hallucination detector inside LLMs
with Mudith Jayasekara, Max Kirkby, Sviatoslav Chalnev and Rune Chi Zhao
Jan 2025
Resurrecting the salmon: seeing clearer inside LLMs with domain-specific SAEs
with Mudith Jayasekara and Max Kirkby
Jan 2025
Why mechanistic interpretability needs a paradigm inversion
with Mudith Jayasekara and Max Kirkby