Records of the !mmortal Data Scientist

Notes on AI, math, and the long road. Live slow, die whenever.

Long-form pieces on machine learning interpretability, kernel methods, contrastive learning, and the geometry of neural network representations. Most posts ship with interactive visualisations you can play with in the browser. The posts run in series that build on each other; see the map for how they connect.

New here? Two doors, either one works cold: What a Finite Kernel Buys an MLP turns a neuron into something you can look at, and Attention Is Explainable Because It Is a Kernel explains why half a transformer is readable and the other half is not. Everything else grows out of those two.

Series, in reading order

Each series is a narrative: start at part 1.

Latest

Every post by date →

More posts