blogs

longer pieces i've written โ€” mostly things i became fascinated by and wanted to understand end to end. click read to open the full article.


residual streams in transformer models

jan 2025 ยท ~15 min read

what residual streams are, why they matter for interpretability, and how attention & FFN outputs accumulate into a shared communication channel across layers โ€” worked through with a concrete "capital of France" example.

read

a dummy's questions on the road to mech interp

2024 ยท ~10 min read

beginner questions i worked through while learning mechanistic interpretability โ€” activation vector spaces, polysemantic neurons, bigrams & skip-trigrams, QK/OV circuits, and the "no privileged basis" of the residual stream.

read

i'll be graduating soon

7 jan 2025 ยท ~2 min read

looking back at four years of college โ€” the robotics lab, learning ML alone when nobody else was, and the people who'll stay in my life long after. one last time.

read

hi there!

23 may 2024 ยท ~1 min read

a short hello and a note on what i plan to write about here โ€” thoughts on the parts of tech i'm currently fascinated by.

read

the platonic representation hypothesis

on twitter

a thread on the platonic representation hypothesis โ€” how models trained on different data and modalities seem to converge toward a shared statistical model of reality.

read on twitter