When we use language models to measure politics — or ask what politics they already contain — what are we actually measuring? This is hands-on interpretability and evaluation work. I study how LLMs represent political ideology internally, using sparse autoencoders and representation analysis, and I show where LLM-based measurement fails construct validity: models key on surface features rather than the construct they are meant to capture. The work is grounded in mechanistic interpretability, including graduate training in Neural Mechanics (with David Bau).

I also teach this work: most recently a guest lecture, “Anthropology of Machines: From Black Box to White Box”, at SICSS-Istanbul 2026.