PhD Student — University of Alberta

Devon
Yanitski

I study minds and language models: psychometrics from language embeddings, the neural correlates of human memory, and LLM interpretability tools applied to human minds.

Department of Psychology, University of AlbertaComputational Memory Lab

Psychometrics from embeddings

Can the geometry of an embedding space stand in for human survey responses? Semantic Factor Analysis recovers latent psychological structure from item text alone.

What's latent in language models

Opening up language models — sparse-autoencoder feature steering, and what clamping a single interpretable feature reveals about the computations running inside.

loading the rat…

Qwen3.5-2B with a slider to steer the sparse-autoencoder feature corresponding to rats (#26631). A rat-themed take on Golden Gate Claude. This is a small LLM and may produce repetitive output.

sparse autoencodersfeature steeringQwen3.5-2BRats
GitHub →
Paperunpublished

A Computational Analog of OCD in LLMs Through Sparse-Autoencoder Feature Steering

Devon Yanitski

OCD is defined by intrusive, unwanted thoughts a person cannot dismiss on command and that grow stronger when suppressed. We use sparse-autoencoders (SAEs) to amplify features in LLMs that represent obsessive concepts. This produced the defining marks of clinical obsession: the model could not suppress the theme when instructed to, whereas the same theme from ordinary prompting stopped on command, and telling the model not to think about the theme raised the feature's activity further.

computational psychiatrySAE steeringOCDinterpretability

Are LLMs dual-process thinkers?

Do language models choose like utilitarians, and does reasoning shift their judgements the way it shifts ours? Measuring when LLMs replicate human behaviour.

Paperunpublished

Reasoning in LLMs Causes More Utilitarian Judgements

Devon Yanitski

Does the human dual-process finding — that reasoning increases utilitarian moral judgment — extend to LLMs? Across 122 open-source models run on a 40-item moral-dilemma battery, switching on reasoning raised utilitarian responding and lowered altruism and deontology: the human dual-process signature, reproduced in a system with no affect, kin, or evolved morality.

moral psychologyreasoning122 LLMsprocess dissociation
Paperunpublished

LLMs Validate The Cognitive Reflection Test

Devon Yanitski

122 open-source LLMs took an updated Cognitive Reflection Test (Meyer et al., 2024), each item answered twice — once on intuition, once with reasoning. Holding arithmetic ability fixed, engaging reasoning alone moves the score, supporting the claim that the CRT is separable from ability — a cleaner demonstration than human data, where CRT scores confound ability, disposition, and motivation.

cognitive reflection testreasoning122 LLMsMeyer 2024

Improving LLM-assisted research

I am building the scaffolding that lets AI agents carry out autonomous psychology research — starting with how they ground themselves in the literature.

Viewing the live brain

Real-time EEG tooling — Muse Visualizer turns the raw waveforms and oscillatory band activity from a consumer headband into highly customizable, artistic live representations of the brain.

Let’s talk!

Always happy to talk LLMs, EEG, psychometrics, or a strange side project.