Skip to Main Content

Hi, I'm Baimam Boukar 👋

I am a research and software engineer working on mechanistic interpretability and the science of evaluations for advanced language models. At Jinesis AI Lab, University of Toronto, I study when probes, sparse autoencoders, and causal interventions generalize across model families, layers, tasks, and data distributions. I recently completed my master's in Applied Machine Learning at Carnegie Mellon University.

Profile

Recent updates

Jul 9, 2026 Intellibra won Cameroon's Social Entrepreneurship Prize (1st Place)

See All

Research Focus

I study how language models represent, conceal, and act on information.

Mechanistic Interpretability

How models encode and use internal representations.

Deception & Situational Awareness

When models detect oversight, hide intent, or behave strategically.

White-box Control

Whether probes, steering, and patching provide reliable causal control.

Evaluation Science

How to measure emerging capabilities with trustworthy controls.


Selected Work

View All

Built from scratch with