Ali Jalal-Kamali, Ph.D.

Researcher in behavioral evaluation with focus on AI safety, alignment, and interpretability.

I work on exposing latent behaviors and making them measurable. Given a system (or a dataset) containing behavioral expressions, I explore what can be discovered about hidden behaviors, how early they can be detected, and their impact on the system’s performance. My work expands from geopolitical actors’ relations to human teams’ interactions to frontier language models’ responses. Each study provides a working pipeline along with the findings, so the evaluation systems can be used by others.

Ph.D. in Computer Science, University of Southern California.

Research projects