Ariana Azarbal

Hello! These days, I'm a researcher through the Anthropic Fellows Program. I have broad interests in scalable oversight, AI psychology, and AI welfare. I recently worked on mitigations for reward hacking and misgeneralization, as a MATS fellow.

I also study Math-CS at Brown, where I help lead BAIST. Lately, I've been enjoying yoga and creative writing.

Recent Research

Writings