Research Scientist
Conducts original research in AI safety, focusing on interpretability and behavioral analysis of transformer-based language models. Designs and runs experiments to study model behavior, develops reproducible research tooling in Python/PyTorch, and translates technical findings for government, industry, and safety partners. Works closely with senior scientists to advance robust AI deployment and contribute to emerging safety standards.