Control Red Team - Research Engineer/Research Scientist
This role involves designing and conducting ML experiments to evaluate the effectiveness of AI control measures such as monitors and sandboxes, with a focus on identifying adversarial attacks that could bypass them. The work combines empirical research, red-teaming of frontier AI systems, and building scalable infrastructure for automated evaluation. You'll produce rigorous safety assessments and contribute to high-impact publications and policy-relevant reporting.