Control Red Team - Research Engineer/Research Scientist
This role involves designing and running machine learning experiments to evaluate the effectiveness of AI control measures, such as monitors and sandboxes, by developing adversarial attacks and conducting security analyses. The work combines empirical research with real-world testing of frontier AI systems, using LLMs to automate evaluation loops and building scalable infrastructure for high-quality, reusable results. The position is embedded within a red team focused on identifying vulnerabilities in AI safety systems to inform both developers and government policy.