Representative projects:
Building a tool to continuously evaluate models and mitigate their risks. From designing the APIs for frontier labs, to building analysis and visualization tools that summarize 10,000+ transcripts into specific conclusions.
Designing and building challenges that measure a models ability to evade discovery, allowing us to see if models can operate on remote systems while avoiding detection by common defensive security tools.
Developing controlled environment frameworks for more secure use of frontier models.
Designing and building agents that improve a models ability to complete complex tasks. Includes many potential avenues, such as incorporating SOTA prompting practices, creating tools for task delegation, and more.
Publishing your research and/or delivering research to our customers.
You may be a good fit if you:
Have strong production programming skills and experience.
Have strong problem-solving and analytical skills.
Work well in a multidisciplinary team and can adapt to rapidly evolving challenges.
Are interested in AI and cybersecurity (experience in machine learning or cybersecurity is a plus but not necessary).
Care about the societal impacts of your work.








