Head of Cyber Safety
$230,000–$280,000 year
Remote
Job Summary
Design and lead adversarial evaluations of frontier LLMs for offensive cyber capabilities, including vulnerability discovery, exploit development, malware generation, and autonomous cyber operations across text, agentic, and multimodal systems. Partner with machine learning engineers to build scalable benchmarks, classifiers, guardrails, and automated detection systems for internal products and frontier AI lab deployments. Develop and maintain the catastrophic cyber harm taxonomy while producing technical risk assessments and actionable recommendations for labs, enterprise customers, and internal stakeholders. Build and mentor a team of cybersecurity subject matter experts, establishing scalable evaluation processes and quality standards. Represent Gray Swan as the company's cybersecurity authority, collaborating with leading AI labs, security researchers, and government partners to reduce real-world security risks.
Required Qualifications
- Deep technical expertise in offensive cybersecurity, vulnerability research, exploit development, penetration testing, malware analysis, reverse engineering, or a closely related field
- Significant experience assessing advanced cyber threats, offensive tooling, or AI-enabled cyber capabilities, especially in critical infrastructure domains
- Hands-on experience conducting adversarial evaluations, AI red-teaming, LLM security research, or building evaluation datasets for frontier AI systems
- Comfortable operating at the intersection of cybersecurity research, AI safety, and machine learning engineering
- Thriving in highly ambiguous, fast-moving environments where you'll define strategy while building entirely new capabilities
- Being a builder who enjoys creating teams, infrastructure, and evaluation systems from scratch
- Experience developing machine learning models, AI security classifiers, or automated cyber detection systems
- Hands-on experience red-teaming frontier language models, jailbreaking, prompt injection research, or agentic AI evaluations
- Experience working with frontier AI labs, national security organizations, or leading cybersecurity research teams
- Background in threat intelligence, autonomous cyber operations, AI agent security, or AI governance
- Strong software engineering experience in Python, Go, Rust, or other systems programming languages
Desired Qualifications
- Experience developing machine learning models, AI security classifiers, or automated cyber detection systems
- Hands-on experience red-teaming frontier language models, jailbreaking, prompt injection research, or agentic AI evaluations
- Experience working with frontier AI labs, national security organizations, or leading cybersecurity research teams
- Background in threat intelligence, autonomous cyber operations, AI agent security, or AI governance
- Strong software engineering experience in Python, Go, Rust, or other systems programming languages
Hiring someone like this?
Get your role in front of qualified candidates on Sorce.