AI EvalsJobs
Hiring?
All roles / Gray Swan AI / Head of Cyber Safety

Head of Cyber Safety

Red Team & SafeguardsRemote4w ago
Source verified: this exact posting was present on Gray Swan AI’s careers feed on . View employer source →

What the posting asks for

Names
Python, Rust
Doctorate
Not mentioned

Read out of the employer's own description. Absence means the posting does not say, not that the answer is no.

The role

About Gray Swan

Gray Swan is on a mission to empower the world to use AI safely and securely. We evaluate AI models for the leading frontier labs along with building real-time threat detection and adaptive adversarial red teaming agents for teams deploying AI.

We're a team of approximately 50 people, well-funded, growing quickly. Our work directly influences how the world deploys AI agents and systems at scale..

Learn more about how we work.

The Role

Come build and lead Gray Swan's cyber safety capability from the ground up, serving as the technical authority on AI-enabled cyber risk across red-teaming, evaluation, benchmarking, defenses, and safety infrastructure development. You'll help define how frontier AI systems are evaluated for offensive cyber capabilities while partnering with leading AI labs to reduce real-world security risks.

This role sits at the intersection of offensive security, AI safety, and machine learning. You'll transform deep cybersecurity expertise into scalable evaluation methodologies, safety infrastructure, and automated defenses that help establish industry standards for frontier model security.

If you have deep expertise in offensive cybersecurity, vulnerability research, or adversarial AI security, experience evaluating frontier models, and are driven to reduce catastrophic cyber risks from increasingly capable AI systems, we'd love to hear from you.

What You’ll Do:

  • Design and lead adversarial evaluations of frontier LLMs for offensive cyber capabilities, including vulnerability discovery, exploit development, malware generation, privilege escalation, social engineering, persistence, and autonomous cyber operations across text, agentic, and multimodal systems.

  • Partner closely with machine learning engineers to translate cybersecurity expertise into scalable benchmarks, classifiers, guardrails, automated detection systems, and evaluation infrastructure for both internal products and frontier AI lab deployments.

  • Develop and maintain Gray Swan’s catastrophic cyber harm taxonomy, continuously evolving cyber evaluation frameworks as frontier model capabilities rapidly advance.

  • Produce technical risk assessments and actionable recommendations for frontier AI labs, enterprise customers, and internal stakeholders, helping guide responsible model deployment and security mitigations.

  • Build, mentor, and lead a world-class team of cybersecurity subject matter experts while establishing scalable evaluation processes, quality standards, and technical infrastructure.

  • Represent Gray Swan as the company's cybersecurity authority, collaborating with frontier AI labs, security researchers, government partners, and the broader AI safety and cybersecurity communities.

Who You Are:

  • Deep technical expertise in offensive cybersecurity, vulnerability research, exploit development, penetration testing, malware analysis, reverse engineering, or a closely related field through industry, research, or equivalent experience.

  • Significant experience assessing advanced cyber threats, offensive tooling, or AI-enabled cyber capabilities, especially in critical infrastructure domains.

  • Hands-on experience conducting adversarial evaluations, AI red-teaming, LLM security research, or building evaluation datasets for frontier AI systems.

  • Comfortable operating at the intersection of cybersecurity research, AI safety, and machine learning engineering.

  • Thrive in highly ambiguous, fast-moving environments where you'll define strategy while building entirely new capabilities.

  • A builder who enjoys creating teams, infrastructure, and evaluation systems from scratch.

Bonus Points If You Have:

  • Experience developing machine learning models, AI security classifiers, or automated cyber detection systems.

  • Hands-on experience red-teaming frontier language models, jailbreaking, prompt injection research, or agentic AI evaluations.

  • Experience working with frontier AI labs, national security organizations, or leading cybersecurity research teams.

  • Background in threat intelligence, autonomous cyber operations, AI agent security, or AI governance.

  • Strong software engineering experience in Python, Go, Rust, or other systems programming languages.

If you don’t have 100% of these, you should still seriously consider applying. We care more about what you can do than your credentials.

You’ll Thrive Here If You:

  • Want visibility across the frontier of AI security by evaluating multiple frontier models and working directly with the world's leading AI labs.

  • Are excited to build the infrastructure that makes AI cyber safety scalable.

  • Feel motivated to reduce catastrophic cyber risks posed by increasingly capable AI systems.

  • Enjoy turning offensive security research into practical defenses that improve the safety of frontier AI.

  • Have a vision for what world-class AI cyber safety should look like—and are excited to build it.

What We Offer:

We offer a competitive compensation package designed to reward impact and incentivize growth. Our compensation philosophy is informed by our current valuation and recent industry data.

Compensation: $230,000 - $280,000 plus performance based bonus and meaningful equity package

Benefits:

  • 401k with up to 4% matching

  • 28 days annual leave (vacation + holidays)

  • Health, dental, and vision coverage

  • Catered lunches (Pittsburgh office)

  • Flexible work arrangements

  • Visa sponsorship available for exceptional candidates

Interview Process

🔎 Application review. We read everything; we’ll respond within 10 days.

🗣 Recruiter Screen. We learn about you; you learn about us.

🧑‍💻 Technical interview or Hiring Manager Interview.

💻 Role related assesment

🗣 Experience & culture interview

😇 Reference checks. We’ll reach out to 3-5 references that you provide.

📃 Offer. If it’s mutual, we move fast.

How to Apply

Submit your resume, link to your portfolio, and answer the questions on the application.

Published by Gray Swan AI on their own careers page and reproduced here unedited. Read it at Gray Swan AI.

Apply at Gray Swan AI → Applications go directly to Gray Swan AI. This board does not sit in between, take a fee from you, or see your application.

What this listing does not tell you

Listed 27 days. Of the 23 red team and safeguards roles this board has watched from listing to removal, 22% were gone from their employer's careers page by day 27, and the median came down after 58 days. That is a description of other listings that have already ended, not a prediction about this one: this board records when a listing disappears, never why, and a posting still up is not on a clock it can see.

Gray Swan AI has 4 roles open on this board, 4 of them in red team and safeguards.

Get the weekly AI Evals Jobs briefNew roles and board updates, with published pay where available. This is the general weekly brief. Or browse them all now.

More roles like this

Matched by discipline, title, listed location and work arrangement.

Same discipline: Red Team & Safeguards · Both have management titles · Shared listed location · Both list remote work; check location eligibility

Same discipline: Red Team & Safeguards · Both have management titles · Both list remote work; check location eligibility

Same discipline: Red Team & Safeguards · Both have management titles · Both list remote work; check location eligibility

Same discipline: Red Team & Safeguards · Both have management titles

Same discipline: Red Team & Safeguards · Both have management titles

Privacy · Terms