Gray Swan Ai logo

Red Team Engineer

Gray Swan Ai

RemoteRemoteJob$110–185K/yrPosted 2w agoVerified open 5 days ago

Most applications go out cold — see where you stand first. No sign-up to start.

At a glance

Compensation
$110–185K/yr
Location
RemoteRemote
Work Authorization
Visa required • Visa sponsorship

Job overview

Gray Swan Ai is hiring a Red Team Engineer. Gray Swan’s Red Team Engineers design realistic AI attack scenarios, conduct red‑team engagements for external clients, and produce clear, actionable reports, while building internal tools and collaborating with product and engineering to improve frontier AI safety.

Key focus areas include Design new challenges and adversarial scenarios for Arena competitions, Work red‑team engagements with external clients covering prompt injection, tool use, and agent exploitation, and Use Slack, Linear, the Arena design editor, and AI tools to deliver high‑quality results on time.

Successful candidates bring 1+ Years AI Red Teaming Experience, Degree Or Experience Equivalent In Computer Science Or Related Field, and Background In Traditional Security Research CTFs Or Penetration Testing. Important skills include Vulnerability Assessments, Technical Communication, Scenario Design, Collaboration, Problem Solving, and Slack. Preferred (not required): Coding For Automating Test Cases.

Skills & qualifications

RequiredNice to have

Skills

Vulnerability AssessmentsTechnical CommunicationScenario DesignCollaborationProblem SolvingSlackLinearArena Design EditorAI ToolsAI Red TeamingAdversarial MLPrompt InjectionTool UseAgent ExploitationMultimodal Attack SurfacesCoding for Automating Test CasesWriting and Scenario Design SkillsCollaborative Communication StyleStrong Intuition About AI FailuresOwnership and Deadline DeliveryPenetration TestingCodingAutomationWritingCollaborative CommunicationQuick PivotsOwnershipTraditional Security ResearchCTFs

Qualifications

1+ Years AI Red Teaming ExperienceDegree or Experience Equivalent in Computer Science or Related FieldBackground in Traditional Security Research CTFs or Penetration TestingExperience as Gray Swan Arena CompetitorExperience With AI Red Teaming Beyond LLM Prompt InjectionFamiliarity With Multimodal Attack SurfacesCoding Experience for Automating or Scaling Test Cases

Benefits

401(k) Match
Medical Insurance
Dental Insurance
Vision Insurance
Paid Time Off

Full job description

About Gray Swan Gray Swan is on a mission to empower the world to use AI safely and securely. We evaluate AI models for the leading frontier labs along with building real-time threat detection and adaptive adversarial red teaming agents for teams deploying AI.

We're a team of approximately 50 people, well-funded, growing quickly. Our work directly influences how the world deploys AI agents and systems at scale..

Learn more about how we work .

The Role Gray Swan Arena is where the world's top red-teamers test frontier AI models, learn about AI security, and win cash prizes. Leading labs use the resulting data to ensure that future LLMs are safe and aligned with human values.

Gray Swan’s Red Team Engineers are the hands-on practitioners behind both our popular Arena challenges and private red team engagements. You’ll design realistic attack scenarios, find vulnerabilities in frontier AI systems, and translate your findings into clear and actionable reports.

If you're a creative AI red teamer with thorough reporting skills, a strong technical background, and a knack for finding attack surfaces, we want to hear from you.

What You’ll Do:

  • Design new challenges and adversarial scenarios for ongoing and upcoming Arena competitions

  • Work red teaming engagements with external clients, covering areas like prompt injection, tool use, and agent exploitation

  • Use Slack, Linear, our Arena design editor, and AI tools to deliver red-teaming results on time and with high quality

  • Build internal tools and processes for the Arena (e.g., cheating detection) and support operational tasks like break verification and prize distribution.

  • Work alongside Product and Engineering teams to create realistic scenarios that highlight attack surfaces in frontier systems

  • Help shape the future of Gray Swan Arena with novel ideas and challenge types

Who You Are:

  • 1+ years of experience in AI red teaming, adversarial ML, or related security research

  • Organized and comfortable working independently across multiple concurrent engagements

  • Proven ability to convey technical ideas in accessible reports

  • Comfortable with quick pivots and changing priorities.

  • Experience as a Gray Swan Arena competitor is a plus.

  • Background in traditional security research, CTFs, or penetration testing.

Bonus Points If You Have:

  • Experience with AI red teaming beyond LLM prompt injection

  • Degree or experience equivalent in Computer Science or related field

  • Familiarity with multimodal attack surfaces

  • Coding experience for automating or scaling test cases

If you don’t have 100% of these, you should still seriously consider applying. We care more about what you can do than your credentials.

You’ll Thrive Here If You:

  • Have writing and scenario design skills with a collaborative communication style for effective cross-functional teamwork.

  • Have strong intuition about how AI systems fail in real-world deployments

  • Take ownership and deliver high-quality work under deadlines

  • Want to be on the cutting edge of AI safety and security, proactively defending against novel threat vectors

What We Offer: We offer a competitive compensation package designed to reward impact and incentivize growth. Our compensation philosophy is informed by our current valuation and recent industry data.

Compensation: $110,000 - $185,000 (depending on level & location) plus performance based bonus and meaningful equity package

Benefits:

  • 401k with up to 4% matching

  • 28 days annual leave (vacation + holidays)

  • Health, dental, and vision coverage

  • Catered lunches (Pittsburgh office)

  • Flexible work arrangements

  • Visa sponsorship available for exceptional candidates

Interview Process 🔎 Application review. We read everything; we’ll respond within 10 days.

✏️ Online technical screen (15 min). Complete a simple, job-relevant exercise.

🗣 Intro call (30 min). We learn about you; you learn about us.

🧑‍💻 Technical interview (90 min). Live coding with some tasks requiring AIand others not.

🗣 Experience & culture interview (60 min). Conversational exploration of the skills fit.

😇 Reference checks. We’ll reach out to 3-5 references that you provide.

📃 Offer. If it’s mutual, we move fast.

How to Apply Submit your resume, link to your portfolio, and answer the questions on the application.

You've read the whole posting — now see how you match it.