Head of Safety
San Francisco, CAFull-timePosted todayStill listed today
Most applications go out cold — see where you stand first. No sign-up to start.
Watch jobs like this. New roles like this one near San Francisco, CA, by email.
Don't just apply. Show up ready.
Olive works from this exact posting.
At a glance
Olive lists jobs from US employers, including remote roles you can work from the United States.
Job overview
Cognition is building Devin, an autonomous software agent, and seeks a Head of Safety to establish its safety function. The role owns safety strategy, evaluations and red-teaming, training and agent-harness collaboration, deployment policy, and team building. The hire will work with researchers and engineers and represent the company’s safety work to customers, policymakers, and the research community.
Skills & qualifications
Skills
Qualifications
Full job description
We are an applied AI lab building end-to-end software agents. We're the makers of Devin, the first AI software engineer. Our team is extremely talent-dense. Among our founding team, we have world-class competitive programmers, former founders, and leaders from companies at the cutting edge of AI including Scale AI, Palantir, Cursor, Waymo, Tesla, Lunchclub, Modal, Google DeepMind, and Nuro. Building Devin is just the first step—our hardest challenges still lie ahead. If you’re excited to solve some of the world’s biggest problems and build AI that can reason on real-world tasks, apply to join us. Role Mission Devin is one of the most capable autonomous agents in production. It writes and runs code, uses tools, takes actions in real customer systems, and operates for hours without a human in the loop. That makes it one of the most important places in the industry to get safety right, and one of the few where safety work ships to millions of developers rather than staying in a paper. You will build Cognition's safety function from the ground up: the research agenda, the evaluation and red-teaming program, the deployment policies, and the team. You will work directly with the researchers training our models and the engineers building the agent harness. This is a hands-on role for someone who has done serious safety or alignment work at a frontier lab and wants to apply it where agents actually operate.
What You'll Accomplish
-
Own safety end to end: Set the safety strategy for Devin and the models behind it, and own the outcomes.
-
Build the evaluation program: Design evals and red-teaming for agentic risk: unsafe actions, prompt injection, data exfiltration, sandbox escape, reward hacking, and misuse. Make them part of every model and product release.
-
Shape training and the agent harness: Partner with pre-training, post-training, and agent teams so safety is built into how models are trained and how Devin plans, acts, and asks for help.
-
Set deployment policy: Define what Devin is and is not allowed to do, how permissions and oversight work, and how we handle the gray areas, in a way that holds up with enterprise customers.
-
Build the team and the external voice: Hire and lead safety researchers and engineers. Represent Cognition's safety work with customers, policymakers, and the broader research community.
Exceptional Candidates Have Demonstrated
-
Frontier lab safety or alignment experience: Hands-on work on safety, alignment, evaluations, or red-teaming at a frontier AI lab. You have shipped safety work that affected real models or products.
-
Agentic systems depth: You understand how agents fail in practice: tool misuse, specification gaming, long-horizon drift, adversarial inputs. You have ideas about how to measure and mitigate it.
-
Research credibility: A track record of published or widely used safety research, evals, or methods. An advanced degree in Computer Science, Machine Learning, or a related field is a plus.
-
Technical fluency: Comfortable in Python and in the training and inference stack. You can read the code, run the evals, and argue with researchers on the details.
-
Builder, not reviewer: You have started or scaled a safety function and prefer shipping mitigations to writing memos about them.
-
Judgment under uncertainty: You can make clear calls on deployment risk with incomplete information and explain them to engineers, executives, and customers.
Equal Opportunity Cognition is an equal opportunity employer. We do not discriminate on the basis of race, color, religion, sex, sexual orientation, gender identity, national origin, age, disability, veteran status, or any other protected characteristic under applicable law. We are committed to providing reasonable accommodations for candidates with disabilities throughout the hiring process - please let us know if you need any.
Similar jobs, posted recently
Open roles like this one, listed in the last 30 days.
Research Engineer, Safety & AlignmentCognition · San Francisco, CAPosted todayPosted today
Director, Trust & SafetySuno · San Francisco, CA (Hybrid) · $195–256K/yrPosted 4w agoPosted 4w agoSafety Case EngineerAtoms · San Francisco, CA · $149–188K/yrPosted 3w agoPosted 3w ago
Safety and Occupational Health Specialist (Dive Safety Officer)Geological Survey · Salt Lake City, UT · $106–158K/yrPosted 2 days agoPosted 2 days ago
2027 Internship Safety Engineer, Agentic Safety Case AssessmentBedrock Robotics · San Francisco, CAPosted 1w agoPosted 1w ago
You've read the whole posting — now see how you match it.