Meta logo

Research Engineer, Safety Evaluation

Meta

New York, NYPer Diem$219–301K/yrSeen 1w agoSeen in employer's feed 5 days ago

Most applications go out cold — see where you stand first. No sign-up to start.

Watch jobs like this.

At a glance

Compensation
$219–301K/yr
Location
New York, NY
Schedule
Per Diem
Work Authorization
Not specified

Olive lists jobs from US employers, including remote roles you can work from the United States.

Requirements

Credentials this posting asks for.

Bachelor's degree

Job overview

Meta is seeking a Research Engineer to join the Safety Evaluation team within Meta Superintelligence Labs, aiming to make the safety of frontier AI systems measurable by designing evaluations, building infrastructure, and setting technical direction for safety across multiple model families and modalities.

Skills & qualifications

RequiredNice to have

Skills

PythonPyTorchMachine LearningLarge Language ModelsMultimodal ModelsNLPDistributed SystemsData PipelinesEvaluation HarnessesStatisticsExperimental DesignCommunicationMentorship

Qualifications

Bachelor's Degree in Computer Science or Computer Engineering or Relevant Technical Field3+ Years Industry Research or Research-Engineering Experience in ML/AIPhD in Computer Science Machine Learning or Relevant Technical Field

Full job description

Summary:

Meta is seeking a Research Engineer to join the Safety Evaluation team within Meta Superintelligence Labs. Our mission is to make the safety of Meta's frontier AI systems measurable — turning ambiguous notions of "safe" into rigorous, defensible metrics that model developers, product teams, and company leadership rely on to make launch decisions.Safety evaluation is the ground truth for every safety claim Meta makes. This role owns that ground truth: designing the evaluations that detect emerging risks in text, image, voice, video, and agentic systems; building the infrastructure that runs them continuously against training checkpoints and production traffic; and setting the technical direction for how safety is measured across Meta's AI portfolio. You will define measurement standards that outlast any single model generation, and your results will directly gate what ships to billions of people.

Required Skills:

Research Engineer, Safety Evaluation Responsibilities:

  1. Set the technical strategy for safety evaluation across multiple model families and modalities, and drive it to execution across teams

  2. Design, implement, and validate novel evaluations for safety-critical behaviors — policy adherence, adversarial robustness, agentic risk, jailbreak resistance, and emerging harm categories — including for capabilities with no established benchmark

  3. Build and harden the distributed evaluation platform so that hundreds of evals run reliably and continuously against checkpoints throughout large-scale training runs

  4. Own the measurement quality bar: signal-to-noise, statistical power, saturation, contamination, and construct validity — and establish when an eval is trustworthy enough to gate a launch

  5. Create, curate, and analyze high-quality safety datasets, including adversarial, borderline, multilingual, and long-tail cases

  6. convert real-world incidents and red-team findings into durable, repeatable safety signals

  7. Diagnose anomalous eval results mid-training-run, determine whether the cause is a model change or an infrastructure artifact, and communicate a clear answer under time pressure

  8. Own the dashboards and reporting that researchers, product partners, and leadership use to monitor safety during training and post-launch

  9. Translate evolving global safety policy and regulatory standards into concrete, testable measurement criteria, partnering with Policy, Legal, and Integrity

  10. Influence the roadmaps of partner research and product teams

  11. mentor engineers and researchers and raise the evaluation bar across the org

  12. Represent Meta's safety evaluation methodology to internal leadership and, where appropriate, to external audiences and the research community

Minimum Qualifications:

Minimum Qualifications:

  1. Bachelor's degree in Computer Science, Computer Engineering, relevant technical field, or equivalent practical experience

  2. Bachelor's degree in Computer Science, Computer Engineering, a relevant technical field, or equivalent practical experience

  3. 3+ years of industry research or research-engineering experience in ML/AI, including hands-on work with LLMs, multimodal models, or NLP

  4. Demonstrated experience setting technical direction for a large, ambiguous problem area and driving it to delivery across multiple teams

  5. Experience designing and validating evaluations or benchmarks for ML systems, including reasoning about metric reliability and failure modes

  6. Experience building production-grade or research infrastructure that must be reliable at scale — distributed systems, data pipelines, or evaluation harnesses

  7. Programming experience in Python and hands-on experience with frameworks such as PyTorch

  8. Experience communicating complex technical results to non-specialist stakeholders and decision-makers

Preferred Qualifications:

Preferred Qualifications:

  1. Experience translating regulatory or policy requirements into technical measurement criteria

  2. Experience evaluating LLMs across multiple languages and modalities (text, image, voice, video, reasoning, tool use)

  3. Experience operating in an on-call or production-support capacity for live training runs or safety-critical systems

  4. Experience evaluating agentic systems — multi-step tool use, autonomy, and oversight mechanisms

  5. Publications at peer-reviewed venues (e.g. ICLR, NeurIPS, ICML, ACL, CVPR, ICCV, FAccT) with a track record in evaluation, alignment, or AI safety

  6. Experience with large-scale distributed training (hundreds/thousands of GPUs) and evaluating models in-flight during training

  7. Experience with adversarial evaluation and red-teaming, including automated attack generation and jailbreak robustness measurement

  8. experience with observability, monitoring, or experiment-tracking systems

  9. PhD in Computer Science, Machine Learning, or a relevant technical field

  10. Background in statistics and experimental design

Public Compensation:

$219,000/year to $301,000/year + bonus + equity + benefits

Industry: Internet

Equal Opportunity:

Meta is proud to be an Equal Employment Opportunity and Affirmative Action employer. We do not discriminate based upon race, religion, color, national origin, sex (including pregnancy, childbirth, or related medical conditions), sexual orientation, gender, gender identity, gender expression, transgender status, sexual stereotypes, age, status as a protected veteran, status as an individual with a disability, or other applicable legally protected characteristics. We also consider qualified applicants with criminal histories, consistent with applicable federal, state and local law. Meta participates in the E-Verify program in certain locations, as required by law. Please note that Meta may leverage artificial intelligence and machine learning technologies in connection with applications for employment.

Meta is committed to providing reasonable accommodations for candidates with disabilities in our recruiting process. If you need any assistance or accommodations due to a disability, please let us know at [email protected].

Similar jobs, posted recently

Open roles like this one, listed in the last 30 days.

You've read the whole posting — now see how you match it.