Spectrum Labs logo

GenAI Safety Analyst

Spectrum Labs

RemoteRemoteJob$63–90K/yrPosted 1mo agoVerified open 4 days ago

Most applications go out cold — see where you stand first. No sign-up to start.

At a glance

Compensation
$63–90K/yr
Location
RemoteRemote
Work Authorization
Not specified

Job overview

Spectrum Labs is hiring a GenAI Safety Analyst. Alice seeks a driven, detail‑focused professional to join its team as a Generative AI Analyst, diving into cutting‑edge technology to meticulously analyze content infringements and secure emerging Generative AI tools. The role involves collaborating with experts across hate speech, misinformation, intellectual property, and more, while developing adversarial prompts and overseeing high‑quality data outputs.

Key focus areas include Develop adversarial and risky prompt strategies across several areas of abuse to expose potential vulnerabilities in models, Manage projects end‑to‑end from planning through quality assurance to final delivery, and Handle extensive datasets across multiple languages and areas of abuse ensuring precision and meticulous attention to detail.

Successful candidates bring Near-Native English Command. Important skills include AI Safety, Responsible AI, Trust And Safety, Generative AI Models, AI Agents, and Attention To Detail. Preferred (not required): Text-to-Text Models, OSINT, and Self-Starter Attitude.

Skills & qualifications

RequiredNice to have

Skills

AI SafetyResponsible AITrust and SafetyGenerative AI ModelsAI AgentsAttention to DetailOrganizational CapabilitiesJuggling Numerous Tasks ConcurrentlyText-to-Text ModelsOSINTSelf-Starter Attitude

Qualifications

Near-Native English Command

Full job description

Description Alice is seeking a driven, detail-focused professional to become a vital part of our team as a Generative AI Analyst. In this role, you'll dive into the cutting-edge of technology, meticulously analyzing various content infringements to secure the new wave of Generative AI tools. Your duties will include collaborating with experts in diverse fields such as Hate Speech, Misinformation, Intellectual Property and Copyright, among others.

Your tasks will involve writing adversarial; prompts to identify weaknesses in various AI models, including Large Language Models (LLMs), Text-to-Image, Text-to-Video, AI Agents and beyond. You'll also oversee data management to guarantee the highest quality of outputs.

Responsibilities:

  • Developing adversarial and risky prompt strategies across several areas of abuse to expose potential vulnerabilities in models.

  • Managing projects end-to-end, from initial planning and oversight through quality assurance to final delivery.

  • Handling extensive datasets across multiple languages and areas of abuse, ensuring precision and meticulous attention to detail.

  • Ongoing investigation into new tactics for circumventing foundational models' safety measures.

  • Working alongside diverse teams, engineering, product, policy, to tackle new challenges and craft forward-thinking strategies and resolutions.

  • Promoting a culture of knowledge exchange and continual learning within the team. Requirements Must have:

  • Background in AI Safety and/or Responsible AI and/or Trust and Safety

  • Familiarity with recent Generative AI models and agents is essential, though direct technical experience is not a prerequisite.

  • Command of English at a near-native level.

  • Attention to detail, organizational capabilities, and the capacity to juggle numerous tasks concurrently. Additional Wants:

  • Experience with various model types (Text-to-Text, Text-to-Image) is desirable.

  • Prior experience with OSINT (Open Source Intelligence) will be considered an asset.

  • A self-starter attitude, with the energy to excel in a fast-moving and variable environment.

The salary range for this role is $63-90K OTE - Range may vary based on experience. Salary at the time of offer will be commensurate with experience.

About Alice Alice is a trust, safety, and security company built for the AI era. We safeguard the communicative technologies people use to create, collaborate, and interact—whether with each other or with machines.

In a world where AI has fundamentally changed the nature of risk, Alice provides end-to-end coverage across the entire AI lifecycle. We support frontier model labs, enterprises, and UGC platforms with a comprehensive suite of solutions: from model hardening evaluations and pre-deployment red-teaming to runtime guardrails and ongoing drift detection.

You've read the whole posting — now see how you match it.