Sieve logo

Member of Technical Staff, Forward Deployed

Sieve

San Francisco, CAFull-time$150–350K/yrPosted 3mo agoStill listed 3w ago

Most applications go out cold — see where you stand first. No sign-up to start.

Watch jobs like this.

At a glance

Compensation
$150–350K/yr
Location
San Francisco, CA
Schedule
Full-time
Work Authorization
Not specified

Olive lists jobs from US employers, including remote roles you can work from the United States.

Job overview

Sieve is hiring a Member of Technical Staff, Forward Deployed. Sieve is a multi‑modal lab that curates the world’s highest‑quality training datasets across video, audio, images, text, and 3D. Leveraging exabyte‑scale infrastructure and novel multimodal techniques, the company partners with leading AI labs to deliver high‑quality data that powers creative, communication, gaming, AR/VR, and robotics applications.

Key focus areas include Work closely with customers and internal teams to understand data needs, Translate ambiguous requirements into production systems for dataset creation, and Build custom algorithms, model workflows, and large‑scale data pipelines.

Successful candidates bring Onsite In San Francisco. Important skills include Python, PyTorch, ML Frameworks, Building Custom Algorithms, Model Workflows, and Large-Scale Data Pipelines. Preferred (not required): Audio Processing, Multimodal Data Processing, Open Source Contributions, and Active Open Source Contributor.

Skills & qualifications

RequiredNice to have

Skills

PythonPyTorchML FrameworksBuilding Custom AlgorithmsModel WorkflowsLarge-Scale Data PipelinesDataset QualityDataset FilteringDataset LabelingDataset EvaluationEdge CasesClean CodeMaintainable CodeVideoMedia TechnologiesFrontier AI ApplicationsCustomer InteractionBias to ActionUntangling Messy RequirementsShipping FastDelivering End-to-End OutcomesAudio ProcessingMultimodal Data ProcessingOpen Source ContributionsCustom AlgorithmsDataset Quality IntuitionData FilteringData LabelingData EvaluationClean Maintainable CodeTranslating Ambiguous NeedsActive Open Source Contributor

Qualifications

Onsite in San FranciscoEarly Hire at Startup Experience

Benefits

Medical Insurance
401(k) Match
Dental Insurance
Vision Insurance
Paid Time Off

Full job description

ABOUT US

Sieve is a multi-modal lab curating the world's highest-quality training datasets — spanning video, audio, images, text, and 3D. We combine exabyte-scale data infrastructure and novel multimodal understanding techniques that push the frontier of foundation models. Video alone makes up 80% of internet traffic, and across modalities, data has become the enabling medium powering creativity, communication, gaming, AR/VR, and robotics. Sieve exists to solve the biggest bottleneck in the growth of these applications: high-quality training data.

We partner with top AI labs and did $XXM last quarter alone, as a team of ~30 people. We also raised our Series A from Tier 1 firms such as Matrix Partners https://matrix.vc/, Swift Ventures https://www.swift.vc/, Y Combinator https://www.ycombinator.com/, and AI Grant https://aigrant.com/.

 

WHY NOW

Sieve is one of the most capital-efficient teams in AI — roughly 30 people serving the world's leading AI labs across every major data modality. You'll join early, own problems end-to-end, and watch your work ship directly into the models defining the frontier.

ABOUT THE ROLE

As a Forward Deployed Engineer at Sieve, you’ll work on highly specific dataset problems for frontier AI labs. We're looking for someone with a strong bias to action who likes working closely with customers, untangling messy requirements, and shipping fast.

You’ll work closely with customers and internal teams to understand exactly what data is needed, then turn ambiguous requirements into production systems that can find, generate, filter, transform, evaluate, and package high-quality video datasets at scale.

REQUIREMENTS

  • Comfortable working directly with customers or external teams to translate ambiguous needs into concrete technical systems

  • Strong Python developer with hands-on experience in PyTorch or similar ML frameworks

  • Experience building custom algorithms, model workflows, or large-scale data pipelines

  • Strong intuition for dataset quality, filtering, labeling, evaluation, and edge cases

  • Able to break customer-level goals down into the models, heuristics, infrastructure, and QA steps needed to deliver

  • Writes clean, maintainable code and can move quickly without creating brittle systems

  • Deep passion for video, media technologies, and frontier AI applications

  • Motivated by delivering end-to-end outcomes, not just training models or writing research code

  • Bonus: Experience with large-scale video, audio, or multimodal data processing

  • Bonus: Active contributor to open source projects

  • Bonus: Experience as an early hire at a startup

  • In-person at our SF HQ

BENEFITS

  • 401k + Full Health Insurance

  • Breakfast, Lunch, and Dinner covered and your choice of snacks

  • Ubers covered home

*all roles at Sieve require you to be onsite in San Francisco 5 days per week

Similar jobs, posted recently

Open roles like this one, listed in the last 30 days.

You've read the whole posting — now see how you match it.