Monarch logo

ML Engineer

Monarch

Emeryville, CAFull-timePosted 1w agoStill listed 1 day ago

Most applications go out cold — see where you stand first. No sign-up to start.

Watch jobs like this.

At a glance

Compensation
No compensation found
Location
Emeryville, CA
Schedule
Full-time
Work Authorization
Not specified

Olive lists jobs from US employers, including remote roles you can work from the United States.

Job overview

The role offers the chance to work on critical technical and moral challenges by building reliable data pipelines that transform assay recordings into reproducible models and actionable recommendations, while collaborating closely with researchers in an in‑office, full‑time environment in Emeryville.

Skills & qualifications

RequiredNice to have

Skills

PythonMachine LearningGoogle CloudPyTorchJAXTensorFlowWorkflow OrchestrationComputer VisionMolecular Machine LearningActive LearningScientific Data PlatformsReliability Instinct

Qualifications

Production Software Engineering ExperienceTestingObservabilityData ValidationVersion ControlReproducible Computational WorkflowsLarge Video DatasetsStructured Scientific DataCollaboration

Full job description

We offer opportunities to do your life’s work while helping solve one of the most important technical and moral challenges of our time. Full-time, in-office in Emeryville, California. Compensation includes equity. Build the reliable systems that carry our data from an assay recording to a reproducible model, an evaluated prediction, and a usable recommendation for the next experiment. Key Responsibilities

  • Own pipelines for ingesting, validating, versioning, and joining assay videos, metadata, compound records, model features, and experimental outcomes

  • Build reproducible training and evaluation infrastructure with clear data lineage, model versioning, automated tests, and auditable outputs

  • Turn research prototypes into dependable batch and online systems that can rank compounds and surface recommendations through our tools

  • Monitor data quality, distribution shift, calibration, latency, cost, and failures as the number of labs and assays grows

  • Design interfaces between computer vision, molecular models, active-learning systems, and the lab workflow

  • Improve developer and researcher velocity without weakening scientific reproducibility or access controls Qualifications

  • Strong production software engineering experience in Python and modern machine-learning or data systems

  • Experience deploying and operating model-training, feature, evaluation, or inference pipelines in a cloud environment

  • Fluency with testing, observability, data validation, version control, and reproducible computational workflows

  • Ability to work with large video datasets and structured scientific data

  • Ability to collaborate closely with researchers while making sound engineering tradeoffs Desired Attributes

  • Experience with PyTorch, JAX, or TensorFlow and workflow-orchestration tools

  • Experience on Google Cloud or with large-scale object-storage pipelines

  • Familiarity with computer vision, molecular machine learning, active learning, or scientific data platforms

  • Instinct for simple systems, explicit failure modes, and measurable reliability

Similar jobs, posted recently

Open roles like this one, listed in the last 30 days.

You've read the whole posting — now see how you match it.