
Machine Learning Engineer
San Francisco, CAFull-timeSeen 1mo agoStill listed 2 days ago
Most applications go out cold — see where you stand first. No sign-up to start.
Watch jobs like this. New roles like this one near San Francisco, CA, by email.
Don't just apply. Show up ready.
Olive works from this exact posting.
At a glance
Olive lists jobs from US employers, including remote roles you can work from the United States.
Job overview
Osmosis seeks a Machine Learning Engineer to develop high‑performance distributed training infrastructure for reinforcement learning at scale, collaborating with the founding team and design partners to advance post‑training and continual learning systems in a fast‑paced, customer‑driven environment.
Skills & qualifications
Skills
Full job description
About Osmosis At Osmosis, we help companies use cutting-edge reinforcement learning techniques to fine-tune open-source language models that beat foundation models on performance, latency, and cost. We’ve raised $7M in funding from Y Combinator, top institutional investors like CRV and Audacious Ventures, as well as angel investors including Paul Graham (Y Combinator), Erik Bernhardsson (Modal Labs), Misha Laskin (Reflection AI), and Guillermo Rauch (Vercel). About the Role We're looking for a Machine Learning Engineer to contribute to high-performance distributed training infrastructure for RL at scale. You'll work directly with our founding team and design partners to push the boundaries of what's possible with post-training and continual learning systems. This role requires expertise in RL algorithms, distributed training, and low-level optimization. You'll have exceptional agency to make impactful decisions while working in a fast-paced, customer-driven environment. Responsibilities You’ll contribute to work in areas like:
- Distributed Training Infrastructure: implement new RL algorithms and build scalable post-training pipelines
- Resource Management & Optimization: design infrastructure systems for efficient GPU utilization and dynamic resource allocation
- Customer-Facing Work: work directly with customers on production deployments and custom model development
Technology
- Backend: Python FastAPI, Golang
- Frontend: React, TypeScript, Next.js
- Cloud Infrastructure: AWS Fargate, Docker, Kubernetes, AWS SageMaker
- ML Frameworks: Verl / slime / Megatron-LM / SkyRL, PyTorch (FSDP experience is a plus), vLLM / SGLang
- Databases: DynamoDB, S3
Similar jobs, posted recently
Open roles like this one, listed in the last 30 days.
Machine Learning EngineerLatent · San Francisco, CA · $225–300K/yrPosted 3 days agoPosted 3 days ago
Machine Learning EngineerWitness AI · Mountain View, CA · $36–60K/yrPosted 4w agoPosted 4w ago
Machine Learning EngineerAndromeda Surgical · South San Francisco, CA · $150–250K/yrPosted 2w agoPosted 2w ago
Machine Learning Engineerreddit · San Francisco, CA · $230–322K/yrPosted 1 day agoPosted 1 day ago
Machine Learning Engineer InternCoinbase · San Francisco, CA (Hybrid) · $60/hrPosted 3w agoPosted 3w ago
You've read the whole posting — now see how you match it.