Research Scientist/Research Engineer, Midtraining
Menlo Park, CAJob$250–350K/yrPosted 3 days agoStill listed 2 days ago
Most applications go out cold — see where you stand first. No sign-up to start.
Watch jobs like this. New roles like this one near Menlo Park, CA, by email.
Don't just apply. Show up ready.
Olive works from this exact posting.
At a glance
Olive lists jobs from US employers, including remote roles you can work from the United States.
Requirements
Credentials this posting asks for.
Job overview
Periodic Labs is an AI and physical sciences company building state‑of‑the‑art models to accelerate breakthroughs in materials, energy and beyond, seeking a Midtraining Research Engineer to curate data, generate synthetic data, build evaluations and run large‑scale training experiments.
Skills & qualifications
Skills
Qualifications
Full job description
We're an AI and physical sciences company building state-of-the-art models to accelerate breakthroughs across materials, energy, and beyond. Backed by world-class investors and growing rapidly, we operate at the pace the frontier requires. Our team brings deep expertise, genuine ownership, and a drive to push the boundaries of what's scientifically possible.
About the Role We're training frontier models to develop deep scientific knowledge and reasoning for scientific discovery. As a Midtraining Research Engineer, you'll take base models and improve their scientific reasoning: curating and generating data, building evals, and running large-scale training experiments. Your work will also lay the groundwork for our pre-training efforts down the line.
What You'll Do
-
Identify, process, and curate novel sources of scientific data for large-scale model training.
-
Generate high-quality synthetic data to fill gaps in scientific knowledge and reasoning.
-
Build evaluations that correlate with downstream scientific task performance, working closely with RL researchers, physicists, and chemists.
-
Develop and apply techniques such as self-distillation and on-policy distillation to improve model capability.
-
Design and run large-scale training experiments, partnering with supercompute engineers to scale efficiently across thousands of GPUs.
-
Build tools for yourself and the team to investigate how data choices shape model intelligence.
You Will Thrive in This Role If You Have
-
Experience training LLMs on curated mixes of trillions of tokens.
-
Experience on a dedicated evals team supporting a large production training run.
-
Hands-on use of self-distillation, on-policy distillation, or similar methods in a real training pipeline.
-
Experience with scaling laws and compute-optimal hyperparameters.
-
Comfort working across data, evals, and training infrastructure.
Especially Strong Candidates May Also Have
-
Experience optimizing throughput and reliability for large-scale distributed training runs.
-
A background in AI for science or training on specialized domain data (e.g., protein, materials, or other scientific datasets).
-
Experience creating evals or synthetic data for non verifiable tasks and tracking performance over live runs.
Mechanics
-
Minimum education: Bachelor's degree or similar experience
-
Location: Menlo Park, CA
-
Compensation: $250,000–$350,000 + equity
-
Visa sponsorship: Yes, we sponsor visas and will do everything we can to assist in this process.
Similar jobs, posted recently
Open roles like this one, listed in the last 30 days.
Research Computer ScientistVeterans Health Administration · Palo Alto, CAPosted 1w agoPosted 1w ago
PhD Research Intern, Quantum Simulation and AI - 2027NVIDIA · Santa Clara, CA · $38–94/hrPosted 1w agoPosted 1w agoMTS - Agent ResearchEtched · San Jose, CA · $175–275K/yrPosted 1w agoPosted 1w ago
2027 Summer Intern, PhD, Safety Research, Human Behavior AnalyticsWaymo · Mountain View, CA (Hybrid) · $85/hrPosted 2w agoPosted 2w ago
AI Research Scientist - SafetyBosch Group · Sunnyvale, CA · $165–185K/yrPosted 3w agoPosted 3w ago
You've read the whole posting — now see how you match it.