Research Scientist, Scaling RL
Menlo Park, CAJob$250–350K/yrPosted 3 days agoStill listed 2 days ago
Most applications go out cold — see where you stand first. No sign-up to start.
Watch jobs like this. New roles like this one near Menlo Park, CA, by email.
Don't just apply. Show up ready.
Olive works from this exact posting.
At a glance
Olive lists jobs from US employers, including remote roles you can work from the United States.
Requirements
Credentials this posting asks for.
Job overview
Periodic Labs is an AI and physical sciences company building state-of-the-art models to accelerate breakthroughs across materials, energy, and beyond, seeking a Research Scientist to study scaling of reinforcement learning and develop advanced algorithms.
Skills & qualifications
Skills
Qualifications
Full job description
About Periodic Labs We're an AI and physical sciences company building state-of-the-art models to accelerate breakthroughs across materials, energy, and beyond. Backed by world-class investors and growing rapidly, we operate at the pace the frontier requires. Our team brings deep expertise, genuine ownership, and a drive to push the boundaries of what's scientifically possible.
About the Role We're training frontier models to develop deep scientific knowledge and reasoning for scientific tasks. You’ll study how RL scales with training compute, develop better algorithms, and take ideas from controlled experiments to our largest runs like Periodic Neon.
What You'll Do
-
Design experiments to understand how RL performance scales with compute, model size, data, and reward quality, building on work such as ScaleRL
-
Develop better RL algorithms, spanning policy optimization, advantage estimation, exploration, and credit assignment for long-horizon RL tasks
-
Build adaptive sampling and curriculum methods that adjust task difficulty, problem selection, and the number of rollouts as models improve
-
Study bias and stability during RL training, including importance-sampling corrections and methods to tackle policy staleness and training–inference mismatch, as discussed here.
-
Improve compute efficiency across training and inference through experiments with hyperparameters, such as length penalties, rollout counts, batch sizes, and update schedules.
You Will Thrive in This Role If You Have
-
Hands-on experience training LLMs with reinforcement learning
-
Strong attention to detail and rigorous approach to answer questions scientifically.
-
Coming up with small-scale RL setups that transfers to large-scale training runs.
-
Comfort working across a complex training stack to implement, debug, and test new research ideas.
Mechanics Minimum experience: 5+ years Minimum education: Bachelor’s degree or similar experience Location: Menlo Park, CA Compensation: $250,000-$350,000 base + equity Visa sponsorship: Yes, we sponsor visas and will do everything we can to assist in this process.
Similar jobs, posted recently
Open roles like this one, listed in the last 30 days.
Staff Research ScientistGeneral Motors · Sunnyvale, CA (Hybrid) · $219–335K/yrPosted 2w agoPosted 2w ago
Research Computer ScientistVeterans Health Administration · Palo Alto, CAPosted 1w agoPosted 1w ago
AI Research Scientist - SafetyBosch Group · Sunnyvale, CA · $165–185K/yrPosted 3w agoPosted 3w ago
PhD Research Intern, Quantum Simulation and AI - 2027NVIDIA · Santa Clara, CA · $38–94/hrPosted 1w agoPosted 1w ago
Neuromorphic/AI Research ScientistIntel · Santa Clara, CA (Hybrid) · $171–315K/yrPosted 2w agoPosted 2w ago
You've read the whole posting — now see how you match it.