Member of Technical Staff - Research
Most applications go out cold — see where you stand first. No sign-up to start.
Don't just apply. Show up ready.
Olive works from this exact posting — no sign-up to start.
At a glance
Requirements
Credentials this posting asks for.
Job overview
Armadin is hiring a Member of Technical Staff - Research. Armadin is building autonomous proactive security using AI, bringing together researchers and engineers to create models that detect and remediate threats before breaches occur, leveraging post‑training strategies and real‑world security tradecraft.
Key focus areas include Own post‑training strategy and drive roadmap for fine‑tuning and reward modeling, Push efficiency and build benchmarks to evaluate model reliability at scale, and Ground research in reality by translating security tradecraft into production capabilities.
Important skills include Reinforcement Learning, Research, Collaboration, Machine Learning, Artificial Intelligence, and SFT.
Skills & qualifications
Skills
Qualifications
Benefits
Full job description
About Us
At Armadin, we're a group of engineers, researchers, and hackers on a mission to redefine what proactive security can do in the AI era. Cyberattacks are becoming autonomous and relentless and we believe defending against threats before they materialize is one of the most powerful ways to protect the institutions the world depends on.
We're building autonomous proactive security from the ground up, reinforcing the tradecraft of elite red teamers into purpose-built security models and agents that discover risk and remediate it before organizations are breached.
Led by Kevin Mandia, founder of Mandiant ($5.4B exit to Google), our team brings together researchers and engineers from Google, xAI, Meta, Stanford, and MIT to reinvent security for an adversary that never sleeps.
What You'll Do
-
Own Post-Training Strategy: Drive the post-training roadmap (fine-tuning, preference optimization, reward modeling, RL, distillation) to make models more capable, reliable, and aligned.
-
Push Efficiency & Evals: Make models serve reliably at scale, and build the benchmarks that measure quality and catch regressions.
-
Ground Research in Reality: Turn our security experts' tradecraft into model capabilities and ship your techniques into production.
-
Stay at the Frontier: Track post-training and efficiency research and bring the best ideas in.
What You'll Bring
-
Research Pedigree: PhD in CS, ML, or related field, or equivalent experience with a strong track record.
-
Post-Training Depth: Hands-on experience with SFT, RLHF/RLAIF, DPO, and reward modeling.
-
Practical Results: You train, fine-tune, and evaluate models, and make research work in practice.
-
Engineering Strength: You implement your own ideas and run them at scale.
-
Range & Judgment: You drive strategy broadly or go deep on one hard problem as needed.
-
Collaboration: You work well across teams and care about the outcome, not just your slice.
Even Better With
-
Publications in top ML venues (e.g., NeurIPS, ICML, ICLR, ACL).
-
RL, preference optimization, or reward modeling at scale.
-
Research from prototype to production in a fast-moving company.
-
Enterprise products built for reliability, scale, and trust.
Benefits & Perks
🏥 Full Health, Dental, & Vision Coverage
📈 Meaningful Equity Ownership
🥙 In-Office Meals
✂️ Haircuts at the Office 🎉 Company Sponsored Conferences & Events
💸 401(k), HSA, and FSA Plans
🌴 Flexible PTO
You've read the whole posting — now see how you match it.