Ema logo

AI Resident

Ema

Bay Area, CAHybridFull-time$4,000/moPosted 3w agoVerified open 6 days ago

Most applications go out cold — see where you stand first. No sign-up to start.

At a glance

Compensation
$4,000/mo
Location
Bay Area, CAHybrid
Schedule
Full-time
Work Authorization
Not specified

Job overview

Ema is hiring an AI Resident. Ema is building an Agentic AI platform that transforms enterprise productivity, enabling organizations to delegate repetitive tasks to a universal AI employee. The AI Resident will own a hard problem end‑to‑end, writing proposals, building systems, designing evaluations, and shipping production code with a senior mentor, while contributing to research on inference, reward modeling, and efficiency.

Key focus areas include Own a hard problem end‑to‑end, from proposal to production deployment, Write system proposals and design evaluation frameworks, and Build and ship code within a large production codebase.

Important skills include ML Fundamentals, Strong Engineering, Python, PyTorch, Discipline To Ship In A Large Production Codebase, and Post-Training (SFT/DPO/GRPO-Family RL). Preferred (not required): Hands-On Post-Training With Open Models, TRL, VeRL, and OpenRLHF.

Skills & qualifications

RequiredNice to have

Skills

ML FundamentalsStrong EngineeringPythonPyTorchDiscipline to Ship in a Large Production CodebasePost-Training (SFT/DPO/GRPO-Family RL)Reward ModelingLLM JudgesAgent and Tool-Use SystemsRetrieval and MemoryEval DesignStatistical LiteracyHonest MeasurementHands-on Post-Training With Open ModelsTRLVeRLOpenRLHFDebugging Reward-Hacked RunBuilt or Trained in Interactive Agent EnvironmentsSWEWebTool-Use GymsLarge-Scale Trace AnalysisData CurationSynthetic Data WorkServing and Efficiency ExperienceVLLM/SGLangDistillationQuantizationMulti-Node GPU TrainingInfra FluencyPublicationsOpen-Source WorkWritingSecurity InstinctsPrompt InjectionData Governance

Qualifications

Demonstrated Depth in ML or Agent Systems

Full job description

ABOUT EMA

Ema is building the world’s leading Agentic AI platform to transform enterprise productivity. We enable organizations to delegate repetitive tasks to Ema, the Universal AI Employee, delivering 10x gains in workforce efficiency, across functions. Founded by former executives from Google, Coinbase, Flipkart, and Okta, our team includes engineers from premier tech companies and graduates of Stanford, MIT, UC Berkeley, CMU, and IITs.

We are backed by industry leading investors including Accel, Naspers/Prosus, Section32, and angels like Sheryl Sandberg and Dustin Moskovitz. Headquartered in Silicon Valley and with offices in London, Bangalore and Vancouver and Bangalore, Ema is at the frontier of what Agentic AI can do in production — we ship real systems that run real business processes at scale.

The residency

You own one hard problem end to end. You write the proposal, build the system, design the evaluation, ship behind a gate, and finish with a write-up of what turned out to be true, including the parts that didn't work. You'll sit in the production codebase with a senior mentor and real production data. Recent residents have shipped self-improving harnesses, inference-cost work, agent memory, and eval infrastructure. Your project gets scoped with you, not handed to you.

The problem space

The loop we care about: production traces become data, data becomes training and evaluation, and better agents produce better traces. Projects live somewhere on that loop.

  • Harness and inference-time work. Context engineering, tool and skill design, orchestration, and deciding where extra inference compute actually pays. Self-improvement loops run behind hard fences.

  • Post-training for agents. SFT on curated trajectories, preference optimization, RL on real agent tasks. Reward design where outcomes are verifiable, process vs. outcome supervision, distilling frontier behavior into cheaper models.

  • Environments and rewards. Turning enterprise workflows into training and eval environments: fixture tenants, simulated users (some of whom get impatient and leave), verifiable rewards, and defenses against reward hacking. Agents will exploit a lazy grader.

  • Data engines. Mining production agent-steps into training and eval corpora: failure mining, labeling with calibrated judges, synthetic augmentation that stays useful.

  • Evaluation. Behavior-level benchmarks from real workflows, LLM judges calibrated against human labels, reliability statistics for stochastic agents.

  • Efficiency. Routing, ensembles, caching, small-model specialization. Quality per dollar is a research metric here.

What we're looking for

  • No specific degree required. Strong undergrads, grad students, and self-taught builders are all welcome; what matters is demonstrated depth in ML or agent systems.

  • Solid ML fundamentals and strong engineering: Python, PyTorch, and the discipline to ship in a large production codebase.

  • Real depth in at least one of: post-training (SFT/DPO/GRPO-family RL), reward modeling or LLM judges, agent and tool-use systems, retrieval and memory, eval design. One area you can teach us beats five you've touched.

  • Statistical literacy. You can size an experiment, and you know 25 samples at one seed is a datapoint, not a result.

  • Honest measurement as a habit. You'd rather kill your own feature with a clean experiment than ship it on a hunch.

Nice to have

  • Hands-on post-training with open models (TRL, veRL, OpenRLHF, or your own loop). Bonus points if you've debugged a reward-hacked run.

  • Built or trained in interactive agent environments (SWE, web, or tool-use gyms).

  • Large-scale trace analysis, data curation, or synthetic data work.

  • Serving and efficiency experience: vLLM/SGLang, distillation, quantization.

  • Multi-node GPU training, or the infra fluency to get there fast.

  • Publications, open-source work, or writing that shows how you think.

  • Security instincts: prompt injection, data governance, why a self-improving agent needs a fence.

Logistics

  • SF Bay Area, on-site/hybrid, half/full-time for the term. Flexible start.

  • Salary: $4,000 month

Compensation offered will be determined by factors such as location, level, job-related knowledge, skills, and experience. Certain roles may be eligible for variable compensation, equity, and benefits.

Ema Unlimited is an equal opportunity employer and is committed to providing equal employment opportunities to all employees and applicants for employment without regard to race, color, religion, sex, national origin, age, disability, sexual orientation, gender identity, or genetics.

You've read the whole posting — now see how you match it.