
Member of Technical Staff, Architecture & Scaling
San Jose, CA · HybridFull-time$180–450K/yrPosted 2 days agoStill listed 2 days ago
Most applications go out cold — see where you stand first. No sign-up to start.
Watch jobs like this. New roles like this one near San Jose, CA, by email.
Don't just apply. Show up ready.
Olive works from this exact posting.
At a glance
Olive lists jobs from US employers, including remote roles you can work from the United States.
Job overview
Hark is an artificial intelligence company developing multimodal models and next‑generation AI hardware. The role focuses on architecture and scaling of the largest models, including research on model design, optimization, training recipes, and cross‑modal work, with a direct line to production and frontier compute runs.
Skills & qualifications
Skills
Full job description
About Hark
Hark is an artificial intelligence company building advanced, personalized intelligence. One that is proactive, multimodal, and capable of interacting with the world through speech, text, vision, and persistent memory.
We're pairing that intelligence with next-generation hardware to create a universal interface between humans and machines. While today's AI largely operates through chat boxes and decade-old devices, Hark is focused on what comes next: agentic systems that interact naturally with people and the real world.
To get there, we're developing multimodal models and next-generation AI hardware together - designed from the ground up as a single, unified interface for a new era of intelligent systems.
About the Role
You'll work on the architecture and scaling of our largest models: what we train, how we train it, and how to spend the next order of magnitude of compute well. This is empirical research with a direct line to production. The recipes you set are the recipes our frontier runs use.
Responsibilities
-
Conduct research on model architecture, optimization, and scaling to improve the capability and efficiency of our largest models.
-
Establish strong baselines and run controlled experiments to determine which ideas actually scale to frontier training runs.
-
Set training recipes at scale: learning-rate schedules, context length, token budgets, and compute allocation.
-
Explore new architectures, including mixture-of-experts, hybrid attention, and long-context extension.
-
Diagnose instability in large runs: loss spikes, divergence, numerical issues, and the infrastructure failures that look like research problems.
-
Work across modalities, since our models are multimodal from pretraining forward.
Requirements
-
Hands-on experience training large models, at a scale where compute allocation and stability decisions carry real cost.
-
Strong empirical instincts: you design the experiment that distinguishes between two hypotheses instead of the one that confirms the first.
-
Fluency with scaling laws and how to use small-scale results to make a frontier-scale bet.
-
Deep familiarity with a modern training stack and distributed training across large GPU clusters.
-
Strong engineering. Research here means writing the code and reading the profiler, not handing off a spec.
-
A record of work that shipped into real models, whether that shows up as papers, systems, or production runs.
Bonus Qualifications
-
Experience with mixture-of-experts routing, sparse architectures, or long-context methods.
-
Work on data mixtures, curriculum, or tokenizer design.
-
Kernel-level optimization or mixed-precision training experience.
-
Experience with efficiency work aimed at constrained inference targets, including on-device.
Compensation
The US base salary range for this full-time position is between $180,000 - $450,000 annually.
The pay offered for this position may vary based on several individual factors, including job-related knowledge, skills, and experience. The total compensation package may also include additional components/benefits depending on the specific role. This information will be shared if an employment offer is extended.
Similar jobs, posted recently
Open roles like this one, listed in the last 30 days.
Staff ML Infrastructure Engineer - Embodied AI Scaling FoundationsGeneral Motors · Sunnyvale, CA (Hybrid) · $172–335K/yrPosted 2w agoPosted 2w ago
Member of Technical Staff, Frontier RL EnvironmentsHark · San Jose, CA · $180–450K/yrPosted 3w agoPosted 3w ago
Senior Architecture Energy Modeling EngineerNVIDIA · Santa Clara, CA · $136–265K/yrPosted 4w agoPosted 4w ago
Staff Technical Program Manager - SecuritySnowflake Inc. · Menlo Park, CA (Hybrid) · $194–278K/yrPosted 1w agoPosted 1w ago
Architecture Energy Modeling Engineer - New College Grad 2026NVIDIA · Santa Clara, CA · $116–190K/yrPosted 4w agoPosted 4w ago
You've read the whole posting — now see how you match it.