The Subvocal Company logo

Head of ML

The Subvocal Company

San Francisco, CAFull-time$150–200K/yrSeen 1mo agoStill listed 2 days ago

Most applications go out cold — see where you stand first. No sign-up to start.

Watch jobs like this.

At a glance

Compensation
$150–200K/yr
Location
San Francisco, CA
Schedule
Full-time
Work Authorization
Not specified

Olive lists jobs from US employers, including remote roles you can work from the United States.

Job overview

The role leads the central machine learning effort at Subvocal, turning weak physiological signals into continuous language and guiding data collection, model development, and evaluation.

Skills & qualifications

RequiredNice to have

Skills

Machine LearningSpeech RecognitionPythonPytorchCudaDistributed GPU TrainingSignal ProcessingData CollectionExperiment TrackingEvaluationConformersMamba‑Style ArchitecturesCtcTransducersPretrained Speech ModelsLanguage ModelsDeep Learning for SpeechTime SeriesBiosignalsNeuroscienceBcisRadar

Qualifications

Phd-Level Research Ability

Full job description

We are looking for the person who will own the central machine learning problem at Subvocal: turning weak, noisy, highly variable physiological signals into continuous language. We have already built prototypes that decode subvocal speech at more than 200 words per minute. The much harder problem now is generalization. A model that works on one person, in one session, with one device placement is not a product. It needs to work when the same person returns the next day, when the hardware moves slightly, and eventually when a completely new person puts it on and gives us only a few minutes of calibration data. You will lead that effort end to end. You will work directly with the founders and our sensing and hardware leads to decide what data we collect, how we represent the signal, which model families we pursue, and how we evaluate whether we are actually making progress. Some of the problems you will work on include:

  • Learning general representations from large amounts of unlabeled and weakly labeled physiological time-series data.
  • Building continuous sequence models using approaches such as Conformers, Mamba-style architectures, CTC, transducers, and pretrained speech or language models.
  • Separating speech-related information from anatomy, placement, session, device, and environmental variation.
  • Adapting a large cross-user model to a new person from roughly 15 minutes of calibration data.
  • Designing honest user-held-out, session-held-out, and device-held-out evaluations.
  • Scaling training across thousands of hours and thousands of people.
  • Getting the complete system to run continuously with low enough latency for real-time use.

Our current ML stack is primarily Python, PyTorch, CUDA, and distributed GPU training, with custom infrastructure for signal processing, data collection, experiment tracking, and evaluation. You might be a great fit if you have unusually strong experience in deep learning for speech (ASR), time series, biosignals, neuroscience, BCIs, radar, or another domain where signals are noisy and data distributions shift constantly. We are especially interested in people with PhD-level research ability, whether or not that came through a formal PhD, who are also comfortable writing production-quality code and moving quickly when the research direction changes. This is not a role where you will be handed a model architecture and asked to improve it incrementally. You will help decide how the problem should be framed in the first place, build the initial ML organization around you, and directly determine whether this technology becomes a real product. This is a full-time, in-person role in San Francisco.

Similar jobs, posted recently

Open roles like this one, listed in the last 30 days.

You've read the whole posting — now see how you match it.