Karumi logo

AI/ML Harness Engineer

Karumi

New York, NYFull-timeSeen 1mo agoStill listed 4 days ago

Most applications go out cold — see where you stand first. No sign-up to start.

Watch jobs like this.

At a glance

Compensation
No compensation found
Location
New York, NY
Schedule
Full-time
Work Authorization
Not specified • Visa sponsorship

Olive lists jobs from US employers, including remote roles you can work from the United States.

Job overview

Join the AI engineering team to build core intelligence for the platform, creating voice AI agents, browser automation, and multimodal large language model systems that interact with users in real time while ensuring production reliability and performance.

Skills & qualifications

RequiredNice to have

Skills

PythonLarge Language ModelsSpeech to TextText to SpeechDeepgramElevenLabsWhisperPlaywrightPuppeteerSeleniumPython AsyncPrompt EngineeringRetrieval SystemsObservability ToolsComputer VisionReal-Time SystemsAutonomous AgentsMulti-Step AI WorkflowsFine TuningNLPMachine Learning Research

Full job description

The Opportunity Join our AI engineering team in the US to build the core intelligence behind our platform. You'll work at the intersection of voice AI, browser automation, and large language models - creating agents that can listen, speak, navigate interfaces, and interact naturally with users in real-time. This role combines cutting-edge AI with practical systems work. You'll design voice experiences, build browser agents that understand and control web applications, and optimize LLM behavior for production reliability. We ship working AI features that solve real problems, balancing innovation with pragmatic constraints. We sponsor visas for qualified candidates. Core Responsibilities

  • Build and optimize voice AI systems using speech-to-text and text-to-speech models

  • Design browser agents that navigate, understand, and interact with web applications

  • Implement browser automation with computer vision and DOM understanding

  • Engineer prompt systems and LLM workflows for consistent, intelligent behavior

  • Create evaluation frameworks to measure voice quality, agent accuracy, and user experience

  • Integrate multimodal AI - combining voice, vision, and language understanding

  • Build real-time AI pipelines where latency and reliability are critical

  • Manage the AI Infrastructure and take care of it

  • Monitor and improve AI system performance in production environments Technical Requirements

  • Production experience with LLMs (OpenAI, Anthropic, or open-source models)

  • Hands-on work with speech AI (STT/TTS systems like Deepgram, ElevenLabs, Whisper)

  • Experience with browser automation (Playwright, Puppeteer, Selenium) or computer vision

  • Strong Python skills with async programming and real-time systems

  • Understanding of prompt engineering, retrieval systems, and agent frameworks

  • Ability to debug complex AI behaviors and build observability tools

  • Software engineering fundamentals for production AI systems Nice to Have

  • Experience building autonomous agents or multi-step AI workflows

  • Knowledge of computer vision for UI understanding and visual grounding

  • Fine-tuning or training language models for specialized tasks

  • Real-time audio processing and streaming architectures

  • Background in NLP, machine learning research, or AI systems Why Karumi

  • Meaningful equity stake in a backed, fast-growing company\

  • Work on cutting-edge voice AI and browser agents in production\

  • Shape how AI systems interact with users and software interfaces\

  • Small team with direct impact on core product capabilities\

  • Gym

  • Visa sponsorship available

Similar jobs, posted recently

Open roles like this one, listed in the last 30 days.

You've read the whole posting — now see how you match it.