Rhoda AI logo

Senior DevOps Engineer

Rhoda AI

Mountain View, CAFull-timeNo compensation foundPosted 1mo agoVerified open 4 days ago

Most applications go out cold — see where you stand first. No sign-up to start.

At a glance

Compensation
No compensation found
Location
Mountain View, CA
Schedule
Full-time
Work Authorization
Not specified

Job overview

Rhoda AI is hiring a Senior DevOps Engineer. Rhoda AI is seeking a Senior DevOps Engineer to maintain and improve CI/CD pipelines, manage infrastructure for build artifacts, and enhance cloud and GPU clusters. This role involves building essential DevOps infrastructure to support research and software development, improving MLOps training and deployment paths, and hardening security and multi-tenancy for generalist intelligent robots.

Key focus areas include Maintain and improve CI/CD pipeline, Manage and maintain infrastructure around build artifacts and dependencies storage, and Build and improve on cloud and GPU cluster.

Successful candidates bring Strong Knowledge Of Software Engineering Best Practices, Experience With Software Build Systems, and Experience With Procedural CI And CD. Important skills include Python, C++, Rust, Go, Software Engineering Best Practices, and Design Patterns. Preferred (not required): EtherCAT, Modbus, gRPC, and MLOps.

Skills & qualifications

RequiredNice to have

Skills

PythonC++RustGoSoftware Engineering Best PracticesDesign PatternsDockerContainerized EnvironmentsSoftware Build Systems for Cloud InfrastructureSoftware Build Systems for Embedded SystemsProcedural CICDBuild Artifact Delivery SystemOperating in Linux EnvironmentSelf-Starter MentalityPrioritize IndependentlyEffective Communication SkillsWork Cross-FunctionallyEtherCATModbusgRPCMLOpsTCP/IPDNSFirewallsSecurity Best Practices for Embedded IoT DevicesKubernetesCloud System Orchestration

Qualifications

5+ Years in DevOps / SRE / Platform / Infra With Ownership of Production Systems

Full job description

At Rhoda AI, we’re building the next generation of generalist intelligent robots. We own the full robotics stack from high-performance hardware and robot systems to the infrastructure and state-of-the-art foundation world models that control our robots. Our robots are designed to be generalists capable of operating in complex, real-world environments and handling long-tail edge cases, made possible by our cutting edge research and end-to-end system design. We've raised over $450M and are investing aggressively in model research, infrastructure, hardware development, and manufacturing scale-up to make generalist robotics a reality.

Responsibility:

  • Maintain and improve our CI/CD pipeline: instrument and monitor our CI jobs, and iteratively improve on it.

  • Manage and maintain infrastructure around build artifacts and dependencies storage.

  • Build and improve on our cloud and GPU cluster, and create essential devops infrastructure to support our daily research and software development effort.

  • Improve our MLOps training and deployment path: checkpoint registry/DB, model artifact promotion, and rollout to inference/robot endpoints.

  • Improve our infrastructure observability.

  • Harden security & multi-tenancy

Qualifications

  • 5+ years in DevOps / SRE / platform / infra, with ownership of production systems.

  • Proficient in a programming language such as Python, C++, Rust, Go

  • Strong knowledge of software engineering best practices and design patterns

  • Experience with docker and containerized environments

  • Experience with software build systems for cloud infrastructure and embedded systems.

  • Experience with procedural CI and CD, and build artifact delivery (OTA update) system.

  • Comfortable with operating in Linux environment

  • Self-starter mentality — comfortable with ambiguity, able to prioritize independently, and willing to jump in wherever needed

  • Effective communication skills; able to work cross-functionally in a fast-moving team

Preferred Qualifications

  • Knowledge of industrial communication protocols (EtherCAT, Modbus, gRPC, etc.)

  • Experience with ML Ops

  • Understanding of networking fundamentals (TCP/IP, DNS, firewalls) and security best practices for embedded IoT devices

  • Knowledge with kubernetes and cloud system orchestration

You've read the whole posting — now see how you match it.