Robinhood Markets logo

Staff Software Engineer, Observability

Robinhood Markets

Menlo Park, CAHybridJob$230–270K/yrPosted 3w agoVerified open 3 days ago

Most applications go out cold — see where you stand first. No sign-up to start.

At a glance

Compensation
$230–270K/yr
Location
Menlo Park, CAHybrid
Work Authorization
Not specified

Job overview

Robinhood Markets is hiring a Staff Software Engineer, Observability . Robinhood’s Observability team seeks a Staff Software Engineer to lead the full‑stack observability platform, defining roadmap, owning the control plane, and ensuring 99.9% uptime. The role involves high‑ownership technical decisions, cost‑efficient telemetry ingestion, and collaboration with engineering and the Command Center, based in Menlo Park with hybrid work.

Key focus areas include Define and execute full‑stack observability roadmap establishing technical vision, Own and evolve observability control plane including telemetry pipelines and cost management, and Lead and collaborate with a team of six engineers to build scalable self‑service solutions.

Successful candidates bring 8+ Years Software Engineering Experience, Vector Data Pipeline Or Equivalent, and In-Person Attendance At Least 3 Days Per Week. Important skills include Kubernetes, Go, Python, Architect Distributed Systems, Operate Distributed Systems, and Integrating Observability Agents. Preferred (not required): High-Throughput Log Routing, High-Throughput Metrics Routing, and AI-Driven Approaches.

Skills & qualifications

RequiredNice to have

Skills

KubernetesGoPythonArchitect Distributed SystemsOperate Distributed SystemsIntegrating Observability AgentsIntegrating Observability LibrariesInstrumentation Into Production CodebasesObservability Control Plane OwnershipTelemetry Pipeline OwnershipIngestion Cost ManagementCardinality ControlSignal RoutingDefine Observability RoadmapExecute Observability RoadmapEstablish Technical VisionOwn Observability Control PlaneEvolve Observability Control PlaneLead Engineering TeamCollaborate With Engineering TeamBuild Scalable Observability SolutionsBuild Self-Service Observability SolutionsEstablish SLOs for Observability SystemsMaintain SLOs for Observability SystemsPartner With Robinhood Command CenterPartner With Engineering TeamsAlign on Dependency MappingAlign on Incident Response WorkflowsAlign on Observability StandardsAWSVector Data PipelinePrometheusGrafanaHoneycombHumioSentryIntegrate Observability AgentsIntegrate Observability LibrariesInstrumentationTelemetry Ingestion Cost ManagementTechnical LeadershipObservability Control PlaneTelemetry PipelineHigh-Throughput Log RoutingHigh-Throughput Metrics RoutingArchitectural Decision MakingStrategy for Cost-Efficient Telemetry IngestionBuilding in-House SolutionsIntegrating AI-Driven ApproachesDefining Observability RoadmapExecuting Observability RoadmapEstablishing SLOsPartnering With Command CenterAligning on Dependency MappingAligning on Incident Response WorkflowsAligning on Observability StandardsEstablishing Technical VisionTelemetry Ingestion PipelinesCost Management StrategiesEnsuring Observability Components Highly AvailableBuilding Scalable Self-Service Observability SolutionsEstablishing SLOs for Observability SystemsMaintaining SLOs for Observability SystemsPartnering With Robinhood Command CenterMetrics InfrastructureLogs InfrastructureTraces InfrastructureAlerting InfrastructureEstablish SLOsMaintain SLOsDependency MappingIncident Response WorkflowsObservability StandardsIntegrate InstrumentationDrive Strategy for Cost-Efficient Telemetry IngestionBuild in-House SolutionsIntegrate AI-Driven ApproachesObservabilityTelemetry IngestionTracingMetricsLoggingAlertingSLOsCost ManagementSLO DefinitionVectorDistributed SystemsPublic Cloud EnvironmentsTracesIncident ManagementAI‑Driven ApproachesLeadershipTeam CollaborationObservability AgentsAI ToolsControl PlaneDistributed Systems ArchitectureDistributed Systems at ScalePublic CloudSLO Management

Qualifications

8+ Years Software Engineering ExperienceIn-Person Attendance 3 Days Per Week

Benefits

Medical Insurance
401(k) Match
Paid Time Off
Parental Leave

Full job description

Join us in building the future of finance.

Our mission is to democratize finance for all. An estimated $124 trillion of assets will be inherited by younger generations in the next two decades. The largest transfer of wealth in human history. If you’re ready to be at the epicenter of this historic cultural and financial shift, keep reading.

About the team + role

We are building an elite team, applying frontier technologies to the world's biggest financial problems. We're looking for bold thinkers. Sharp problem-solvers. Builders who are wired to make an impact. Robinhood isn't a place for complacency, it's where ambitious people do the best work of their careers. We're a high-performing, fast-moving team with ethics at the center of everything we do. Expectations are high, and so are the rewards.

The Observability team's mission is to build and own Robinhood's full-stack observability platform — the foundation that keeps every product, service, and customer experience running reliably at scale. We design and operate the systems that give engineers deep visibility into how Robinhood's infrastructure behaves, ensuring that when something goes wrong, the right people know immediately and can act fast. Our work spans metrics, logs, distributed tracing, and alerting pipelines, and we partner closely with the Robinhood Command Center to ensure our observability systems meet or exceed 99.9% uptime. We believe in monitoring our own monitors — the observability infrastructure is a product, not just a tool.

As a Staff Software Engineer on the Observability team, you will be the technical lead shaping the roadmap for how Robinhood observes itself at scale. You will own the observability control plane end-to-end, making architectural decisions that directly impact the reliability and operational health of Robinhood's entire product surface. You'll lead a team of six engineers, drive the strategy for cost-efficient telemetry ingestion, build in-house solutions where off-the-shelf tooling falls short, and integrate AI-driven approaches to accelerate progress toward full-stack observability. This is a high-ownership, high-visibility role where your technical decisions set the direction for the entire organization!

This role is based in our Menlo Park, CA office, with in-person attendance expected at least 3 days per week.

At Robinhood, we believe in the power of in-person work to accelerate progress, spark innovation, and strengthen community. Our office experience is intentional, energizing, and designed to fully support high-performing teams.

What you'll do

  • Define and execute the full-stack observability roadmap, establishing a clear technical vision for metrics, logs, traces, and alerting infrastructure across Robinhood's engineering organization.
  • Own and evolve the observability control plane, including telemetry ingestion pipelines, cost management strategies, and the tooling that ensures observability components remain highly available.
  • Lead and collaborate with a team of six engineers to build scalable, self-service observability solutions that enable product and infrastructure teams to move faster with greater confidence.
  • Establish and maintain SLOs for observability systems that meet or exceed Robinhood's 99.9% uptime target, ensuring the observability platform is as reliable as the services it monitors.
  • Partner with the Robinhood Command Center and engineering teams across the organization to align on dependency mapping, incident response workflows, and observability standards.

What you bring

  • 8+ years of software engineering experience, with a proven track record of owning and delivering large-scale observability or infrastructure platform initiatives.
  • Deep expertise in Kubernetes and public cloud environments (AWS preferred), with the ability to architect and operate distributed systems at scale regardless of specific vendor tooling.
  • Strong coding proficiency in one or more languages (Go, Python, or similar) with experience integrating observability agents, libraries, and instrumentation directly into production codebases.
  • Demonstrated experience owning an observability control plane or telemetry pipeline — including ingestion cost management, cardinality control, and signal routing — in a high-traffic production environment.
  • Experience with the Vector data pipeline (or equivalent high-throughput log/metrics routing tools) and familiarity with tools such as Prometheus, Grafana, Honeycomb, Humio, or Sentry is a plus.

What we offer

  • Challenging, high-impact work to grow your career
  • Performance driven compensation with multipliers for outsized impact, bonus programs, equity ownership, and 401(k) matching
  • Top Tier benefits to fuel your work, including 100% paid health insurance for employees with 90% coverage for dependents
  • Access to the best AI tools on the market and continuous AI skill-building for every employee, technical or not
  • Lifestyle wallet - a highly flexible benefits spending account for wellness, learning, and more
  • Employer-paid life & disability insurance, fertility benefits, and mental health benefits
  • Time off to recharge including company holidays, paid time off, sick time, parental leave, and more!
  • Exceptional office experience with catered meals, events, and comfortable workspaces.

In addition to the base pay range listed below, this role is also eligible for bonus opportunities + equity + benefits.

Base pay for the successful applicant will depend on a variety of job-related factors, which may include education, training, experience, location, business needs, or market demands. The expected base pay range for this role is based on the location where the work will be performed and is aligned to one of 3 compensation zones. For other locations not listed, compensation can be discussed with your recruiter during the interview process.

Base Pay Range:

Zone 1 (Menlo Park, CA; New York, NY; Bellevue, WA; Washington, DC)

$230,000—$270,000 USD

Zone 2 (Denver, CO; Westlake, TX; Chicago, IL)

$203,000—$238,000 USD

Zone 3 (Lake Mary, FL; Clearwater, FL; Gainesville, FL)

$180,000—$211,000 USD

Click here to learn more about our Total Rewards, which vary by region and entity.

If our mission energizes you and you’re ready to build the future of finance, we look forward to seeing your application.

Robinhood provides equal opportunity for all applicants, offers reasonable accommodations upon request, and complies with applicable equal employment and privacy laws. Inclusion is built into how we hire and work—welcoming different backgrounds, perspectives, and experiences so everyone can do their best. Please review the Privacy Policy for your country of application.

You've read the whole posting — now see how you match it.