Platform Engineer (SRE) - AI Control Plane
San Francisco, CAFull-time$95–220K/yrPosted 4mo agoStill listed 4 days ago
Most applications go out cold — see where you stand first. No sign-up to start.
Watch jobs like this. New roles like this one near San Francisco, CA, by email.
Don't just apply. Show up ready.
Olive works from this exact posting.
At a glance
Olive lists jobs from US employers, including remote roles you can work from the United States.
Job overview
Speakeasy is hiring a Platform Engineer (SRE) - AI Control Plane. Speakeasy provides an AI control plane for enterprises, enabling secure scaling of AI usage with centralized management, fine‑grained permissions, threat detection and observability. The role involves owning product decisions, collaborating with GTM, and delivering high‑impact reliability, performance and availability improvements for an early‑stage product.
Key focus areas include Identify architectural changes to improve reliability, performance and availability, Foster a culture of reliability across Speakeasy's engineering organization, and Design and implement key operational processes such as deployments, upgrades, rollbacks, and postmortem review.
Successful candidates bring Participate In On-Call Rotation, Respond To Production Incidents, and Work In Person In San Francisco Office. Important skills include Reliability, Performance Improvement, Availability Improvement, Operational Process Design, Deployment, and Upgrades.
Skills & qualifications
Skills
Qualifications
Full job description
Speakeasy provides the AI control plane for enterprises.
We help companies securely scale usage of AI with a platform to centrally manage MCPs, Skills, and Assistants your whole company can use, with fine-grained permissions, threat detection, and full observability.
We are a fast growing venture backed team based out of San Francisco and London. We have a track record of building developer tools and infrastructure widely used from Fortune 500 to well known startups.
Working at Speakeasy as an engineer means being a product owner, collaborating closely with GTM and owning product decisions from inception to customer success. We care deeply about craft, quality and execution.
About the AI control plane
Our platform enables enterprises and fast-moving startups to get the most out of Claude, Codex, Cursor and other providers by
-
Leveraging enterprise identity systems to secure access to MCP servers
-
Securing agent sessions to ensure sensitive data and enterprise policies are respected.
-
Providing easy to use primitives to build new mcps, skills and assistants (secure claw)
-
Deeply understand AI usage within an organisation from tool use, token spend, and complete agent sessions.
About the Role This is a unique, high-impact opportunity to join a passionate team about an intuitive craft to enable solving green field and hard product problems.
You’ll be working on an early product with a fast moving team that’s done this before and collaborating with founders (and customers) directly on a daily basis. Some of the ways you will have impact:
-
Identify architectural changes to improve reliability, performance and availability.
-
Foster a culture of reliability across Speakeasy's engineering organization.
-
Design and implement key operational processes such as deployments, upgrades, rollbacks, and postmortem review.
-
Join a core engineering team and participate in on-call rotation, responding to production incidents.
-
Build monitoring systems that ensure the highest quality service for our customers.
-
Debug production issues across all services and levels of the stack.
You’re a good fit if...
-
You want to join a talent-dense team made up of ex-founders and domain experts across various developer tools, languages, and infrastructure (in fact over a quarter of the team ).
-
Ownership excites you in the full lifecycle of development: building, shipping, running support, maintaining infrastructure and measuring impact.
-
You have a record of full agency in improving a product's reliability and uptime.
-
You have an exceptional ability to learn, pickup new frameworks, and can work across the stack from backend to frontend.
-
Ability to participate in on-call rotation and respond to production incidents.
-
Ability to work in person in our San Francisco Office
Similar jobs, posted recently
Open roles like this one, listed in the last 30 days.
Site Reliability EngineerMaintainX · Remote · US · $120–249K/yrPosted 3w agoPosted 3w ago
Senior Software Engineer, Backend PlatformHandshake · San Francisco, CA · $208–260K/yrPosted 2w agoPosted 2w ago
Data Center Hardware Quality & Reliability EngineerOpenAI · San Francisco, CA (Hybrid) · $226–285K/yrPosted 4 days agoPosted 4 days ago
Senior Software Engineer, ReliabilityRoblox · San Mateo, CA · $243–295K/yrPosted 2 days agoPosted 2 days ago
Senior Software Engineer, Application GatewayRoblox · San Mateo, CA · $197–243K/yrPosted 2w agoPosted 2w ago
You've read the whole posting — now see how you match it.