
Sr Manager - Infrastructure, SRE, & AI Platforms - Services Special Projects
Cupertino, CA · HybridFull-timeSeen 3w agoSeen in employer's feed 1 day ago
Most applications go out cold — see where you stand first. No sign-up to start.
Watch jobs like this. New roles like this one near Cupertino, CA, by email.
Don't just apply. Show up ready.
Olive works from this exact posting.
At a glance
Olive lists jobs from US employers, including remote roles you can work from the United States.
Requirements
Credentials this posting asks for.
Job overview
Apple seeks a Senior Infrastructure, SRE & AI Platforms Manager to define long‑term technical strategy, shape organizational structure, and drive operational roadmap for global, mission‑critical platforms, balancing innovation, reliability, performance, and cost efficiency while fostering a culture of automation and accountability.
Skills & qualifications
Skills
Qualifications
Full job description
Weekly Hours: 40
Role Number: 200681287-0836
Summary
We are looking to hire a Senior Infrastructure, SRE & AI Platforms Manager to help set the long-term technical strategy, organizational structure, and operational roadmap for global, mission-critical infrastructure platforms on the Services Special Projects team.
This position requires a rare blend of deep technical domain expertise—spanning distributed systems, Kubernetes, and AI workload orchestration—and proven organizational leadership managing large, globally distributed engineering teams.
Description
In this role, you will be responsible for defining and building infrastructure strategy that balances continuous innovation with high reliability, performance, and cost efficiency. You will lead a growing, multi-tiered team of engineers who are responsible for foundational platforms that power large-scale consumer and enterprise workloads.
Beyond operational delivery, you will establish standards for operational excellence, Site Reliability Engineering (SRE), and capacity planning. You will be a key strategic partner, translating complex business imperatives into scalable platform designs while cultivating a strong engineering culture focused on automation, technical ownership, accountability, and continuous improvement.
Minimum Qualifications
-
MS Degree in Computer Science or related degree and 12+ years of experience of progressive engineering leadership experience building, scaling, and operating mission-critical infrastructure platforms and global services.
-
Management & Leadership Scope: 6+ years managing multi-layered engineering organizations (manager-of-managers) with a proven track record of hiring, developing, and retaining top-tier technical talent across global sites.
-
Cloud & Distributed Compute Expertise: Demonstrated hands-on and architectural mastery of cloud-native infrastructure, Kubernetes platform engineering, and hybrid cloud operations (AWS, GCP, private data centers).
-
Accelerated Computing & AI Infrastructure: Direct operational and architectural experience running large-scale systems for AI/ML training and inference workloads, including utilization optimization, scheduling, and high-performance storage/networking.
-
SRE & Production Operations: Deep background in Site Reliability Engineering (SRE) principles, telemetry, observability frameworks, disaster recovery, and managing 24/7 high-availability infrastructure at scale.
-
Technical Communication: Exceptional ability to seamlessly bridge executive strategy and low-level technical trade-offs—communicating vision to executive stakeholders while driving detailed technical discussions with principal engineers.
Preferred Qualifications
-
Large-Scale Enterprise Provenance: Experience leading core infrastructure or foundational platform SRE for a global, tier-1 technology organization operating at massive scale.
-
Multi-Engine Database & Data Infrastructure: Familiarity overseeing diverse open-source and proprietary storage/data ecosystems (e.g., Cassandra, FoundationDB, Kafka, Redis, PostgreSQL).
-
Financial & Capacity Governance: Proven competency managing large-scale infrastructure investments, capital expenditures, operational budgets, capacity forecasting, and cloud optimization strategies.
Apple is an equal opportunity employer that is committed to inclusion and diversity. We seek to promote equal opportunity for all applicants without regard to race, color, religion, sex, sexual orientation, gender identity, national origin, disability, Veteran status, or other legally protected characteristics.
Similar jobs, posted recently
Open roles like this one, listed in the last 30 days.
Capital Projects - Electrical Infrastructure EngineerEtched · San Jose, CAPosted 2 days agoPosted 2 days ago
- Associate Site Reliability Engineer/Site Reliability EngineerC3 AI · Redwood City, CA · $90–160K/yrPosted 3 days agoPosted 3 days ago
Senior Integrated Marketing Specialist - AI InfrastructureNVIDIA · Santa Clara, CA · $108–173K/yrPosted 4 days agoPosted 4 days ago
Technical Product Manager - AI Infra ResilienceNVIDIA · Santa Clara, CA · $208–328K/yrPosted 3w agoPosted 3w ago
Senior Technical Program Manager - EDA Chip InfrastructureNVIDIA · Santa Clara, CA · $200–322K/yrPosted 1w agoPosted 1w ago
You've read the whole posting — now see how you match it.