Site Reliability Operations Engineer
Seattle, WAPer DiemPosted 3w agoStill listed 5 days ago
Most applications go out cold — see where you stand first. No sign-up to start.
Watch jobs like this. New roles like this one near Seattle, WA, by email.
Don't just apply. Show up ready.
Olive works from this exact posting.
At a glance
Olive lists jobs from US employers, including remote roles you can work from the United States.
Requirements
Credentials this posting asks for.
Job overview
Digital Enterprise Technology (DET) at Salesforce connects people and technology to transform the future of work. The Site Reliability Operations Engineer joins the internal DET team to support global employees, combining incident command, reliability engineering, and hands‑on technical support to keep critical systems running across time zones.
Skills & qualifications
Skills
Qualifications
Benefits
Full job description
Description The Experience
Digital Enterprise Technology (DET) connects people and technology to transform the future of work at Salesforce. Guided by our core values of Trust, Customer Success, Equality, Innovation, and Sustainability, we deliver business outcomes that fuel growth, drive competitive advantage, and empower our employees and customers globally. DET's scope stretches beyond traditional IT. We are strategic partners, advocating for the best outcomes for our customers, always innovating, and helping to shape the future of work. DET oversees technology strategy, Salesforce on Salesforce, customer and partner enablement, applications engineering, infrastructure, collaboration, enterprise operations, architecture, and program enablement. DET is Customer Zero, the best example of Salesforce products delivered globally, at scale, sustainably.
As a Site Reliability Operations Engineer you'll be part of our internal DET Site Reliability Operations team supporting our employees globally. This role combines incident command, reliability engineering, and hands-on technical support. You'll help keep critical systems running while working with teams across different time zones.
What You’ll Actually Be Doing...
-
Respond to and manage major incidents affecting internal business operations. Serve as Incident Commander to coordinate technical teams, establish impact, and drive rapid service restoration.
-
Monitor and troubleshoot enterprise systems including infrastructure, applications, and network components. Use your technical skills to diagnose complex problems across multiple platforms and vendors before they impact users.
-
Work with teams globally to improve incident response by creating and improving runbooks, developing SOPs, and driving automation.
-
Coordinate emergency changes and infrastructure updates to resolve incidents. Work with cross-functional teams to maintain business continuity during critical situations.
-
Analyze incident data and KPI metrics to identify trends. Develop actionable recommendations to reduce impact duration and improve performance, then present findings to stakeholders.
-
Lead problem management activities, investigating recurring incidents, documenting root cause analyses, and tracking known errors.
-
Participate in on-call rotation as part of regional coverage. Handle escalations during your shift and serve as Duty Manager for high severity incidents when needed.
-
Track on-call burden and surface toil reduction opportunities with measurable impact.
You’re Our Person If...
-
5-8 years in IT operations, incident management, or site reliability work. Experience in a 24x7 high availability environment with enterprise systems preferred.
-
Demonstrated ability to manage high severity incidents under pressure. Establish impact, evaluate solutions with subject matter experts, and make decisions that balance technical and business needs.
-
Strong verbal and written communication skills to explain complex technical issues to both technical and executive audiences. Create clear incident updates and status reports.
-
Demonstrated technical troubleshooting ability across Windows and Linux servers, networking, cloud platforms, and virtualization technologies. Diagnose problems quickly using logs, monitoring tools, and common diagnostic approaches.
-
Experience with cloud platforms (e.g. AWS) and monitoring of IT infrastructure. You should know core cloud concepts and be comfortable with monitoring tools.
-
Understanding of ITIL framework, particularly incident, problem, and change management processes.
-
A related technical degree required.
Even Better If...
-
Salesforce platform experience and certifications
-
Industry certifications like ITIL, AWS, CCNA, MCSA, or RHCE
-
Scripting ability in Python, Bash, PowerShell, or similar languages to help automate, reduce manual work, and improve efficiency.
-
Experience with monitoring and visualization tools like Splunk, Grafana, or Tableau. Ability to analyze data and identify trends for improving reliability.
-
Background with automation tools like Puppet or Chef
In the United States, compensation offered will be determined by factors such as location, job level, job-related knowledge, skills, and experience. Certain roles may be eligible for incentive compensation, equity, and benefits. Salesforce offers a variety of benefits to help you live well including: time off programs, medical, dental, vision, mental health support, paid parental leave, life and disability insurance, 401(k), and an employee stock purchasing program. More details about company benefits can be found at the following link: https://www.salesforcebenefits.com.
Similar jobs, posted recently
Open roles like this one, listed in the last 30 days.
Site Reliability Operations EngineerSalesforce · Seattle, WA · $94–142K/yrPosted 3w agoPosted 3w ago
Senior/Lead Software Engineer, Site Reliability (Agentforce Operations)Salesforce · San Francisco, CA · $149–314K/yrPosted 2 days agoPosted 2 days ago
Senior Staff Site Reliability OperationsNVIDIA · Seattle, WA · $184–265K/yrPosted 3w agoPosted 3w ago
Senior Quality Engineer, Applied AIAnduril Industries · Seattle, WA · $146–194K/yrPosted 4 days agoPosted 4 days ago
RF Engineer II (R6053)Shield AI · Seattle, WA · $110–170K/yrPosted 1w agoPosted 1w ago
You've read the whole posting — now see how you match it.