Site Reliability Engineer
Marina del Rey, CA · HybridJob$100–130K/yrPosted 4 days agoStill listed 2 days ago
Most applications go out cold — see where you stand first. No sign-up to start.
Watch jobs like this. New roles like this one near Marina del Rey, CA, by email.
Don't just apply. Show up ready.
Olive works from this exact posting.
At a glance
Olive lists jobs from US employers, including remote roles you can work from the United States.
Job overview
Zefr is seeking a Site Reliability Engineer to collaborate with its Engineering and Data Science teams and help maintain reliable, scalable infrastructure. The role supports cloud infrastructure, CI/CD, observability, and core SRE practices, with opportunities to grow alongside an experienced team. The engineer will help deploy and support a multi-cloud microservices architecture and contribute to automation and continuous improvement.
Skills & qualifications
Skills
Qualifications
Benefits
Full job description
What we do: Zefr is the leader in AI-powered content classifications for brands and advertisers. Zefr’s platform is purpose built for multi-modal content understanding on open platforms like YouTube, TikTok, Meta and Snap, with pre-bid activation and verification solutions. Our products safeguard media and AI investments, while maximizing performance and efficacy on those channels. Headquartered in Los Angeles with global offices across New York, Chicago, London, Toronto, Singapore, and more, Zefr is redefining what trust and transparency means for social media in the age of AI.
What you’ll do: As a Site Reliability Engineer at Zefr, you’ll join our Engineering team and grow your skills in cloud infrastructure, CI/CD, Observability, and core SRE concepts, helping deliver high-quality, reliable, and scalable solutions. You’ll work closely with the rest of Zefr’s Engineering and Data Science teams, ensuring the infrastructure behind our services is robust, efficient, and scalable. We are looking for a passionate engineer to come collaborate and grow alongside our experienced SRE team. You bring curiosity, a strong work ethic, and a passion for continuous improvement and automation.
-
Support and build systems and tools that enable other engineers to deploy and manage product features both quickly and safely.
-
Help deploy and support our multi-cloud, micro-service architecture, deployed via Github Actions, ArgoCD & Kubernetes.
-
Write and maintain Infrastructure as Code with Terraform and Terragrunt, contributing changes through pull requests and code review.
-
Build and improve CI/CD pipelines and release workflows.
-
Help maintain the health of production environments, including monitoring application performance and resource utilization.
-
Participate in 24/7 on-call rotation, responding to system performance issues and outages alongside senior teammates.
-
Debug issues at the application and infrastructure level.
-
Write and maintain clear documentation and runbooks.
-
Contribute to our DevOps culture and philosophy of continuous improvement.
Technology Stack at Zefr:
-
Cloud Providers: Google Cloud Platform (primary), Amazon Web Services
-
Infrastructure as Code (IaC): Terraform, Terragrunt
-
Containerization & Orchestration: Docker, Kubernetes (GKE)
-
CI/CD: GitHub Actions, Argo CD
-
Primary Language: Python
-
Observability: Prometheus, OpenTelemetry, Chronosphere, Pagerduty
-
Application Languages/Frameworks: Python, FastAPI, Flask, Node.js, React
-
Workflow Orchestration: Apache Airflow, Ray
-
Relational Databases: PostgreSQL
-
NoSQL Databases: DynamoDB
-
Search Databases: OpenSearch
-
Data Warehousing: Snowflake
What we’re looking for:
-
1-3 years of experience supporting Cloud Infrastructure in a production environment using AWS and/or GCP
-
Hands-on experience with containers and Kubernetes
-
Competency in Python and shell scripting
-
Strong problem-solving skills
-
Strong written and verbal communication, organization, and documentation skills
Nice to have:
-
Familiarity with modern CI/CD pipelines and GitOps (Github Actions, GitLab, Argo CD)
-
Exposure to Monitoring and Observability tools (Prometheus, Grafana, Chronosphere, Datadog, OpenTelemetry)
-
Experience collaborating across departments to execute complex, high-impact projects.
Benefits (for US based employees):
-
Flexible PTO
-
Medical, dental, and vision insurance with FSA options
-
Company-paid life insurance
-
Paid parental leave
-
401(k) with company match
-
Professional development opportunities
-
13 paid holidays off
-
Summer Fridays (we leave early)
-
In-office and hybrid work options available
-
In-office lunches and lots of free food
-
Optional in-person and virtual events (we like to celebrate!)
Compensation (for US based employees): The anticipated salary for this position is between $100,000 and $130,000. Within the range, individual pay is determined by factors such as job-related skills, experience, and relevant education or training. If your compensation expectations fall outside of this range, it may still be worth having a conversation.
Zefr is an equal opportunity employer that embraces diversity and inclusion in the workplace. We are committed to building a team that represents a variety of backgrounds, skills, and perspectives because we know this only makes us better. We strongly encourage women, persons of color, LGBTQIA+ individuals, persons with disabilities, members of ethnic minorities, foreign-born residents, and veterans to apply even if you do not meet 100% of the qualifications.
Similar jobs, posted recently
Open roles like this one, listed in the last 30 days.
Next Gen GEO OPIR Deputy Chief Engineer (DCE)RTX Corporation · el segundo, CA · $146–277K/yrPosted 5 days agoPosted 5 days ago
Senior Build Reliability EngineerHermeus · El Segundo, CA · $120–160K/yrPosted 2w agoPosted 2w ago
Spacecraft Command and Data Handling (C&DH) Engineer - Millennium Space SystemsThe Boeing Company · El Segundo, CA · $99–176K/yrPosted 1w agoPosted 1w ago
Software Test & Verification Engineer (Experienced/Senior)The Boeing Company · El Segundo, CA · $135–183K/yrPosted 1w agoPosted 1w ago
Entry Level Command and Data Handling Network Engineer (Autonomous Navigation Systems)The Boeing Company · El Segundo, CA · $81–109K/yrPosted 1w agoPosted 1w ago
You've read the whole posting — now see how you match it.