ML Infrastructure Engineer
San Francisco, CAFull-time$190–250K/yrPosted 10mo agoStill listed 3 days ago
Most applications go out cold — see where you stand first. No sign-up to start.
Watch jobs like this. New roles like this one near San Francisco, CA, by email.
Don't just apply. Show up ready.
Olive works from this exact posting.
At a glance
Olive lists jobs from US employers, including remote roles you can work from the United States.
Job overview
Gridware is hiring a ML Infrastructure Engineer. Gridware, a San Francisco‑based technology company focused on climate‑tech, is seeking a Senior ML Infrastructure Engineer to design, build, and maintain scalable infrastructure for model deployment and monitoring. The role collaborates with data and platform teams, develops observability systems, and enhances CI/CD pipelines to ensure reliable production of machine‑learning models.
Key focus areas include Design, build, and maintain infrastructure, tooling, and workflows for reliable, scalable ML model deployment, Develop monitoring and observability systems to track model performance, data drift, data quality, and system health, and Create and maintain end‑to‑end testing frameworks and simulation environments for model validation.
Successful candidates bring 5+ Years Of Experience Building Production ML Infrastructure and Authorized To Work In The Country Of Employment. Important skills include ML Infrastructure Design, ML Infrastructure Building, ML Infrastructure Maintenance, Tooling Development, Workflow Development, and Monitoring System Development. Preferred (not required): Feature Stores, Model Registries, Centralized Metadata Systems, and MLFlow.
Skills & qualifications
Skills
Qualifications
Benefits
Full job description
About Gridware Gridware is a San Francisco-based technology company dedicated to protecting and enhancing the electrical grid. We pioneered a groundbreaking new class of grid management called active grid response (AGR), focused on monitoring the electrical, physical, and environmental aspects of the grid that affect reliability and safety. Gridware’s advanced Active Grid Response platform uses high-precision sensors to detect potential issues early, enabling proactive maintenance and fault mitigation. This comprehensive approach helps improve safety, reduce outages, and ensure the grid operates efficiently. The company is backed by climate-tech and Silicon Valley investors. For more information, please visit www.Gridware.io . Role Description As a Senior ML Infrastructure Engineer, you will work directly in the Automation org with the core ML, Ops, and Analytics teams to help improve and build out the infrastructure around model deployment and monitoring. This role is essential to helping scale out the amount of time saving’s Gridware brings to customers. Responsibilities
- Design, build, and maintain the infrastructure, tooling, and workflows that enable reliable, scalable deployment of ML models to production.
- Develop monitoring and observability systems to track model performance, data drift, data quality, and overall system health.
- Create and maintain end-to-end testing frameworks and simulation environments to validate models and pipelines prior to deployment.
- Work closely with Data Engineering and Platform Engineering teams to ensure ML systems integrate cleanly with broader Gridware infrastructure and operational standards.
- Improve CI/CD pipelines for ML workloads, ensuring reproducibility, safe rollout, and automated rollback strategies. Required Skills
- 5+ years of experience building production ML infrastructure
- Strong software engineering skills and proficiency in Python
- Experience with cloud platforms (AWS) and container orchestration (Kubernetes)
- Familiarity with feature stores, model registries, or centralized metadata systems (i.e. MLFlow) At this time, Gridware is unable to provide visa sponsorship or immigration support for this role. We’re only able to consider candidates who are currently authorized to work in the country of employment without visa sponsorship now or in the future. This describes the ideal candidate; many of us have picked up this expertise along the way. Even if you meet only part of this list, we encourage you to apply! Gridware Technologies Inc. is an equal opportunity employer. All qualified applicants will receive consideration for employment without regard to any characteristic protected by applicable federal, state, or local law. Benefits Health, Dental & Vision (Gold and Platinum with some providers plans fully covered) Paid parental leave Alternating day off (every other Monday) “Off the Grid”, a two week per year paid break for all employees. Commuter allowance Company-paid training 190000 - 250000 USD a year
Senior ML Engineer Base Salary- $190,000-$210,000 Staff ML Engineer Base Salary- $245,000-$250,000.
Similar jobs, posted recently
Open roles like this one, listed in the last 30 days.
Research Engineer - ML InfrastructureChai Discovery · San Francisco, CAPosted 2w agoPosted 2w ago
ML Infra Engineer, PlatformPhysical Intelligence · San Francisco, CAPosted 1 day agoPosted 1 day ago
Senior Software Engineer, ML Infra - Asset SafetyRoblox · San Mateo, CA · $279–329K/yrPosted 2w agoPosted 2w ago
2027 Internship Onboard Infrastructure Engineer, ML InferenceBedrock Robotics · San Francisco, CAPosted 4w agoPosted 4w ago
Infrastructure EngineerAEGIS · San Francisco, CAPosted 2w agoPosted 2w ago
You've read the whole posting — now see how you match it.