Medal logo

Site Reliability / Infrastructure Engineer

Medal

New York, NYFull-timeNo compensation foundPosted 1mo agoVerified open 5 days ago

Most applications go out cold — see where you stand first. No sign-up to start.

At a glance

Compensation
No compensation found
Location
New York, NY
Schedule
Full-time
Work Authorization
Not specified

Job overview

Medal is hiring a Site Reliability / Infrastructure Engineer. Medal's infrastructure supports billions of video clips and social features at massive scale. The role focuses on reliability, incident response, and scaling, requiring ownership of on‑call rotation, postmortem leadership, and close collaboration with engineering teams to meet infrastructure needs.

Key focus areas include Own on‑call rotation and respond to incidents, Drive postmortems and implement preventive fixes, and Collaborate with engineering teams to meet infrastructure requirements.

Successful candidates bring Experience Owning Infrastructure-As-Code At Scale, Experience Running ES For User-Facing Features, and GCP Depth. Important skills include Terraform, Elasticsearch, Kubernetes, Amazon VPC, IAM, and Cloud Logging. Preferred (not required): Electron, React, Redux, and Styled Components.

Skills & qualifications

RequiredNice to have

Skills

TerraformElasticsearchKubernetesAmazon VPCIAMCloud LoggingMySQLPostgreSQLIncident ManagementCommunicationGitHub ActionsGreat JudgmentElectronReactReduxStyled ComponentsC#C++SwiftKotlinJavaRedisRabbitMQSaltStackCircleCIInfrastructure as CodeGCPDatabase ScalingCI/CDJudgment

Qualifications

Experience at Startups

Full job description

About the Company General Intuition is the frontier lab for acting in space and time. We build large action models and world models that can perceive, predict, and act across virtual and physical environments. General Intuition builds on the strength of Medal, the world's largest and fastest-growing platform for gaming clips, where millions of gamers capture, share, and discover new games every year. We recently raised $320M at a $2.3B valuation led by Khosla Ventures with participation from General Catalyst, Eric Schmidt, and Jeff Bezos, to discover the next generation of real-world intelligence.

The Role Medal's infrastructure handles billions of clips, video ingestion pipelines, and social features at a massive scale most engineers never get to touch. The work centers on reliability, incident response, scaling, and making sure our infrastructure keeps up with our growth. You'll own the on-call rotation, drive postmortems, and work directly with engineering teams to meet their infra needs. The right person probably came through startups and scale-ups, has been in the room when things broke at 2am, has scaled databases under pressure, and knows the difference between a durable fix and a patch that buys you a week.

What We're Looking For

  • Infrastructure-as-code: Strong fluency in Terraform, with real experience owning infrastructure-as-code at scale

  • Elasticsearch depth: Hands-on experience running ES for user-facing features, not just as a log sink

  • GCP depth: You know it maybe a little too well: Kubernetes, VPC, IAM, Cloud Logging, and the managed services ecosystem

  • Database scaling: Deep, hands-on experience scaling and sharding relational databases (MySQL, Postgres) in production

  • Incident response instincts: You can work a P0 calmly, communicate clearly under pressure, and run a postmortem that prevents recurrence

  • CI/CD: You've worked with GitHub Actions in a production environment

  • Communication (crucial!): You flag issues clearly and rapidly during incidents and lead/write actionable postmortems

  • Experience at startups: You are comfortable in an environment of rapid growth where scaling up is a priority

  • Great judgment: You know the difference between a durable, sustainable fix and a patch that buys you a week

Our Stack Electron, React, Redux, Styled Components & other modern web-based technologies C# and C++ for native Windows recording & more Swift for iOS, Kotlin for Android Java, Redis, RabbitMQ, Kubernetes for backend Terraform, Salt, GitHub Actions, CircleCI for IaC and CI/CD

You've read the whole posting — now see how you match it.

Site Reliability / Infrastructure Engineer at Medal | Olive Jobs