Insight Global logo

Principal Engineer - Resiliency

Insight Global

Austin, TX · HybridJob$180–210K/yrSeen 1 day agoSeen in employer's feed 1 day ago

Most applications go out cold — see where you stand first. No sign-up to start.

Watch jobs like this.

At a glance

Compensation
$180–210K/yr
Location
Austin, TXHybrid
Work Authorization
Not specified

Olive lists jobs from US employers, including remote roles you can work from the United States.

Requirements

Credentials this posting asks for.

AWS Solutions Architect Professional Certification

Job overview

The Principal Engineer will lead resiliency strategy and reference architecture for a large-scale AWS environment, establishing availability standards, driving high‑availability designs, and overseeing business continuity and disaster recovery initiatives across infrastructure, containers, and data platforms.

Skills & qualifications

RequiredNice to have

Skills

AWSMulti‑Account AWSMulti‑Region AWSAWS NetworkingAWS IAMRoute 53Load BalancingGlobal AcceleratorMulti‑AZ ArchitectureBC/DR StrategyAutomated FailoverRTO/RPORecovery TestingAWS Well‑Architected FrameworkEKSECS/FargateAuroraRDSDynamoDBElastiCacheS3 ReplicationTerraformInfrastructure as CodeCI/CDState ManagementGuardrailsDrift DetectionChaos EngineeringObservabilitySLOsError BudgetsIncident ResponseCross‑Functional LeadershipAWS Resiliency ToolsInfrastructure Orchestration PlatformsSRE PracticesAWS Solutions Architect Professional Certification

Qualifications

12+ Years Engineering Experience7+ Years Resiliency Architecture ExperienceAWS Solutions Architect Professional Certification

Full job description

Job Description

Seeking a Principal Engineer to own the resiliency strategy and reference architecture for a large-scale AWS environment. This individual will establish availability standards, lead business continuity and disaster recovery initiatives, and drive high-availability architecture across infrastructure, containers, data platforms, and engineering teams.

Key Responsibilities

  • Own AWS resiliency strategy, availability standards, and multi-AZ/multi-region architecture.

  • Define appropriate active-active, active-passive, warm-standby, and pilot-light patterns.

  • Lead BC/DR planning, including RTO/RPO targets, automated failover, runbooks, DR testing, and game days. Conduct AWS Well-Architected Reviews and drive remediation efforts.

  • Build resiliency across ECS/Fargate, EKS, Aurora, RDS, - DynamoDB, ElastiCache, and S3.

  • Automate recovery using IaC, self-healing, drift detection, and fault injection.

  • Establish SLOs, error budgets, health checks, dependency mapping, and observability standards.

  • Lead availability incident response and convert recurring failures into architectural improvements.

  • Partner with engineering, product, finance, and business leaders to balance reliability, cost, and operational complexity.

This is a hybrid position in Austin Texas and pays between $180,000 and $210,000 per year.

Skills and Requirements

12+ years of engineering experience, including 7+ years owning large-scale resiliency or high-availability architecture. Deep hands-on AWS experience across multi-account, multi-region environments, networking, IAM, Route 53, load balancing, and Global Accelerator. Proven ownership of multi-AZ/multi-region architecture, BC/DR strategy, automated failover, RTO/RPO, and recovery testing. Expertise in the AWS Well-Architected Framework and production container platforms, including EKS and ECS/Fargate. Strong data resiliency experience with Aurora/RDS and at least one of DynamoDB, ElastiCache, or S3 replication. Advanced Terraform/IaC and infrastructure CI/CD experience, including state management, guardrails, and drift detection. Hands-on experience with chaos engineering, observability, SLOs, error budgets, and incident response. Strong cross-functional leadership with the ability to balance availability, cost, risk, and operational complexity. Experience with AWS resiliency tools, infrastructure orchestration platforms, and SRE practices. AWS Solutions Architect Professional certification or equivalent expertise. Experience modeling redundancy costs and presenting tradeoffs to business stakeholders. Background integrating AWS environments after a merger or acquisition. Experience supporting regulated or mission-critical environments.

We are a company committed to creating diverse and inclusive environments where people can bring their full, authentic selves to work every day. We are an equal employment opportunity/affirmative action employer that believes everyone matters. Qualified candidates will receive consideration for employment without regard to race, color, ethnicity, religion, sex (including pregnancy), sexual orientation, gender identity and expression, marital status, national origin, ancestry, genetic factors, age, disability, protected veteran status, military or uniformed service member status, or any other status or characteristic protected by applicable laws, regulations, and ordinances. If you need assistance and/or a reasonable accommodation due to a disability during the application or the recruiting process, please send a request to [email protected].

Similar jobs, posted recently

Open roles like this one, listed in the last 30 days.

You've read the whole posting — now see how you match it.