The Hartford logo

Staff Infrastructure Engineer

The Hartford

Hartford, CT · HybridJob$117–175K/yrSeen 1w agoSeen in employer's feed 1w ago

Most applications go out cold — see where you stand first. No sign-up to start.

Watch jobs like this.

At a glance

Compensation
$117–175K/yr
Location
Hartford, CTHybrid
Work Authorization
US work authorization required

Olive lists jobs from US employers, including remote roles you can work from the United States.

Job overview

The Staff Infrastructure Engineer leverages agentic AI to identify, assess, and resolve critical service issues and major incidents, partnering with cross‑functional teams to restore services and improve event management within the Technology Command Center of an insurance company.

Skills & qualifications

RequiredNice to have

Skills

SplunkDynatraceITSIMoogsoftThousandEyesAI‑Driven Data AnalysisPrompt EngineeringInterrogation SkillsITSMEvent CorrelationAlert PrioritizationService Impact AnalysisBasic Automation/Workflow EnablementStrong Analytical ThinkingOperational JudgmentAbility to Work Under Pressure

Qualifications

8+ Years Experience in Monitoring and Observability ToolsAuthorization to Work in US Without Sponsorship

Full job description

Staff Infrastructure Engineer - IEB07CE

We’re determined to make a difference and are proud to be an insurance company that goes well beyond coverages and policies. Working here means having every opportunity to achieve your goals – and to help others accomplish theirs, too. Join our team as we help shape the future.

The Staff Infrastructure Engineer is responsible for leveraging agentic AI to quickly identify, assess, and resolve critical service issues and major incidents, ensuring swift detection and response. The successful candidate will demonstrate engineering-level technical depth in AI, cloud computing concepts, and infrastructure platforms, partnering with cross-functional teams for service restoration and continuously improving the event management process. This position is critical in enhancing enterprise resilience and implementing proactive AI-driven operations within the Technology Command Center.

This role will have a Hybrid work schedule, with the expectation of working in an office (Hartford, CT or Charlotte, NC) 3 days a week (Tuesday through Thursday).

Key Responsibilities

  • Manage technical and executive communication for major incidents.

  • Perform alert triage, correlation, and initial impact assessment to separate actionable events from non-actionable noise.

  • Support major incident detection and escalation by validating symptoms, confirming affected services, and engaging resolver teams.

  • Use standard operating procedures, runbooks, and decision frameworks to investigate, prioritize, and escalate events.

  • Maintain situational awareness during active events; document event patterns and operational observations.

  • Promote events to incidents when thresholds are met; partner with internal and external teams during restoration activities.

  • Identify opportunities to improve monitoring effectiveness, event quality, automation, and operational readiness.

Partnership & Work Expectations

  • Collaborate with infrastructure, application, reliability engineering, service desk, and vendor teams to ensure effective event response and service restoration.

  • Support continuous improvement by identifying recurring alert issues, documenting operational insights, and recommending enhancements to monitoring, runbooks, and automation.

Schedule & Role Expectations

  • May require shift coverage, off-hours support, and escalation activities aligned to a 24x7 operational model and follow-the-sun support structure.

Required Qualifications

  • 8+ years' experience in monitoring and observability tools such as Splunk, Dynatrace, ITSI, Moogsoft, ThousandEyes, or similar.

  • Experience in AI-driven data analysis and problem solving.

  • AI tool proficiency including prompt engineering and Interrogation skills

  • Familiarity with ITSM and ticketing platforms, including incident creation, categorization, escalation, and documentation.

  • Understanding of event correlation, alert prioritization, service impact analysis, and basic automation/workflow enablement.

  • Strong analytical thinking and operational judgment.

  • Ability to work under pressure during high-impact events.

Candidate must be authorized to work in the US without company sponsorship. The company will not support the STEM OPT I-983 Training Plan endorsement for this position.

Compensation

The listed annualized base pay range is primarily based on analysis of similar positions in the external market. Actual base pay could vary and may be above or below the listed range based on factors including but not limited to performance, proficiency and demonstration of competencies required for the role. The base pay is just one component of The Hartford’s total compensation package for employees. Other rewards may include short-term or annual bonuses, long-term incentives, and on-the-spot recognition. The annualized base pay range for this role is:

$116,800 - $175,200

Equal Opportunity Employer/Sex/Race/Color/Veterans/Disability/Sexual Orientation/Gender Identity or Expression/Religion/Age

Similar jobs, posted recently

Open roles like this one, listed in the last 30 days.

You've read the whole posting — now see how you match it.