Maximus logo

Site Reliability Engineer (Systems Manager)

Maximus

Colorado Springs, COContract / Other$130–150K/yrSeen todaySeen in employer's feed today

Most applications go out cold — see where you stand first. No sign-up to start.

Watch jobs like this.

At a glance

Compensation
$130–150K/yr
Location
Colorado Springs, CO
Schedule
Contract / Other
Work Authorization
US work authorization required

Olive lists jobs from US employers, including remote roles you can work from the United States.

Requirements

Credentials this posting asks for.

Secret clearance

Job overview

Maximus seeks a Systems Manager to ensure availability, reliability, and performance of critical federal systems, designing, implementing, and managing scalable, highly‑available infrastructure while collaborating with cross‑functional teams and supporting 24x7 operations.

Skills & qualifications

RequiredNice to have

Skills

TerraformChefAnsibleJenkinsGitLabMonitoring ToolsBackup and Disaster RecoveryCapacity PlanningIncident Management

Qualifications

Active Secret ClearanceU.S. Citizenship15+ Years Relevant ExperienceResidency Within Commutable DistanceWillingness to Participate in Rotational on‑Call Schedule

Full job description

Maximus is seeking a Systems Manager to provide expertise to a federal client in support of their mission critical systems in defense of our Homeland.

The Systems Manager will ensure the availability, reliability, and performance of critical systems and applications, utilizing a diverse set of relevant technologies, for a federal client's operations. As a crucial member of our team, the Systems Manager will play a pivotal role in designing, implementing, and managing scalable and highly available systems.

This is an on-site position based in Colorado Springs, CO requiring an active Secret clearance.

Maximus TCS (Technology and Consulting Services) Internal Job Profile Code: TCS137, T5, Band 8

Job-Specific Essential Duties and Responsibilities:

  • Collaborate with cross-functional teams to design, implement, and maintain highly available, fault-tolerant systems, leveraging a range of technologies.

  • Monitor and analyze system performance, utilizing a diverse set of monitoring and alerting tools, and take proactive actions to prevent and resolve incidents.

  • Conduct system capacity planning and scaling, utilizing performance testing and optimization techniques across various technologies.

  • Implement and maintain robust backup and disaster recovery solutions, leveraging relevant technologies, to ensure data integrity and system availability.

  • Collaborate with development teams to ensure application reliability, performance, and scalability in production environments.

  • Stay updated with industry best practices and emerging trends in site reliability engineering and incorporate relevant technologies. Motivate the team to adhere to IT best practices and deliver outstanding customer service and satisfaction.

  • Assist in creating and maintaining a training program to increase business, customer service, and technical knowledge.

  • Participate in the organization’s change management process.

  • Maintain detailed records of incidents, including reporting, classification, and resolution steps for daily and weekly metric reporting.

  • Gather and present PMR metric reporting to management, contract leadership, and government stakeholders.

  • Use monitoring tools to proactively identify and address potential issues and generate reports on incident trends.

  • Collaborate with the Operations Support Center Lead to ensure continuity of daily operations.

  • Other duties as assigned.

Job-Specific Minimum Requirements:

  • Active Secret clearance is required.

  • Due to contract requirements, only U.S. citizens can be considered.

  • 15+ years of relevant experience supporting IT platforms, environments, and infrastructure is required.

  • Experience in some or all the following: Infrastructure as Code (IaC) tools (Terraform, Chef, Ansible), Jenkins, GitLab

  • Candidates must reside within a commutable distance for daily onsite work and on-call requirements.

  • This contract supports systems that require 24x7x365 uptime. Candidates must be willing and able to meet recall requirements, including participation in a rotational on-call schedule.

#techjobs #clearance

Minimum Requirements

TCS137, T5, Band 8

Maximus is an equal opportunity employer. We evaluate qualified applicants without regard to race, color, religion, sex, age, national origin, disability, veteran status, genetic information and other legally protected characteristics.

Minimum Salary

$130,000

Maximum Salary

$150,000

Similar jobs, posted recently

Open roles like this one, listed in the last 30 days.

You've read the whole posting — now see how you match it.