Alloy Enterprises logo

Senior Data Center Systems Engineer -- AI Infrastructure & Performance Optimization

Alloy Enterprises

New Freedom, PAFull-timeNo compensation foundPosted 1mo agoVerified open 5 days ago

Most applications go out cold — see where you stand first. No sign-up to start.

At a glance

Compensation
No compensation found
Location
New Freedom, PA
Schedule
Full-time
Work Authorization
Not specified

Requirements

Credentials this posting asks for.

Master's degree

Job overview

Alloy Enterprises is hiring a Senior Data Center Systems Engineer -- AI Infrastructure & Performance Optimization. The Senior Data Center Systems Engineer will join Johnson Controls' Data Center Thermal Solutions team, focusing on AI/HPC infrastructure and performance optimization. This role involves bridging computing systems with data center infrastructure to optimize workload and cooling efficiency. The ideal candidate will possess deep expertise in GPU-based AI workloads and a comprehensive understanding of data center infrastructure systems.

Key focus areas include Deploy, configure, and manage AI/HPC compute clusters running diverse workloads., Deploy and commission liquid cooling hardware for high-density AI infrastructure., and Design and execute workload management strategies to evaluate thermal and power characteristics..

Successful candidates bring Phd Or Masters Degree In Electrical Engineering Mechanical Engineering Or Computer Science and 5+ Years Professional Experience In Ai/Hpc Infrastructure Data Center Operations Or Related Technical Roles. Important skills include AI/HPC Computing Systems, Cooling Efficiency Optimization, GPU-Based AI Workloads, Data Center Infrastructure Systems, Deploy AI/HPC Compute Clusters, and Configure AI/HPC Compute Clusters.

Skills & qualifications

RequiredNice to have

Skills

AI/HPC Computing SystemsCooling Efficiency OptimizationGPU-Based AI WorkloadsData Center Infrastructure SystemsDeploy AI/HPC Compute ClustersConfigure AI/HPC Compute ClustersManage AI/HPC Compute ClustersLLM TrainingLLM InferenceVLM ApplicationsHPC SimulationsStress TestingDeploy Liquid Cooling HardwareCommission Liquid Cooling HardwareDirect-to-Chip Cold PlatesCDUsImmersion SystemsHybrid Cooling SolutionsWorkload Management StrategiesThermal Characteristics EvaluationPower Characteristics EvaluationCPU-Based AI NodesGPU-Based AI NodesSystem-Level Digital Twin SimulationsIT Telemetry IntegrationMechanical Infrastructure DataElectrical Infrastructure DataCooling Efficiency ModelingComputing Efficiency ModelingExtract Telemetry DataAnalyze Telemetry DataIT NodesGPUsCPUsThermal SystemsChillersCooling TowersPumpsPerformance OptimizationEnergy Efficiency OptimizationCooling System Performance ValidationNPI Development InitiativesDCIMBAS PlatformsHolistic Infrastructure MonitoringPredictive AnalyticsOperational OptimizationTechnical Report AuthoringIndustry White Paper AuthoringIndustry Conference Representation

Qualifications

PhD in Electrical EngineeringPhD in Mechanical EngineeringPhD in Computer ScienceMaster's Degree in Electrical EngineeringMaster's Degree in Computer Science5-7 Years Professional Experience in AI/HPC Infrastructure5-7 Years Professional Experience in Data Center Operations5-7 Years Professional Experience in Related Technical RolesHands-on Experience Deploying and Managing LLM, VLM, Training, Inference, and HPC Simulation Workloads on GPU-Based AI InfrastructurePrior Experience in Data Center EnvironmentsDegree in or Computer Science

Full job description

Senior Data Center Systems Engineer -- AI Infrastructure & Performance Optimization

Job Description

Position Summary

Johnson Controls' Data Center Thermal Solutions team is seeking a Senior Data Center Systems Engineer for the Advanced Technology & Innovation group. This role bridges AI/HPC computing systems with data center infrastructure, focusing on optimizing the interdependencies between computing workloads and cooling efficiency. The ideal candidate will be highly self-driven, able to work independently with limited supervision, and bring deep expertise in GPU-based AI workloads combined with comprehensive understanding of data center infrastructure systems.

Key Responsibilities

  • Deploy, configure, and manage AI/HPC compute clusters running diverse workloads including LLM training and inference, VLM applications, HPC simulations, and stress testing scenarios.

  • Deploy and commission liquid cooling hardware including direct-to-chip cold plates, CDUs, immersion systems, and hybrid cooling solutions for high-density AI infrastructure.

  • Design and execute workload management strategies to evaluate thermal and power characteristics across CPU and GPU-based AI nodes under varied operational conditions.

  • Develop system-level digital twin simulations integrating IT telemetry with mechanical and electrical infrastructure data to model cooling and computing efficiency interdependencies.

  • Extract and analyze telemetry data from IT nodes, GPUs, CPUs, and thermal systems including chillers, cooling towers, pumps, and CDUs to optimize performance and energy efficiency.

  • Collaborate with thermal engineers and controls teams to validate cooling system performance under real-world AI/HPC workloads and support NPI development initiatives.

  • Integrate and leverage DCIM and BAS platforms for holistic infrastructure monitoring, predictive analytics, and operational optimization.

  • Author technical reports and industry white papers documenting research findings, best practices, and performance benchmarks for AI data center infrastructure.

  • Represent Johnson Controls at industry conferences, technical forums, and customer engagements to share innovations and establish thought leadership in AI infrastructure optimization.

  • Interface with hyperscalers, OEMs, and technology partners to align infrastructure solutions with evolving AI hardware and workload requirements.

  • Contribute to technical publications, patent filings, and industry standards related to data center systems optimization. Qualifications

  • PhD or Master's degree in Electrical Engineering, Mechanical Engineering, or Computer Science .

  • 5-7 years of professional experience in AI/HPC infrastructure, data center operations, or related technical roles.

  • Prior experience working in data center environments strongly preferred .

  • Proven hands-on experience deploying and managing LLM, VLM, training, inference, and HPC simulation workloads on GPU-based AI infrastructure.

  • Deep understanding of AI hardware architectures (NVIDIA H100/H200, AMD MI300, etc.) and their thermal/power characteristics across diverse workloads.

  • Strong knowledge of data center mechanical systems (cooling distribution, heat rejection), electrical systems (power distribution, UPS), and networking architectures (InfiniBand, Ethernet fabrics).

  • Proficiency with DCIM platforms, BAS integration, and digital twin modeling tools for infrastructure simulation and optimization.

  • Experience with stress testing tools, workload orchestration platforms (Kubernetes, Slurm, etc.), and telemetry collection frameworks.

  • Demonstrated ability to work independently , manage complex projects, and deliver results with minimal supervision.

  • Track record of innovation and cross-functional collaboration in data center or AI infrastructure environments.

  • Strong technical communication and presentation skills for conference participation and white paper authorship.

Johnson Controls International plc. is an equal employment opportunity and affirmative action employer and all qualified applicants will receive consideration for employment without regard to race, color, religion, sex, national origin, age, protected veteran status, genetic information, sexual orientation, gender identity, status as a qualified individual with a disability or any other characteristic protected by law. To view more information about your equal opportunity and non-discrimination rights as a candidate, visit EEO is the Law . If you are an individual with a disability and you require an accommodation during the application process, please visit here .

You've read the whole posting — now see how you match it.

Senior Data Center Systems Engineer -- AI Infrastructure… | Olive Jobs