IBM logo

Data Engineer – IBM Quantum

IBM

Yorktown Heights, NYJobNo compensation foundPosted 3mo agoSeen in employer's feed 4 days ago

Most applications go out cold — see where you stand first. No sign-up to start.

At a glance

Compensation
No compensation found
Location
Yorktown Heights, NY
Work Authorization
Not specified

Job overview

IBM is hiring a Data Engineer – IBM Quantum. The Data Engineer will design and operate data pipelines for IBM Quantum, focusing on quantum hardware performance, system reliability, user workloads, and platform operations. This role involves collaborating with various teams to transform diverse technical datasets into trusted analytics assets, guiding decision-making across IBM Quantum’s roadmap. The engineer will specialize in data integration, building solutions to transfer data from operational and external environments to business intelligence.

Key focus areas include Design Data Integration Solutions, Create and implement Extract, Transform, and Load (ETL) processes, and Develop and maintain efficient ETL processes.

Important skills include Data Integration, ETL Processes, Data Flow Optimization, Data Pipeline Design, Data Pipeline Maintenance, and Data Connectors. Preferred (not required): Presto, Lakehouse Solutions And Architectures, IBM Watsonx.data, and Distributed Analytics Engines.

Skills & qualifications

RequiredNice to have

Skills

Data IntegrationETL ProcessesData Flow OptimizationData Pipeline DesignData Pipeline MaintenanceData ConnectorsETL/ELT WorkflowsData QualitySQLPostgreSQLPrestoComplex QueriesPerformance OptimizationApache AirflowOrchestration WorkflowsPythonPandasData TransformationsData ValidationsLarge-Scale Batch ProcessingKafkaIBM Event StreamsStreaming Architecture ConceptsCollaborationGit-Based Version ControlCode ReviewsSoftware Engineering Best PracticesLakehouse Solutions and ArchitecturesIBM Watsonx.dataDistributed Analytics EnginesTrinoApache SparkData Modeling TechniquesData Governance ConceptsQuantum ComputingAdvanced Hardware SystemsCryogenicsDeep-Technology Platforms

Qualifications

Experience Operating Data Pipelines in Cloud-Based or Distributed EnvironmentsExperience Working With Hardware Telemetry, Infrastructure Monitoring Data, or High-Volume Operational Datasets

Full job description

Introduction

IBM Quantum is building the world’s leading quantum computing systems, software, and cloud services. The Data Engineer in this role will design and operate the data pipelines that power insight into quantum hardware performance, system reliability, user workloads, and platform operations. You will work closely with quantum hardware, firmware, cloud, and product teams to turn diverse technical datasets into trusted analytics assets that guide decision-making across IBM Quantum’s roadmap.

Your role and responsibilities

As a seasoned Data Engineer specializing in Data Integration, you will design and build solutions to transfer data from operational and external environments to the business intelligence environment. Your expertise will ensure the seamless flow of data throughout the business intelligence solution's lifecycle. Your primary responsibilities will include:

  • Design Data Integration Solutions: Create and implement Extract, Transform, and Load (ETL) processes to facilitate data transfer between environments

  • Develop ETL Processes: Build and maintain efficient ETL processes to ensure accurate and timely data flow, adhering to best practices and industry standards.

  • Ensure Seamless Data Flow: Monitor and troubleshoot data integration issues, collaborating with stakeholders to resolve problems and optimize data flow.

  • Optimize Data Integration Solutions: Continuously evaluate and improve data integration solutions, identifying opportunities for process improvements and efficiency gains.

Required technical and professional expertise

  • Design, build, and maintain scalable, reliable data pipelines supporting analytics, operational dashboards, and hardware performance insights for IBM Quantum systems.

  • Contribute towards building IBM Quantum’s Lakehouse by implementing scalable data connectors.

  • Develop and operate ETL/ELT workflows and tooling with a focus on data quality, accuracy, timeliness, and continuous improvement.

  • Apply advanced SQL skills using PostgreSQL and Presto to support analytical workloads, including complex queries and performance tuning.

  • Build and operate orchestration workflows in Apache Airflow, including dependency management, retries, backfills, monitoring, and operational reliability.

  • Implement data transformations and validations using Python (e.g., pandas and related libraries).

  • Support large-scale batch processing for high-volume, heterogeneous datasets, including system telemetry, experiment metadata, cloud operations data, and device performance metrics.

  • Work with streaming platforms such as Apache Kafka or IBM Event Streams to consume event-driven data from distributed quantum systems and services.

  • Apply streaming architecture concepts including topics, partitions, consumer groups, and schema evolution.

  • Integrate multiple technical data sources—quantum hardware telemetry, calibration data, experiment logs, job execution data, user activity, system health metrics—into trusted analytical datasets.

  • Collaborate with quantum hardware, software, product, SRE, and analytics teams to translate requirements into robust, production-ready data solutions.

  • Use Git-based version control, contribute via code reviews, and follow industry-standard software engineering best practices.

Preferred technical and professional experience

  • Experience with Lakehouse solutions and architectures, including IBM watsonx.data

  • Experience with distributed analytics engines such as Presto/Trino, or Apache Spark

  • Familiarity with data modeling techniques for analytical and reliability engineering use cases.

  • Exposure to data governance concepts such as access control, dataset ownership, lineage, and lifecycle management.

  • Experience operating data pipelines in cloud-based or distributed environments (e.g., hybrid cloud, containerized systems).

  • Experience working with hardware telemetry, infrastructure monitoring data, or high-volume operational datasets.

  • Interest in or exposure to quantum computing, advanced hardware systems, cryogenics, or other deep-technology platforms.

IBM is committed to creating a diverse environment and is proud to be an equal-opportunity employer. All qualified applicants will receive consideration for employment without regard to race, color, religion, sex, gender, gender identity or expression, sexual orientation, national origin, caste, genetics, pregnancy, disability, neurodivergence, age, veteran status, or other characteristics. IBM is also committed to compliance with all fair employment practices regarding citizenship and immigration status.

You've read the whole posting — now see how you match it.