Software Engineer, Data
Palo Alto, CAFull-time$180–220K/yrPosted 7mo agoStill listed 2w ago
Most applications go out cold — see where you stand first. No sign-up to start.
Watch jobs like this. New roles like this one near Palo Alto, CA, by email.
Don't just apply. Show up ready.
Olive works from this exact posting.
At a glance
Olive lists jobs from US employers, including remote roles you can work from the United States.
Requirements
Credentials this posting asks for.
Job overview
HeyGen is hiring a Software Engineer, Data. HeyGen is seeking a Software Engineer with data engineering responsibilities to build foundational data layers for next-generation features. This role focuses on enabling AI models in real-time, constructing robust multimedia pipelines, and powering engaging user experiences. The team is developing cutting-edge features like PPT-to-video converters and interactive, conversational video capabilities.
Key focus areas include Design, develop, and maintain robust batch and real-time data pipelines, Ingest and transform massive multi-modal data to train and run AI models, and Collaborate with ML engineers to implement data structures and APIs.
Successful candidates bring Bachelor's Degree In Computer Science, Bachelor's Degree In Engineering, and Bachelor's Degree In Related Field. Important skills include Python, Apache Spark, Kafka, Snowflake, Databricks, and SQL. Preferred (not required): Go, Iceberg, Computer Vision, and Generative AI Data Processing.
Skills & qualifications
Skills
Qualifications
Benefits
Full job description
About HeyGen At HeyGen, our mission is to make visual storytelling accessible to all. Over the last decade, visual content has become the preferred method of information creation, consumption, and retention. But the ability to create such content, in particular videos, continues to be costly and challenging to scale. Our ambition is to build technology that equips more people with the power to reach, captivate, and inspire audiences. Learn more at www.heygen.com . Visit our Mission and Culture doc here .
Position Summary
A Software Engineer with data engineering responsibilities to bridge the gap between core application development and large-scale data infrastructure. You will help build the data foundational layers for our next-generation features. This role is not just about moving data—it’s about enabling AI models to function in real-time, building robust pipelines for multimedia, and powering engaging user experiences. This team is currently working on cutting-edge features including PPT-to-video converters and interactive, conversational video capabilities.
Core Responsibilities
-
Build & Scale Data Pipelines: Design, develop, and maintain robust batch and real-time data pipelines (using Python, Go, Spark, Kafka) that ingest and transform massive multi-modal data—text, audio, and video—to train and run AI models.
-
Power Intelligent Features: Collaborate with ML engineers to implement data structures and APIs for new, exciting features like PPT-to-video automation and interactive AI avatars that require low-latency data fetching.
-
Data Lakehouse Infrastructure: Architect and manage data lakehouse solutions (e.g., Snowflake, Databricks, Apache Iceberg) to store and query unstructured media data efficiently, enhancing storage and computation efficiency.
-
Data Reliability & Observability: Implement data quality checks, data contracts, and monitoring to ensure high reliability of data, preventing downtime in production video generation.
-
Productize Data: Transform raw data into structured, actionable data products that can be easily consumed by front-end applications, API endpoints, and AI agents. Qualifications
-
Bachelor’s/Master’s degree in Computer Science, Engineering, or a related field.
-
3-5+ years of experience as a Backend Software Engineer with heavy data processing responsibilities.
-
Strong proficiency in Python (for ETL/scripting) and SQL (for data modeling).
-
Experience with cloud platforms (AWS/GCP) and data technologies like Kafka, Spark, and Snowflake/Databricks.
-
Experience or interest in Computer Vision/Generative AI data processing.
-
Proactive, "owner" mindset; ability to operate in a fast-paced, startup environment. What HeyGen Offers
-
Competitive salary and benefits package.
-
Dynamic and inclusive work environment focused on innovation and creativity.
-
Opportunities for professional growth and skill development.
-
Collaborative culture that values teamwork and employee input.
-
Access to state-of-the-art technologies and tools. Salary Range $180,000 – $220,000 + equity + benefits Please note that the salary information is a general guideline only. HeyGen considers factors such as scope and responsibilities of the position, candidate's work experience, education/training, key skills, and internal equity, as well as location, market and business considerations when extending an offer. As part of our total rewards package, HeyGen offers comprehensive benefits including equity, a 401k plan, health benefits, generous PTO, a parental leave program and emotional health resources.
HeyGen is an Equal Opportunity Employer. We celebrate diversity and are committed to creating an inclusive environment for all employees.
Join us at HeyGen and be part of a team that's reshaping the world of video creation through innovative technology!
Similar jobs, posted recently
Open roles like this one, listed in the last 30 days.
Software Engineer - Data InfrastructureLuma AI · Redwood City, CAPosted 4w agoPosted 4w ago
Staff Software Engineer, Data, AutonomyRivian · Palo Alto, CA · $207–258K/yrPosted 4w agoPosted 4w ago
Staff Software Engineer, Data WarehouseCommure · Mountain View, CA (Hybrid) · $210–275K/yrPosted 4 days agoPosted 4 days agoStaff Software Engineer, Data OpsRivian · Palo Alto, CA · $207–258K/yrPosted 4w agoPosted 4w ago
Staff Software Engineer, Data Warehouse Foundation Pinterest · Remote · US · $177–365K/yrPosted 2w agoPosted 2w ago
You've read the whole posting — now see how you match it.