Skip to content
NEXTMOVEFDE careers · United States

Software Engineer, Data

AI summary of the role

Software Engineer bridging application development and large-scale data infrastructure to enable real-time AI features like PPT-to-video and interactive avatars.

What you’ll do

  • Design, develop, and maintain robust batch and real-time data pipelines (Python, Go, Spark, Kafka) for multi-modal data (text, audio, video).
  • Collaborate with ML engineers to implement data structures and APIs for PPT-to-video automation and interactive AI avatars.
  • Architect and manage data lakehouse solutions (Snowflake, Databricks, Apache Iceberg) for unstructured media data.
  • Implement data quality checks, data contracts, and monitoring to ensure high reliability of data.

What you’ll bring

  • 3-5+ years as a Backend Software Engineer with heavy data processing responsibilities.
  • Strong proficiency in Python (ETL/scripting) and SQL (data modeling).
  • Experience with cloud platforms (AWS/GCP) and data technologies (Kafka, Spark, Snowflake/Databricks).
  • Experience or interest in Computer Vision/Generative AI data processing.

Technologies

Python · Go · Spark · Kafka · Snowflake · Databricks · Apache Iceberg · AWS · GCP · SQL

About HeyGen

AI video platform turning text and a single photo into lifelike avatar videos in 175+ languages, sold to creators and enterprises.

Series A

Source and classification

Internal deployment & tooling · Evidence for this classification:

About HeyGen At HeyGen, our mission is to make visual storytelling accessible to all. Over the last decade, visual content has become the preferred method of information creation, consumption, and retention. But the ability to create such content, in particular videos, continues to be costly and challenging to scale. Our ambition is to build technology that equips more people with the power to reach, captivate, and inspire audiences. Learn more at www.heygen.com. Visit our Mission and Culture doc here. Position Summary A Software Engineer with data engineering responsibilities to bridge the gap between core application development and large-scale data infrastructure. You will help build the data foundational layers for our next-generation features. This role is not just about moving data—it’s about enabling AI models to function in real-time, building robust pipelines for multimedia,
More from the job description

About HeyGen At HeyGen, our mission is to make visual storytelling accessible to all. Over the last decade, visual content has become the preferred method of information creation, consumption, and retention. But the ability to create such content, in particular videos, continues to be costly and challenging to scale. Our ambition is to build technology that equips more people with the power to reach, captivate, and inspire audiences. Learn more at www.heygen.com. Visit our Mission and Culture doc here. Position Summary A Software Engineer with data engineering responsibilities to bridge the gap between core application development and large-scale data infrastructure. You will help build the data foundational layers for our next-generation features. This role is not just about moving data—it’s about enabling AI models to function in real-time, building robust pipelines for multimedia, and powering engaging user experiences. This team is currently working on cutting-edge features including PPT-to-video converters and interactive, conversational video capabilities. Core Responsibilities Build & Scale Data Pipelines: Design, develop, and maintain robust batch and real-time data pipelines (using Python, Go, Spark, Kafka) that ingest and transform massive multi-modal data—text, audio, and video—to train and run AI models. Power Intelligent Features: Collaborate with ML engineer [... source excerpt omitted ...] computation efficiency. Data Reliability & Observability: Implement data quality checks, data contracts, and monitoring to ensure high reliability of data, preventing downtime in production video generation. Productize Data: Transform raw data into structured, actionable data products that can be easily consumed by front-end applications, API endpoints, and AI agents. Qualifications Bachelor’s/Master’s degree in Computer Science, Engineering, or a related field. 3-5+ years of experience as a Backend Software Engineer with heavy data processing responsibilities. Strong proficiency in Python (for ETL/scripting) and SQL (for data modeling). Experience with cloud platforms (AWS/ [... source excerpt omitted ...] es and tools. Salary Range $180,000 – $220,000 + equity + benefits Please note that the salary information is a general guideline only. HeyGen considers factors such as scope and responsibilities of the position, candidate's work experience, education/training, key skills, and internal equity, as well as location, market and business considerations when extending an offer. As part of our total rewards package, HeyGen offers comprehensive benefits including equity, a 401k plan, health benefits, generous PTO, a parental leave program and emotional health resources. HeyGen is an Equal Opportunity Employer. We celebrate diversity and are committed to creating an inclusive environment fo

How jobs are selected

Employer postings · Data from · Sources