ML Engineer (Data Engine)
Most applications go out cold — see where you stand first. No sign-up to start.
Don't just apply. Show up ready.
Olive works from this exact posting — no sign-up to start.
At a glance
Job overview
Higgsfield AI is the fastest‑scaling generative AI company, generating millions of AI‑powered videos daily for Fortune 500 brands. The role offers a competitive USD salary, equity, relocation support, and a collaborative on‑site environment in Almaty where the team works together five days a week.
Skills & qualifications
Skills
Benefits
Full job description
Why work at Higgsfield AI? Higgsfield AI is the fastest-scaling generative AI company in history, hitting $500M in annual revenue run rate, 25M+ users worldwide, 6M+ generations per day, and powering 390 of Fortune 500 brands. We're building at the absolute frontier of AI-powered video creation and next-generation creative tools. Joining Higgsfield means becoming part of a high-impact team shaping the future of AI-native experiences, at a company that isn't just moving fast, but rewriting what fast looks like. What you’ll do
-
Analyze motion and physical dynamics in video, including optical flow, camera vs. object motion, temporal consistency, and physical plausibility such as dynamics, collisions, gravity, and deformation.
-
Train models that will be used in post-training, including reward models.
-
Build video classifiers for motion types, shot types, genres, and quality, using both classical approaches and VLM/embedding-based methods.
-
Develop methods for creating multi-shot video sequences: shot-boundary detection, assembling coherent multi-shot sequences with consistent characters and scenes, and generating captions for individual shots.
-
Work with audio: speech/music/sound-event detection, audio-visual synchronization evaluation (lip-sync and sound-to-action alignment), audio-quality filtering, and audio-track captioning.
-
Apply classical and learned data-curation techniques, including near-duplicate detection, heuristic filtering, VLM-based captioning, synthetic data generation, and dataset versioning.
-
Build and run large-scale distributed processing pipelines.
Requirements
-
Strong PyTorch skills and hands-on model training experience, including fine-tuning and deploying classifiers, embedding models, and VLMs.
-
Deep knowledge of video-processing methods, including motion analysis (optical flow, tracking, camera-motion estimation), shot-boundary detection, and temporal consistency and quality evaluation.
-
Experience building and validating classifiers on video data, including annotation workflows, active learning, threshold calibration, and precision/recall measurement at scale.
-
Strong understanding of data-curation techniques such as deduplication, filtering, sampling, and dataset balancing, as well as how these choices affect model training.
-
Experience processing large-scale unstructured datasets, particularly video, audio, and images.
What We Offer
-
Competitive base salary in USD, based on your experience, skills, and the scope of the role.
-
Equity participation through the company’s stock option program, giving you the opportunity to share in Higgsfield’s long-term growth.
-
Relocation support to Almaty for candidates moving from another city or country.
-
A highly collaborative, fast-paced environment where you can work directly with experienced leaders and have a meaningful impact on the product and company.
-
Opportunities for professional growth, ownership, and career development as the company scales.
-
Company-provided equipment, meals, transportation, or other office benefits.
This is a fully on-site role based in our Almaty office. Our team works from the office five days per week for the full working day. We believe in-person collaboration is an important part of how we move quickly, solve complex problems, and build strong teams.
You've read the whole posting — now see how you match it.