Software Engineer, Systems ML - Compilers / Kernels
Menlo Park, CAJob$122–181K/yrSeen 1 day agoSeen in employer's feed 1 day ago
Most applications go out cold — see where you stand first. No sign-up to start.
Watch jobs like this. New roles like this one near Menlo Park, CA, by email.
Don't just apply. Show up ready.
Olive works from this exact posting.
At a glance
Olive lists jobs from US employers, including remote roles you can work from the United States.
Job overview
The Software Engineer will work on Meta’s MTIA Software team within the PyTorch AI framework organization. The role focuses on developing compiler, framework, runtime, or kernel software for machine-learning workloads on specialized hardware. The engineer will analyze deep-learning models with AI researchers and partner with hardware design teams to improve performance and support training and inference on current and next-generation MTIA platforms.
Skills & qualifications
Skills
Qualifications
Full job description
Summary:
In this role, you will be a member of the MTIA (Meta Training & Inference Accelerator) Software team and part of the bigger PyTorch AI framework organization. MTIA Software Team has been developing a comprehensive AI Compiler strategy that delivers a highly flexible platform to train & serve new DL/ML model architectures, combined with auto-tuned high performance for production environments across specialized hardware architectures. The compiler stack, DL graph optimizations, and kernel authoring for specific hardware, directly impacts performance and deployment velocity of both AI training and inference platforms at Meta.You will be working on one of the core areas such as PyTorch framework components, AI compiler and runtime, high-performance kernels and tooling to accelerate machine learning workloads on the current & next generation of MTIA AI hardware platforms. You will work closely with AI researchers to analyze deep learning models and lower them efficiently on MTIA hardware. You will also partner with hardware design teams to develop compiler optimizations for high performance. You will apply software development best practices to design features, optimization, and performance tuning techniques. You will gain valuable experience in developing machine learning compiler frameworks and will help in driving next generation hardware software codesign for AI domain specific problems.
Required Skills:
Software Engineer, Systems ML - Compilers / Kernels Responsibilities:
-
Development of the SW stack with one of the following core focus areas: AI compiler stack, frameworks, high-performance kernel development and acceleration onto the next generation of hardware architectures
-
Contribute to the development of the PyTorch AI framework core compilers to support new state of the art inference and training AI hardware accelerators and optimize their performance
-
Analyze deep learning networks, develop & implement compiler optimization algorithms
-
Collaborating with AI research scientists to accelerate the next generation of deep learning models such as Recommendation systems, Generative AI, Computer vision, NLP etc
-
Performance tuning and optimizations of deep learning framework & software components
Minimum Qualifications:
Minimum Qualifications:
-
Currently has, or is in the process of obtaining a Bachelor's degree in Computer Science, Computer Engineering, relevant technical field, or equivalent practical experience. Degree must be completed prior to joining Meta
-
Proven C/C++ programming skills
-
Experience in AI framework development or accelerating deep learning models on hardware architectures
-
Bachelor's degree in Computer Science, Computer Engineering, relevant technical field, or equivalent practical experience
Preferred Qualifications:
Preferred Qualifications:
-
Experience working with frameworks like PyTorch, Caffe2, TensorFlow, ONNX, TensorRT
-
OR AI frameworks: Experience in developing training and inference framework components. Experience in system performance optimizations such as runtime analysis of latency, memory bandwidth, I/O access, compute utilization analysis and associated tooling development
-
Experience adhering to and implementing responsible, ethical AI practices (e.g., risk assessment, bias mitigation, quality and accuracy reviews)
-
A Bachelor's degree in Computer Science, Computer Engineering, relevant technical field and 7+ years of experience in AI framework development or accelerating deep learning models on hardware architectures OR a Master's degree in Computer Science, Computer Engineering, relevant technical field and 4+ years of experience in AI framework development or accelerating deep learning models on hardware architectures OR a PhD in Computer Science, Computer Engineering, or relevant technical field and 3+ years of experience in AI framework development or accelerating deep learning models on hardware architectures
-
Knowledge of GPU, CPU, or AI hardware accelerator architectures
-
OR AI Compiler: Experience with compiler optimizations such as loop optimizations, vectorization, parallelization, hardware specific optimizations such as SIMD. Experience with MLIR, LLVM, IREE, XLA, TVM, Halide is a plus
-
Demonstrated ability to integrate AI tools to optimize/redesign workflows and drive measurable impact (e.g., efficiency gains, quality improvements)
-
Demonstrated ongoing AI skill development (e.g., prompt/context engineering, agent orchestration) and staying current with emerging AI technologies
-
OR AI high performance kernels: Experience with CUDA programming, OpenMP / OpenCL programming or AI hardware accelerator kernel programming. Experience in accelerating libraries on AI hardware, similar to cuBLAS, cuDNN, CUTLASS, HIP, ROCm etc
Public Compensation:
$121,992/year to $181,000/year + bonus + equity + benefits
Industry: Internet
Equal Opportunity:
Meta is proud to be an Equal Employment Opportunity and Affirmative Action employer. We do not discriminate based upon race, religion, color, national origin, sex (including pregnancy, childbirth, or related medical conditions), sexual orientation, gender, gender identity, gender expression, transgender status, sexual stereotypes, age, status as a protected veteran, status as an individual with a disability, or other applicable legally protected characteristics. We also consider qualified applicants with criminal histories, consistent with applicable federal, state and local law. Meta participates in the E-Verify program in certain locations, as required by law. Please note that Meta may leverage artificial intelligence and machine learning technologies in connection with applications for employment.
Meta is committed to providing reasonable accommodations for candidates with disabilities in our recruiting process. If you need any assistance or accommodations due to a disability, please let us know at [email protected].
Similar jobs, posted recently
Open roles like this one, listed in the last 30 days.
ML Runtime and Kernel Engineer - Core MLCerebras · Sunnyvale, CA (Hybrid)Posted 1w agoPosted 1w ago
Machine Learning Engineer, AI Inference Solutions (Early in Career)General Motors · Sunnyvale, CA (Hybrid) · $119–151K/yrPosted 6 days agoPosted 6 days ago
Software Engineer, GPU KernelsRiver AI · Palo Alto, CA · $200–420K/yrPosted 4w agoPosted 4w ago
Senior ML Accelerator Engineer - GPUGeneral Motors · Sunnyvale, CA (Hybrid) · $170–258K/yrPosted 1w agoPosted 1w ago
Software Engineer, Inference SystemsRiver AI · Palo Alto, CA · $200–420K/yrPosted 4w agoPosted 4w ago
You've read the whole posting — now see how you match it.