Intel logo

AI Frameworks Engineer - Model and Kernel Optimization

Intel

Shanghai, Shanghai, ChinaFull-timePosted 3w agoStill listed today

Most applications go out cold — see where you stand first. No sign-up to start.

Watch jobs like this.

At a glance

Compensation
No compensation found
Location
Shanghai, Shanghai, China
Schedule
Full-time
Work Authorization
Not specified

Olive lists jobs from US employers, including remote roles you can work from the United States.

Requirements

Credentials this posting asks for.

Master's degree

Job overview

Intel's AI engineering team in Shanghai seeks a passionate engineer to deliver high‑performance deep learning solutions, optimizing model performance, designing deployment architectures, developing high‑performance kernels for Intel accelerators, and collaborating with teammates and architects to solve accuracy and memory challenges.

Skills & qualifications

RequiredNice to have

Skills

C++PythonDeep Learning FundamentalsLLMsAIGCModel ArchitecturesPyTorchvLLMSGLangGPU Kernel DevelopmentProblem Solving

Qualifications

Master's or Ph.D. In Computer Science Artificial Intelligence Software Engineering or Related FieldProficient English

Full job description

Job Details:

Job Description: Artificial Intelligence (AI) is transforming our lives and is everywhere. At INTEL, we are part of this AI revolution. Our software stack integrates seamlessly into customer frameworks, which are used by millions of end users. As our AI engineering team in Shanghai continues to grow, we are looking for a passionate engineer to help us deliver high-performance and high-quality deep learning solutions to our customers. Our team's work includes:

  • Optimizing performance for key use cases/models, debugging and resolving issues related to accuracy and memory management
  • Designing and developing model deployment architectures, such as implementing new features on vLLM/SGLang to accelerate inference (e.g., P/D disaggregation)
  • Developing and debugging high-performance kernels for INTEL accelerators
  • Communicating with direct teammates, collaborators, and architects to discuss issues, propose solutions, provide status updates, and gather feedback
  • Applying innovative ideas to enhance our products

Qualifications:

  • A Master's or Ph.D. degree in Computer Science, Artificial Intelligence, Software Engineering, or a related field
  • Strong programming skills in C++ and Python
  • Solid understanding of deep learning fundamentals and hands-on experience
  • Proficient in both written and spoken English
  • Passionate about problem-solving and proactive thinking
  • Nice to have: o Experience with LLMs or AIGC and a deep understanding of model architectures o Or experience with PyTorch, vLLM, or SGLang o Or experience in GPU kernel development

Job Type: College Grad

Shift: Shift 1 (China)

Primary Location: PRC, Shanghai

Additional Locations:

Posting Statement: All qualified applicants will receive consideration for employment without regard to race, color, religion, religious creed, sex, national origin, ancestry, age, physical or mental disability, medical condition, genetic information, military and veteran status, marital status, pregnancy, gender, gender expression, gender identity, sexual orientation, or any other characteristic protected by local law, regulation, or ordinance. Position of Trust N/A Work Model for this Role This role will require an on-site presence. * Job posting details (such as work model, location or time type) are subject to change. *

ADDITIONAL INFORMATION: Intel is committed to Responsible Business Alliance (RBA) compliance and ethical hiring practices. We do not charge any fees during our hiring process. Candidates should never be required to pay recruitment fees, medical examination fees, or any other charges as a condition of employment. If you are asked to pay any fees during our hiring process, please report this immediately to your recruiter.

Similar jobs, posted recently

Open roles like this one, listed in the last 30 days.

You've read the whole posting — now see how you match it.