Amazon logo

SDE, Neuron Inference, Neuron Inference

Amazon

Seattle, WAInternship$144–194K/yrSeen 3w agoSeen in employer's feed 3 days ago

Most applications go out cold — see where you stand first. No sign-up to start.

Watch jobs like this.

At a glance

Compensation
$144–194K/yr
Location
Seattle, WA
Schedule
Internship
Work Authorization
Not specified

Olive lists jobs from US employers, including remote roles you can work from the United States.

Requirements

Credentials this posting asks for.

Bachelor's degree

Job overview

Amazon seeks a senior software engineer to develop and optimize core building blocks of large language model inference such as attention, MLP, quantization, speculative decoding, and mixture of experts on AWS Neuron devices, collaborating with chip architects, compiler and runtime engineers.

Skills & qualifications

RequiredNice to have

Skills

C#C++JavaPerlObject Oriented Design

Qualifications

3+ Years of Non-Internship Professional Software Development Experience2+ Years of Non-Internship Design or Architecture Experience1+ Years of Software Development Engineer or Related Occupational Experience1+ Years of Designing and Developing Large-Scale Multi-Tiered Multi-Threaded Embedded or Distributed Software Applications Using C# C++ Java or Perl1+ Years of Object Oriented Design ExperienceBachelors Degree or Foreign Equivalent in Computer Science Engineering Mathematics or Related FieldExperience Programming With at Least One Software Programming Language3+ Years of Full Software Development Life Cycle Including Coding Standards Code Reviews Source Control Management Build Processes Testing and Operations ExperienceBachelors Degree in Computer Science or Equivalent

Benefits

Medical Insurance

Full job description

Description

AWS Neuron is the complete software stack for the AWS Inferentia and Trainium cloud-scale machine

learning accelerators. This role is for a senior software engineer in the Machine Learning Inference Applications team. This role is responsible for development and performance optimization of core building blocks of LLM Inference - Attention, MLP, Quantization, Speculative Decoding, Mixture of Experts, etc.

The team works side by side with chip architects, compiler engineers and runtime engineers to deliver performance and accuracy on Neuron devices across a range of models.

Key job responsibilities

Responsibilities of this role include adapting latest research in LLM optimization to Neuron chips to extract best performance from both open source as well as internally developed models. Working across teams and organizations is key.

About the team

Our team is dedicated to supporting new members. We have a broad mix of experience levels and tenures, and we’re building an environment that celebrates knowledge-sharing and mentorship. Our senior members enjoy one-on-one mentoring and thorough, but kind, code reviews. We care about your career growth and strive to assign projects that help our team members develop your engineering expertise so you feel empowered to take on more complex tasks in the future.

Basic Qualifications

  • 3+ years of non-internship professional software development experience

  • 2+ years of non-internship design or architecture (design patterns, reliability and scaling) of new and existing systems experience

  • 1+ years of software development engineer or related occupational experience

  • 1+ years of designing and developing large-scale, multi-tiered, multi-threaded, embedded or distributed software applications, tools, systems, and services using: C#, C++, Java, or Perl experience

  • 1+ years of Object Oriented Design experience

  • Bachelor's degree or foreign equivalent in Computer Science, Engineering, Mathematics, or a related field

  • Experience programming with at least one software programming language

Preferred Qualifications

  • 3+ years of full software development life cycle, including coding standards, code reviews, source control management, build processes, testing, and operations experience

  • Bachelor's degree in computer science or equivalent

Amazon is an equal opportunity employer and does not discriminate on the basis of protected veteran status, disability, or other legally protected status.

Our inclusive culture empowers Amazonians to deliver the best results for our customers. If you have a disability and need a workplace accommodation or adjustment during the application and hiring process, including support for the interview or onboarding process, please visit https://amazon.jobs/content/en/how-we-hire/accommodations for more information. If the country/region you’re applying in isn’t listed, please contact your Recruiting Partner.

The base salary range for this position is listed below. Your Amazon package will include sign-on payments and restricted stock units (RSUs). Final compensation will be determined based on factors including experience, qualifications, and location. Amazon also offers comprehensive benefits including health insurance (medical, dental, vision, prescription, Basic Life & AD&D insurance and option for Supplemental life plans, EAP, Mental Health Support, Medical Advice Line, Flexible Spending Accounts, Adoption and Surrogacy Reimbursement coverage), 401(k) matching, paid time off, and parental leave. Learn more about our benefits at https://amazon.jobs/en/benefits .

USA, WA, Seattle - 143,700.00 - 194,400.00 USD annually

Similar jobs, posted recently

Open roles like this one, listed in the last 30 days.

You've read the whole posting — now see how you match it.