Senior Software Development Engineer, AWS Neuron Inference
Senior Software Development Engineer, AWS Neuron Inference
Job ID : 2693077 Amazon.com Services LLC
All potential applicants are encouraged to scroll through and read the complete job description before applying.
AWS Neuron is the complete software stack for the AWS Inferentia and Trainium cloud-scale machine learning accelerators and the Trn1 and Inf2 servers that use them.
This role is for a software engineer in the Neuron Inference (ML Apps) team for AWS Neuron.
This role is responsible for development, enablement and performance optimization (latency, throughput) of a wide variety of ML model families, including massive scale large language models like Llama3, DBRX, and beyond, as well as Stable Diffusion, Vision Transformers and many more.
The Neuron Inference team works side by side with compiler engineers and runtime engineers to create, build and tune distributed inference solutions with Trn1 / Inf2.
Experience optimizing inference performance for both latency and throughput on these large models using PyTorch or JAX is a must.
Experience with technologies / tools such as vLLM, Hugging Face, multi-modal inference, etc. is highly valued.
Key job responsibilities
This role will help lead the efforts building distributed inference support into Pytorch, Tensorflow using XLA and the Neuron compiler and runtime stacks.
This role will help tune these models to ensure highest performance and maximize the efficiency of them running on the AWS Trainium and Inferentia silicon.
Strong software development using C++ / Python and ML knowledge are both critical to this role.
A day in the life :
- As you design and code solutions to help our team drive efficiencies in software architecture, you’ll create metrics, implement automation and other improvements, and resolve the root cause of software defects.
- Build high-impact solutions to deliver to our large customer base.
- Participate in design discussions, code review, and communicate with internal and external stakeholders.
- Work cross-functionally to help drive business decisions with your technical input.
- Work in a startup-like development environment, where you’re always working on the most important stuff.
BASIC QUALIFICATIONS
- 5+ years of non-internship professional software development experience
- 5+ years of programming with at least one software programming language experience
- 5+ years of leading design or architecture (design patterns, reliability and scaling) of new and existing systems experience
- 5+ years of full software development life cycle, including coding standards, code reviews, source control management, build processes, testing, and operations experience
- Experience as a mentor, tech lead or leading an engineering team
PREFERRED QUALIFICATIONS
Bachelor's degree in computer science or equivalent
Amazon is committed to a diverse and inclusive workplace. Amazon is an equal opportunity employer and does not discriminate on the basis of race, national origin, gender, gender identity, sexual orientation, protected veteran status, disability, age, or other legally protected status.
J-18808-Ljbffr