Finding the best job has never been easier
Share
AWS Neuron is the complete software stack for the AWS Inferentia and Trainium machine
This role is responsible for development, enablement and performance optimization (latency, throughput) of a wide variety of ML model families, including massive scale large language models like Llama3, DBRX, Stable Diffusion, Vision Transformers and many more.The Neuron Inference team works side by side with compiler and runtime engineers to create, build and tune distributed inference solutions with Trn1/Inf2. Experience optimizing inference performance for both latency and throughput on these large models using PyTorch or JAX is a must. Experience with technologies/tools such as vLLM, Hugging Face, multi-modal inference, etc. is highly valued.Key job responsibilities
This role will help lead the efforts building distributed inference support into Pytorch, Tensorflow using XLA and the Neuron compiler and runtime stacks. This role will help tune these models to ensure highest performance and maximize the efficiency of them running on the AWS Trainium and Inferentia silicon. Strong software development using C++/Python and ML knowledge are both critical to this role.A day in the life
As you design and code solutions to help our team drive efficiencies in software architecture, you’ll create metrics, implement automation and other improvements, and resolve the root cause of software defects. You’ll also:
Participate in design discussions, code review, and communicate with internal and external stakeholders.Work in a startup-like development environment, where you’re always working on the most important stuff.
- 3+ years of non-internship professional software development experience
- 2+ years of non-internship design or architecture (design patterns, reliability and scaling) of new and existing systems experience
- Experience programming with at least one software programming language
- 3+ years of full software development life cycle, including coding standards, code reviews, source control management, build processes, testing, and operations experience
- Bachelor's degree in computer science or equivalent
These jobs might be a good fit