Expoint - all jobs in one place

מציאת משרת הייטק בחברות הטובות ביותר מעולם לא הייתה קלה יותר

Limitless High-tech career opportunities - Expoint

Amazon Senior Software Engineer - Generative AI AGI Inference Engine 
United States, Massachusetts, Boston 
366291395

10.06.2024
DESCRIPTION

Key job responsibilities
As a Senior Software Development Engineer, you will be responsible for designing, developing, testing, and deploying high performance inference capabilities, including but not limited to multi-modality, SOTA model architectures, latency, throughput, and cost. You will collaborate closely with a team of engineers and scientists to influence our overall strategy, and define the team’s roadmap. You will drive system architecture, spearhead best practices, and mentor junior engineers.A day in the life
You will read papers and consult with scientists to get inspiration of emerging techniques, and blend those into our roadmap; You will design and experiment with new algorithms, benchmark the latency and accuracy of your implementations; Most importantly you will implement production grade solutions, and see them through the deployments swiftly; You may need to collaborate with other science and engineering teams to get things done properly; You will hold highest bar in operational excellence and support production systems, and constantly create solutions to minimize the ops load.

BASIC QUALIFICATIONS

- 5+ years of non-internship professional software development experience
- 5+ years of programming with at least one software programming language experience
- 5+ years of leading design or architecture (design patterns, reliability and scaling) of new and existing systems experience
- Experience as a mentor, tech lead or leading an engineering team


PREFERRED QUALIFICATIONS

- 5+ years of full software development life cycle, including coding standards, code reviews, source control management, build processes, testing, and operations experience
- Bachelor's degree in computer science or equivalent
- Experience with Python, PyTorch, and C++ programming and performance optimization
- Experience with Large Language Model inference
- Experience with Trainium and Inferentia Development
- Experience with GPU programming (TensorRT-LLM)