Are you interested in advancing Amazon's Generative AI capabilities? Come work with a talented team of engineers and scientists in a highly collaborative and friendly team. We are building state-of-the-art Generative AI technology that will benefit all Amazon businesses and customers.

Key job responsibilities
As a Software Development Engineer, you will be responsible for designing, developing, testing, and deploying high performance inference capabilities, including but not limited to multi-modality, SOTA model architectures, latency, throughput, and cost. You will collaborate closely with a team of engineers and scientists to influence our overall strategy, and define the team’s roadmap. You will drive system architecture, spearhead best practices, and mentor junior engineers.

A day in the life
You will read papers and consult with scientists to get inspiration of emerging techniques, and blend those into our roadmap; You will design and experiment with new algorithms, benchmark the latency and accuracy of your implementations; Most importantly you will implement production grade solutions, and see them through the deployments swiftly; You may need to collaborate with other science and engineering teams to get things done properly; You will hold highest bar in operational excellence and support production systems, and constantly create solutions to minimize the ops load.

About the team
Our mission is to build best-in-class, fast, accurate, and cost-efficient large language model inference solutions and infrastructure that will enable Amazon businesses to deliver more value to their customers.

We are open to hiring candidates to work out of one of the following locations:

Boston, MA, USA | New York, NY, USA

Related Jobs