Software Development Engineer, Generative AI, Annapurna Labs - Neuroboros Team
Summary
This role involves building AI agents and tools to simplify and accelerate customer adoption of Amazon's Neuron software stack for Trainium machine learning silicon. You will work on agentic evaluation frameworks and software solutions to optimize machine learning workloads on AWS infrastructure.
Key job responsibilities
This role requires collaborating with other Neuron Software teams, Science, AWS AI Services, external partners and customers with a potential high impact on Amazon's top and bottom line. As a member of the team applying Generative AI to accelerate Neuron adoption, you will play a key role in shaping this space with the following technical responsibilities:
* Research implementations that deliver the best possible experiences for customers.
* Deliver on goals to improve the time and effort it takes to port and optimize Machine Learning workloads on Neuron.
* Build and operate agentic evaluation frameworks, monitoring, and feedback loops that ensure the correctness, reliability, and continuous improvement of AI agents
* Design, implement, test, deploy and maintain innovative software solutions to transform service performance, durability, cost, and security.
* Build high-quality, highly available, always-on products.
* Potentially contribute intellectual property through patents
About the team
We pursue the ambitious goal of leveraging and expanding Generative AI technologies to help customers benefit from the scale and price/performance equation offered by Amazon Machine Learning hardware. The creation of the team in NYC is key to Annapurna Labs’ location strategy, with the goal of creating an additional hub attracting top talent with varied backgrounds to work on challenging problems, using and building state-of-the-art tooling.
About Amazon Annapurna Labs:
Amazon Annapurna Labs team (our organization within AWS UC) is responsible for building innovation in silicon and software for our Amazon customers. We are at the forefront of innovation by combining cloud scale with the world’s most talented engineers. Our team covers multiple disciplines including silicon engineering, hardware design, software and operations. Because of our team’s breadth of talent, we have been able to improve AWS cloud infrastructure in high-performance machine learning with Neuron, Inferentia and Trainium ML chips, in networking and security with products such as Nitro, Enhanced Network Adapter (ENA), and Elastic Fabric Adapter (EFA), and in computing with Graviton and F1 EC2 instances.