Shape the Future of AI Accelerators at AWS Neuron
Join the team behind AWS Neuron - the software stack that powers AWS's purpose-built AI accelerators, Inferentia and Trainium. As a Senior Software Engineer on our Machine Learning Applications team, you will optimize the world's most demanding AI models at a scale few engineers ever get to work on.
What You'll Do
• Build and scale distributed inference solutions for leading large language models, including GPT, Kimi, and Qwen
• Partner directly with silicon architects and compiler engineers to shape the next generation of AI acceleration
• Write custom kernels that optimize LLM computation graphs, improving latency and cost for billions of inference requests worldwide
• Optimize state-of-the-art language, vision, and multimodal generative AI models for Neuron hardware
Key job responsibilities
You will drive the Evolution of Distributed AI at AWS Neuron
As a Technical Leader at the forefront of AWS's AI Accelerator, you'll architect the bridge between ML frameworks including PyTorch, JAX and AI hardware. This isn't just about just optimization-it's about revolutionizing how AI models run at scale.
Technical Impact You'll Drive:
• Spearhead distributed inference architecture for PyTorch and JAX using XLA
• Engineer breakthrough performance optimizations for AWS Trainium and Inferentia
• Develop kernels to improve model efficiency on Amazon AI Accelerators
• Transform complex tensor operations into highly optimized hardware implementations
What Makes This Role Unique:
• Direct influence on AWS's AI infrastructure used by thousands of ML applications
• Full-stack optimization from high-level frameworks to hardware-specific primitives
• Creation of tools and frameworks that define industry standards for ML deployment
• Collaboration with both open-source ML communities and hardware architecture teams
Your Technical Arsenal Should Include:
• Deep expertise in Python and ML framework internals
• Strong understanding of distributed systems and ML optimization
• Passion for performance tuning and system architecture
A day in the life
Work/Life Balance
Our team puts a high value on work-life balance. It isn't about how many hours you spend at home or at work; it's about the flow you establish that brings energy to both parts of your life. We believe striking the right balance between your personal and professional life is critical to life-long happiness and fulfillment. We offer flexibility in working hours and encourage you to find your own balance between your work and personal lives.
Mentorship & Career Growth
Our team is dedicated to supporting new members. We have a broad mix of experience levels and tenures, and we're building an environment that celebrates knowledge sharing and mentorship. We care about your career growth and strive to assign projects based on what will help each team member develop into a better-rounded professional and enable them to take on more complex tasks in the future.
BASIC QUALIFICATIONS
- 5+ years of full software development life cycle, including coding standards, code reviews, source control management, build processes, testing, and operations experience
- 5+ years of programming experience using Python or C++ and PyTorch.
- Experience with AI acceleration via quantization, parallelism, model compression, batching, KV caching, vllm serving
- Experience with accuracy debugging & tooling, performance benchmarking of AI accelerators
- Fundamentals of Machine learning and deep learning models, their architecture, training and inference lifecycles along with work experience on optimizations for improving the model execution.
PREFERRED QUALIFICATIONS
- Master's degree in computer science or equivalent
- Master's degree in machine learning or equivalent
- Experience with accuracy debugging & tooling, performance benchmarking of AI accelerators
- Experience in developing CUDA kernels, HPC and inference optimization, tensors operations
The base salary range for this position is listed below. Your Amazon package will include sign-on payments and restricted stock units (RSUs). Final compensation will be determined based on factors including experience, qualifications, and location. Amazon also offers comprehensive benefits including health insurance (medical, dental, vision, prescription, Basic Life & AD&D insurance and option for Supplemental life plans, EAP, Mental Health Support, Medical Advice Line, Flexible Spending Accounts, Adoption and Surrogacy Reimbursement coverage), 401(k) matching, paid time off, and parental leave. Learn more about our benefits at https://amazon.jobs/en/benefits.
USA, WA, Seattle - 168,100.00 - 227,400.00 USD annually