Senior Software Development Engineer, AI/ML, AWS Neuron, Model Inference
Compensation
$193k – $261k /yr
Employment type
Full-time
Work setting
On-site
Location
Cupertino, CA
Schedule
Day shift
Posted
You'll be redirected to the employer's application page.
Job overview
Compensation is $193,000–$261,000 per year. The Senior Software Development Engineer, AI/ML, AWS Neuron, Model Inference role is an onsite position based in Cupertino, CA. 00 - 261,500.00 USD annually, and this role supports the Annapurna Labs team within AWS, focusing on model inference enablement. This role exists to architect and implement distributed inference support for PyTorch and other frameworks. It contributes to the team's goal of maximizing performance and efficiency for large language models running on Trainium and Inferentia silicon.
What you'll do
- Design and optimize ML models for custom hardware
- build infrastructure for systematic model analysis
- implement high-performance kernels
- collaborate with customers to enable and optimize their ML models.
What we're looking for
- Skills & competencies
- full-timeday shiftcupertino caannapurna labssenior software engineerai/mlinferencepytorchllmoptimization$193k-$261k
- Work arrangement
- Weekend coverage required
Benefits & perks
- Health insurance, 401(k) matching, paid time off, parental leave, sign-on payments, and restricted stock units (RSUs).
Why this role
Focus on distributed inference, LLM optimization, and high-performance kernel development.
About the employer
Annapurna Labs (U.S.) Inc. is hiring for this role. Industry: Custom Computer Programming Services. Sector: 54.
Additional details
- Industry sector
- 54
- Industry
- Custom Computer Programming Services
- Occupation code
- 15-1252.00
You'll be redirected to the employer's application page.
Listing ID: e4ae36b5-1fcb-45a5-aec1-526170481587