Senior Machine Learning Compiler Engineer (AWS Neuron)
Employment type
Full-time
Work setting
On-site
Location
Cupertino, CA
Schedule
Day shift
Posted
You'll be redirected to the employer's application page.
Job overview
The Senior Machine Learning Compiler Engineer is an onsite role based in Cupertino, CA. Compensation is not specified for this position, which supports the AWS Neuron SDK. This role exists to build the next-generation compiler for AWS Inferentia and Trainium chips. By solving complex optimization problems, the engineer enables the deployment of cutting-edge ML models on custom AWS hardware.
What you'll do
- Design and maintain software for the Neuron compiler.
- Solve compiler optimization problems for ML models.
- Partner with chip architects and ML teams.
- Work with open-source communities (OpenXLA, MLIR).
- Resolve compiler defects.
What we're looking for
- Skills & competencies
- full-timeday shiftcupertino caamazonmachine learningcompiler engineeringneuronaionsitec++
- Work arrangement
- Weekend coverage required
Benefits & perks
- Sign-on payments, RSUs, health insurance, 401(k), paid time off, parental leave.
Why this role
Focus on optimizing massive-scale LLMs like Llama and Deepseek.
About the employer
Annapurna Labs (U.S.) Inc. is hiring for this role. Industry: Custom Computer Programming Services. Sector: 54.
Additional details
- Industry sector
- 54
- Industry
- Custom Computer Programming Services
- Occupation code
- 15-1252.00
You'll be redirected to the employer's application page.
Listing ID: 1b0c801b-0bc5-453d-bd0b-8da596d3407d