Software Development Engineer, AI/ML, AWS Neuron, Model Inference
Employment type
Full-time
Work setting
On-site
Location
Cupertino, CA
Schedule
Day shift
Posted
You'll be redirected to the employer's application page.
Job overview
The Software Development Engineer, AI/ML, AWS Neuron, Model Inference role is an onsite position based in Cupertino, CA. Compensation is not specified, though the role supports the Inference Enablement and Acceleration team. This role exists to build distributed inference support for PyTorch in the Neuron SDK. The engineer will tune models to maximize efficiency on Trainium and Inferentia silicon, working across the stack from system-level optimizations to framework-level integration.
What we're looking for
- Skills & competencies
- full-timeday shiftcupertino caannapurna labssoftware engineerai/mlmodel inferenceaws neuronpythonc++onsiteno on-callno travel
- Work arrangement
- Weekend coverage required
Benefits & perks
- Health insurance, 401(k) matching, paid time off, parental leave, sign-on payments, restricted stock units (RSUs)
About the employer
Annapurna Labs (U.S.) Inc. is hiring for this role. Industry: Custom Computer Programming Services. Sector: 54.
Additional details
- Industry sector
- 54
- Industry
- Custom Computer Programming Services
- Occupation code
- 15-1252.00
You'll be redirected to the employer's application page.
Listing ID: db8a2dfb-730c-4386-9883-a5a827264ce5