ML Kernel Performance Engineer, AWS Neuron
Employment type
Full-time
Work setting
On-site
Location
Cupertino, CA
Schedule
Day shift
Posted
You'll be redirected to the employer's application page.
Job overview
The ML Kernel Performance Engineer role is an onsite position based in Cupertino, CA. Compensation is not specified in the posting, and this role supports the Annapurna Labs team in maximizing performance for AWS ML accelerators. This role exists to craft high-performance kernels for ML functions. The engineer will work at the hardware-software boundary to ensure optimal performance for demanding workloads.
What you'll do
- Design and implement high-performance compute kernels for ML operations.
- Analyze and optimize kernel-level performance across Neuron hardware.
- Conduct performance analysis using profiling tools.
- Implement compiler optimizations such as fusion, sharding, and tiling.
- Work directly with customers to enable and optimize their ML models.
What we're looking for
- Skills & competencies
- full-timecupertino casoftware engineeringmachine learningkernel optimizationperformanceneuroncompileronsiteno on-callno travel
- Work arrangement
- Weekend coverage required
Benefits & perks
- Health insurance, 401(k) matching, paid time off, parental leave, sign-on payments, and RSUs.
Why this role
Focus on kernel-level performance optimization for ML accelerators.
About the employer
Annapurna Labs (U.S.) Inc. is hiring for this role. Industry: Engineering Services. Sector: 54.
Additional details
- Industry sector
- 54
- Industry
- Engineering Services
- Occupation code
- 15-1252.00
You'll be redirected to the employer's application page.
Listing ID: e4ae368a-215d-4864-a4ac-b857707b03f0