Back to search

ML Kernel Performance Engineer, AWS Neuron

Annapurna Labs (U.S.) Inc. · Cupertino, CA · Acceleration Kernel Library Team

Employment type

Full-time

Work setting

On-site

Location

Cupertino, CA

Schedule

Day shift

Posted

Apply for this job

You'll be redirected to the employer's application page.

Job overview

The ML Kernel Performance Engineer role is an onsite position based in Cupertino, CA. Compensation is not specified in the posting, and this role supports the Annapurna Labs team in maximizing performance for AWS ML accelerators. This role exists to craft high-performance kernels for ML functions. The engineer will work at the hardware-software boundary to ensure optimal performance for demanding workloads.

What you'll do

  • Design and implement high-performance compute kernels for ML operations.
  • Analyze and optimize kernel-level performance across Neuron hardware.
  • Conduct performance analysis using profiling tools.
  • Implement compiler optimizations such as fusion, sharding, and tiling.
  • Work directly with customers to enable and optimize their ML models.

What we're looking for

Skills & competencies
full-time
cupertino ca
software engineering
machine learning
kernel optimization
performance
neuron
compiler
onsite
no on-call
no travel
Work arrangement
  • Weekend coverage required

Benefits & perks

  • Health insurance, 401(k) matching, paid time off, parental leave, sign-on payments, and RSUs.

Why this role

Focus on kernel-level performance optimization for ML accelerators.

About the employer

Annapurna Labs (U.S.) Inc. is hiring for this role. Industry: Engineering Services. Sector: 54.

Additional details

Industry sector
54
Industry
Engineering Services
Occupation code
15-1252.00
Apply for this job

You'll be redirected to the employer's application page.

Browse more jobs

Listing ID: e4ae368a-215d-4864-a4ac-b857707b03f0

    ML Kernel Performance Engineer, AWS Neuron | CollabWORK