Neural Engine Performance Engineer, Platform Architecture

Cupertino Until 9/22/2026 3+ years exp First posted July 24, 2026 Last posted July 24, 2026
Job description

At Apple, Platform Architecture is responsible for connecting our hardware and software into one unified system. Join this team, and you'll collaborate with engineers across Apple to design how all of our technologies work in unison. In this role, you will be part of the Neural Engine IP architecture team and work to improve the performance of the Neural Engine IP through SW optimizations and HW architectural enhancements.

Description

As a Neural Engine Performance Engineer, you will be responsible for analyzing, debugging and optimizing the performance of the Neural Engine.

Minimum Qualifications

BS degree
Experience with C++ and Python
Experience with performance profiling of HW or SW
Experience in at least one hardware IP: ML HW accelerators or processing units such as GPUs, ISPs, Video CODECs, CPUs, or similar

Preferred Qualifications

MS or PhD degree
3+ years of software experience
Experience writing low level software interfaces to hardware
Experience debugging complex system level performance issues
Experience writing automation software for data collection and analysis
Understanding of OS scheduling and memory management
Understanding of ML workloads and deployment for inference
Understanding of SoC cache hierarchy and its performance implications

About this role

Summary

Optimize Neural Engine performance through software and hardware enhancements.

Job title

Neural Engine Performance Engineer, Platform Architecture

Experience level

3+ years

Minimum experience

3+ years exp

Industry

technology

Location requirements

must be based in Cupertino, no remote work specified

Salary

Not specified

Management role

No

Skills & keywords

Required skills

C++Pythonperformance profilinghardware IPsoftware interfaces

Preferred skills

low level softwareperformance debuggingautomation softwareOS schedulingML workloads

Specializations

performance profilinghardware IPML hardware acceleratorssystem performancehardware architecture
Locations

Structured locations inferred from the posting.

Cupertino, CA, USA

On-site City