Description
Arm is hiring an experienced engineer in San Jose to help customers optimize AI workloads for production AI models running on Arm technology, with a strong focus on inference performance, power efficiency, and kernel-level implementation for DNNs. The role is highly customer-facing and collaborative, involving performance diagnosis, technical communication with internal teams and leadership, production-quality reference implementations, and influence over Arm’s IP and software roadmaps. It is a hybrid role, requires strong Python, C++, and kernel-level optimization experience, and does not offer visa sponsorship.
