Description
Cerebras Systems is hiring a Core ML Engineer to translate machine learning research into efficient software for its Wafer-Scale Engine. The role involves designing runtime components and high-performance kernels, implementing and benchmarking ML algorithms, profiling and debugging across frameworks, compilers, runtimes, communication, and kernels, and optimizing computation, memory movement, communication, and concurrency for large-scale training and low-latency inference. The engineer will collaborate with researchers and compiler, runtime, kernel, and inference teams, with a focus on C++ and Python, parallel programming, and modern ML frameworks.
