Description
The Senior Inference Engineer will focus on building and optimizing inference engines for large-scale LLM serving systems. This involves working across various layers of the inference stack, analyzing bottlenecks, implementing optimization techniques, and collaborating with platform engineering teams to enhance serving architectures.
