Description
Lumai is hiring a Software Inference Deployment Engineer to integrate and support its Iris optical AI compute platform in third-party data centre environments. The role begins with software-stack integration, model onboarding, and familiarity with the disaggregated prefill/decode runtime, then moves to customer integration, troubleshooting, training, and feedback. Candidates need software engineering experience in AI infrastructure or accelerator integration, strong Python and PyTorch skills, model-deployment experience, inference-serving familiarity, and Linux, Docker, and cluster experience. The position offers competitive salary, share options, pension, private health insurance, development allowance, subsidised lunches, and 25 paid holiday days.

