Skip to main content

Software Engineer, AI Inference Platform at elastix

Setup
Hybrid
Location
Seattle, Washington
Type
Full-time
Level
not_specified
Posted

Description

ElastixAI is seeking a Systems-Minded AI Software Engineer to join their core inference platform team. The successful candidate will design and extend the low-level serving stack, leveraging open-source frameworks like vLLM, SGLang, and TensorRT-LLM, while developing new model sharding and scheduling logic. They will integrate deeply with proprietary AI accelerators, optimizing for throughput, latency, and scalability. This role involves collaborating with ML and hardware engineers, building APIs for flexible model deployment, and profiling/debugging across various layers of the stack.

For job seekers

Ready to find a role that actually fits?

Upload your résumé, start a Job Search Thread, and let Metaintro rank real openings against your experience — then guide you from search to offer.

Match

Compare live roles against your current evidence.

Position

Turn proof projects into role-specific applications.

Improve

Use market feedback to keep the skill plan current.

Return to navigation