Skip to main content

Cortex Real-Time Inference Engineering Lead at Cortex

Compensation

$70/hr

Location
Alexandria, Virginia
Level
lead/mgmt
Posted

Description

The Cortex Real-Time Inference Engineering Lead will design, build, and industrialize low-latency, highly available model-serving platforms for real-time predictive workflows. The role covers online inference architecture, production deployment patterns, autoscaling, performance and latency optimization, monitoring, SLOs, CI/CD, and resilience across cloud and on-premises environments. It requires expertise in Kubernetes, model-serving frameworks, APIs, load testing, observability, and automation, with preferred qualifications including a relevant degree, 5+ years of software or DevOps experience, and programming skills in Python, Go, C++, or Java.

For job seekers

Ready to find a role that actually fits?

Upload your résumé, start a Job Search Thread, and let Metaintro rank real openings against your experience — then guide you from search to offer.

Match

Compare live roles against your current evidence.

Position

Turn proof projects into role-specific applications.

Improve

Use market feedback to keep the skill plan current.

Return to navigation