Skip to main content

Inference Engineer at engineering and architecture firm

Department: Engineering

Setup
Hybrid
Location
Bellevue, Washington
Type
Full-time
Level
senior
Posted

Description

The company is hiring Senior and Staff Inference Engineers to build and operate production-grade model-serving and inference systems for high-throughput, low-latency AI workloads. The role focuses on GPU utilization, latency and throughput optimization, scalability, reliability, monitoring, and collaboration with AI training, GPU performance, orchestration, platform engineering, and operations teams. The position is hybrid in Bellevue, Washington, with approximately three days per week in the office, and requires U.S. work authorization; visa sponsorship is not available.

For job seekers

Ready to find a role that actually fits?

Upload your résumé, start a Job Search Thread, and let Metaintro rank real openings against your experience — then guide you from search to offer.

Match

Compare live roles against your current evidence.

Position

Turn proof projects into role-specific applications.

Improve

Use market feedback to keep the skill plan current.

Return to navigation