Skip to main content

Senior Inference Optimization Engineer - Dragonfly Portfolio at Dragonfly

Department: Portfolio

Compensation

$180,000 – $250,000/yr

Setup
Hybrid
Type
Full-time
Level
senior
Posted

Summary from listing

Dragonfly is sourcing a Senior Inference Optimization Engineer for a portfolio company building privacy-first consumer AI infrastructure. The role focuses on improving LLM inference throughput, latency, and cost per token at scale, including GPU infrastructure, batching, attention, KV cache, speculative decoding, quantization, distributed inference, profiling, and load balancing. The position is remote in the United States, with compensation of $180,000–$250,000 for senior-level candidates.

For job seekers

Ready to find a role that actually fits?

Upload your résumé, start a Job Search Thread, and let Metaintro rank real openings against your experience — then guide you from search to offer.

Match

Compare live roles against your current evidence.

Position

Turn proof projects into role-specific applications.

Improve

Use market feedback to keep the skill plan current.

Return to navigation