Skip to main content

AI Infrastructure System Engineer Bangalore at Together AI

Department: EngineeringEducation: education_optional

Location
Bengaluru, Karnataka
Level
mid
Posted

Description

Together AI is hiring an engineer to build and operate GPU fleets for frontier model training and inference. The role focuses on automating GPU cluster provisioning, deployment, upgrades, repairs, and retirement; creating AI infrastructure agents for failure detection and remediation; developing fleet intelligence and validation systems; and improving GPU availability, utilization, performance, and reliability. Candidates need at least three years of experience with distributed systems, infrastructure platforms, or large-scale backend software, along with strong software engineering skills in Python, Go, or Rust and experience with Linux, Kubernetes, Terraform, Ansible, or similar technologies.

For job seekers

Ready to find a role that actually fits?

Upload your résumé, start a Job Search Thread, and let Metaintro rank real openings against your experience — then guide you from search to offer.

Match

Compare live roles against your current evidence.

Position

Turn proof projects into role-specific applications.

Improve

Use market feedback to keep the skill plan current.

Return to navigation