Skip to main content

AI Research Engineer (Kernel & Inference Optimization) - 100% Remote Worldwide at Tether

Department: Data

Setup
Remote
Type
Full-time
Level
senior
Posted

Description

Tether is hiring an AI Model Serving Engineer to design, deploy, and optimize model-serving architectures and inference frameworks for advanced AI systems. The role focuses on high-throughput, low-latency, low-memory, and scalable inference across resource-constrained devices, edge platforms, and GPU clusters, including controlled inference testing, performance benchmarking, bottleneck diagnosis, and integration of optimized serving pipelines into production systems. Candidates need a computer science or related degree, strong AI research and low-level kernel optimization experience, Metal Shading Language expertise, and knowledge of modern model architectures and inference techniques.

Trending job searches

Every query opens live roles, salary samples, and market demand — tap a search to run it instantly.

Get More from Metaintro

Unlock powerful job search, personalized recommendations, and deep career insights from comprehensive, market-leading data and live market signals.

For job seekers

Ready to find a role that actually fits?

Upload your résumé, start a Job Search Thread, and let Metaintro rank real openings against your experience — then guide you from search to offer.

Match

Compare live roles against your current evidence.

Position

Turn proof projects into role-specific applications.

Improve

Use market feedback to keep the skill plan current.

Return to navigation