Description
RadixArk is hiring a Member of Technical Staff, Developer Technology (DevTech) to accelerate LLM inference and training on modern GPU hardware. The role focuses on profiling and optimizing GPU performance, building custom CUDA/ROCm/Triton kernels, enabling new models and hardware, supporting speculative decoding and training systems, and translating expert user problems into reproducible technical guidance and open-source improvements. Candidates need at least four years of experience in GPU systems, LLM infrastructure, or performance engineering, along with strong profiling, debugging, Python, C++, and GPU-programming skills.
