Description
The LLM Engineer (LLM Training) designs and improves LLM training pipelines, trains generative language models for real-world services, and enhances model accuracy and stability through pre-training, post-training, and Self-Refine methodologies. The role requires at least three years of Deep Learning or NLP experience, Python and PyTorch programming, GPU-based distributed training, and model evaluation and optimization. Preferred qualifications include academic papers, conference presentations, Docker and Kubernetes, GPU cluster pipeline management, and LLM fine-tuning experience.
