Description
Liquid AI is hiring a technical owner for enterprise vision-language model post-training projects. The role translates customer requirements into multimodal post-training specifications, designs visual data generation and quality processes, runs supervised fine-tuning, preference alignment, and reinforcement learning workflows, and develops evaluations for visual understanding, grounding, OCR, and document parsing. The position also contributes applied learnings to Liquid’s core multimodal post-training infrastructure. Required experience includes VLM or multimodal data generation and evaluation, vision-language model training or fine-tuning, visual data quality, and multimodal evaluation; the role offers competitive base salary, equity, full medical, dental, and vision premium coverage, 401(k) matching, and unlimited PTO.

