Description
Pika is hiring a staff- or lead-level Research Engineer, Data to architect and scale data engineering systems supporting model training for advanced multimodal foundation models. The role owns large-scale data pipelines, curates and manages text, image, audio, and video datasets, develops ingestion, labeling, filtering, augmentation, and storage tools, and ensures data quality, reliability, privacy, and compliance. Candidates should have at least five years of experience building and scaling data pipelines for machine learning applications, expertise in distributed data systems and ML data curation, strong programming skills, and familiarity with cloud platforms. The position offers a hybrid on-site/remote arrangement based in Palo Alto, California, with competitive salary, equity, health benefits, and 401k matching.
