Description
Sciforium is hiring a Research Engineer focused on Model Evaluation & MLOps to build tools and infrastructure for evaluating, deploying, and operating multimodal foundation models on GPUs. The role includes automated quality and performance benchmarking, model registry and versioning, CI/CD workflows, monitoring, debugging, and collaboration with research, distributed systems, inference, and GPU kernel teams. Candidates need at least two years of professional ML or software engineering experience, strong Python and software engineering skills, hands-on experience with PyTorch, TensorFlow, or JAX, model evaluation or MLOps experience, GPU inference runtime experience, and an MS or PhD in a relevant technical field or equivalent practical experience.
