Description
Mind Robotics is hiring a Research Engineer to build scalable infrastructure for training large machine-learning models in industrial settings. The role focuses on distributed training across hundreds of GPUs, parallelization and compute optimization, experiment and monitoring tools, training-stack debugging, production deployment support, and inference performance. Candidates need strong experience with large-scale ML training, distributed training, Python, and PyTorch or JAX.
