Skip to main content

Researcher, Vision at Sarvam

Department: Engineering

Setup
On-site
Location
Bengaluru, Karnataka
Level
not_specified
Posted

Description

Sarvam is hiring a Vision-Language Model Researcher to work across the full lifecycle of VLM development, including research, data, training, evaluation, production, and model robustness. The role focuses on vision-language architectures, multilingual training methods, data strategies, Indic multimodal benchmarks, failure modes, and interpretability, with close collaboration with engineers. Candidates should have strong research and experimental skills, a track record of impactful work, and strong PyTorch experience; relevant advanced degrees, publications, multilingual or low-resource experience, document understanding, OCR, and large-scale data curation are bonuses.

For job seekers

Ready to find a role that actually fits?

Upload your résumé, start a Job Search Thread, and let Metaintro rank real openings against your experience — then guide you from search to offer.

Match

Compare live roles against your current evidence.

Position

Turn proof projects into role-specific applications.

Improve

Use market feedback to keep the skill plan current.

Return to navigation