Skip to main content

RL Post-Training Engineer at Pluralis Research

Department: Research

Setup
Remote
Location
San Francisco, California
Type
Full-time
Level
senior
Posted

Description

Pluralis Research is hiring an RL Post-Training Engineer to build and ship decentralized post-training systems for large language models. The role owns the end-to-end RL training loop, including rollout ingestion, reward computation, policy updates, weight synchronization, asynchronous and high-latency algorithms, evaluation, and the first public release of a post-trained model. Candidates should have hands-on experience with RL post-training, production Python and PyTorch, and distributed or asynchronous systems; the position is remote-first, offers equity-heavy compensation, and provides optional visa sponsorship to Australia or the United States.

For job seekers

Ready to find a role that actually fits?

Upload your résumé, start a Job Search Thread, and let Metaintro rank real openings against your experience — then guide you from search to offer.

Match

Compare live roles against your current evidence.

Position

Turn proof projects into role-specific applications.

Improve

Use market feedback to keep the skill plan current.

Return to navigation