Jobs and Careers
VM

Member of Technical Staff - RL Infrastructure

Vmax
San Francisco, USAfull_timePosted 20 May 2026

About the role

<h2><strong>About <em>V<sub>max</sub></em></strong></h2> <p><em>V<sub>max</sub></em> is an applied research lab developing AI capable of open-ended learning. We are building systems to exceed humans in all capacities by optimising beyond the local maxima of learning from human expertise.</p> <h2>About the role</h2> <p>This role is for strong infrastructure engineers who can build the systems layer for RL at scale: distributed rollouts, training orchestration, inference, evals, data pipelines, observability, and reliability. You will create the durable platform that enables researchers and applied ML engineers to run, debug, and reproduce large-scale RL experiments.</p> <h2 data-section-id="r8dte7" data-start="6952" data-end="6971">Responsibilities</h2> <ul data-start="6973" data-end="8359"> <li data-section-id="18ncivk" data-start="7232" data-end="7351">Build infrastructure for distributed RL training and inference across thousands of GPUs</li> <li data-section-id="1vp75id" data-start="7634" data-end="7709">Improve the reliability, debuggability, and throughput of RL experiments.</li> <li data-section-id="nt3ocq" data-start="7710" data-end="7839">Build interfaces that allow researchers and applied ML engineers to launch, inspect, compare, and reproduce experiments easily.</li> <li data-section-id="1vop8r7" data-start="7969" data-end="8109">Own infrastructure projects end to end, from architecture and implementation through deployment, documentation, and long-term maintenance.</li> <li data-section-id="a20cdd" data-start="8110" data-end="8235">Identify and eliminate bottlenecks in training, rollout generation, eval execution, data movement, and cluster utilization.</li> <li data-section-id="1x233uz" data-start="8236" data-end="8359">Maintain engineering standards for RL infrastructure, including testing, observability, versioning, and reproducibility.</li> </ul> <h2 data-section-id="w1j6vz" data-start="1480" data-end="1503">Minimum Requirements</h2> <ul> <li data-section-id="1wgb066" data-start="8386" data-end="8463">Strong software engineering experience.</li> <li data-section-id="1g8xqwl" data-start="9427" data-end="9537">Experience building infrastructure for LLM inference and/or RL training.&nbsp;</li> <li data-section-id="1cgittm" data-start="9538" data-end="9644">Experience with GPU clusters, distributed training, model serving, or high-throughput

Apply for this role

Generate a tailored application kit with a matched cover letter, interview prep, and CV highlights — in under 60 seconds.

Apply Now →Generate Application Kit

Free account required — sign up in 30s

Company

Vmax

View company profile →