Member of Technical Staff - RL Algorithms

Vmax · San Francisco

Apply on company site

About Vmax

Vmax is an applied research lab developing AI capable of open-ended learning. We are building systems to exceed humans in all capacities by optimising beyond the local maxima of learning from human expertise.

About the role

RL has become the de-facto method of post-training LLMs. We are limited by the sample efficiency of the current policy gradient algorithms in use today, and are looking for a talented researcher to weave together pre-LLM and post-LLM approaches to learning from experience.

Responsibilities

Minimum Requirements

Nice to have

Role specific location policy

Compensation

The expected salary range for this position is $300,000 - $500,000 USD

Job alert

Get new jobs by email

Save this search and get relevant new jobs when they appear.