Member of Technical Staff - Mechanistic Interpretability

Vmax · San Francisco

Apply on company site

About Vmax

Vmax is an applied research lab developing AI capable of open-ended learning. We are building systems to exceed humans in all capacities by optimising beyond the local maxima of learning from human expertise.

About the role

LLMs are fantastically powerful and there is a rapidly growing corpus of work devoted to understanding their internal representations and computations. We use the tools of mechanistic interpretability to enhance reinforcement learning by generating intrinsic rewards as a supplement or alternative to downstream human-generated verifiers. 

Responsibilities

Minimum Requirements

Nice to have

Role specific location policy

Compensation

The expected salary range for this position is $300,000 - $500,000 USD

Job alert

Get new jobs by email

Save this search and get relevant new jobs when they appear.