Research Fellowship - Mechanistic Interpretability

Vmax · San Francisco

Apply on company site

About Vmax

Vmax is an applied research lab developing AI capable of open-ended learning. We are building systems to exceed humans in all capacities by optimizing beyond the local maxima of learning from human expertise.

About the role

LLMs are fantastically powerful and there is a rapidly growing corpus of work devoted to understanding their internal representations and computations. We use the tools of mechanistic interpretability to enhance reinforcement learning by generating intrinsic rewards as a supplement or alternative to downstream human-generated verifiers. 

This 3 to 6 month fellowship is for PhD students or equivalent early-career researchers who want to work at the intersection of mechanistic interpretability and reinforcement learning. You will own a focused research project, work closely with Vmax technical staff, and contribute to research publications.

Responsibilities

Role Requirements

Nice to have

Role specific location policy

Job alert

Get new jobs by email

Save this search and get relevant new jobs when they appear.