Research Scientist (AI Behaviours)

Whitecircle · Paris · FullTime

Apply on company site

TLDR: We're looking for a research scientist to study how LLM agents fail in the wild, who can elicit deception, misalignment, and unsafe behaviour in concrete experiments, and build out the understanding of how agents break or misbehave in realistic and user-related scenarios.

About us

White Circle is an AI Safety company building the safety, reliability, and optimization layer for AI systems. At the core of our platform are policies – simple natural-language rules that define what an AI model should and shouldn’t do. We automatically test, enforce, and continuously improve these policies at scale.

We’re a small, highly focused team. If you want to work deeply on hard problems, see your work ship to production quickly, and influence how AI safety is actually built – you’re the one we need.

About the team

White Circle's fundamental research team works on the science of how AI systems fail: where agents break, why misalignment and unsafe behaviours emerge, and how to catch them before they reach the real world. We build the evals, benchmarks, environments, and tooling that empirically study the most pressing AI safety concerns — some of which become the guardrails shipped in our products, and some of which become public writeups.

You will:

You’ll fit right in if you:

A big plus:

Why White Circle

 

How we hire

  1. Introductory call with HR (25 min)

  2. Take-home test task

  3. Technical interview with Head of Fundamental Research (60 min)

  4. Final conversation with our CEO (45 min)

Please submit your application in English.

Job alert

Get new jobs by email

Save this search and get relevant new jobs when they appear.