Teaching machines
to understand us.

Expert data, safety evaluations, and benchmarks for how AI understands, responds to, and affects people.

Built around people

Good judgment begins
with human experience.

We’re building a network of psychiatrists, psychologists, therapists, helpline workers and voice actors to help shape scenarios, review responses and create more thoughtful training data.

Meet the work

Behavioral alignment

Your principles,
put into practice.

We help AI teams define how their models should behave in difficult conversations, identify where they fall short, and develop expert training data to improve them. Then we test again on unseen situations.

  1. Define the behavior
  2. Evaluate
  3. Build targeted data
  4. Retest
How it works

Also available independently

01 / EVALUATE

Safety evaluations
& red teaming.

Structured evaluations and challenging conversations that reveal how models handle emotional complexity, changing context and safety boundaries.

Explore the research
02 / IMPROVE

Training data
& human feedback.

Expert demonstrations, preference judgments and review rubrics for fine-tuning, RLHF and more thoughtful model behavior.

Explore the datasets

From the journal.

All essays ↗
Black-and-white photograph of an open doorway into a quiet kitchenPerspective / 2 min read

What happens next.

As machines become more capable, what do we want to become possible?

Black-and-white image of two worn chairs beside a café windowPerspective / 2 min read

The space between words.

We can teach a machine what to say. What would it take to teach it what matters?

Saiki / Start a conversation

What are you
working on?

Expert data, evaluations,
and better model behavior.

I’m interested in

Please leave out confidential or patient information.
Privacy Policy