Safety evaluations
& red teaming.
Structured evaluations and challenging conversations that reveal how models handle emotional complexity, changing context and safety boundaries.
Explore the researchExpert data, safety evaluations, and benchmarks for how AI understands, responds to, and affects people.
Built around people
We’re building a network of psychiatrists, psychologists, therapists, helpline workers and voice actors to help shape scenarios, review responses and create more thoughtful training data.
Meet the workBehavioral alignment
We help AI teams define how their models should behave in difficult conversations, identify where they fall short, and develop expert training data to improve them. Then we test again on unseen situations.
Also available independently
Structured evaluations and challenging conversations that reveal how models handle emotional complexity, changing context and safety boundaries.
Explore the researchExpert demonstrations, preference judgments and review rubrics for fine-tuning, RLHF and more thoughtful model behavior.
Explore the datasets