Adolescent Mental Health Clinical Expert (Contract)

San Francisco · Remote ok$75 – $100

About this role

About the Role

We’re looking for a Clinical AI Safety Contractor to help evaluate how AI systems respond in mental health and other clinically sensitive situations.

You’ll bring clinical judgment to reviewing AI conversations, identifying safety risks, and helping define what appropriate model behavior looks like. This is a flexible, part-time role for clinicians interested in applying their expertise to AI safety and evaluation.

What You'll Do

  • Review and rate AI conversations for clinical safety, appropriateness, and quality

  • Identify clinically meaningful risks and model failure modes

  • Help create realistic scenarios, evaluation criteria, and scoring rubrics

  • Provide expert feedback on how AI should respond in sensitive situations

  • Participate in calibration and review sessions with researchers and engineers

  • Document important edge cases and emerging risks

Requirements

  • Clinical training and professional credentials are required, such as LCSW/LICSW, LMFT, LPC/LPCC/LCPC, clinical psychologist, psychiatrist, MD/DO, or comparable clinical mental health credentials

  • Experience working directly with patients or clients in a mental health setting

  • Strong grounding in clinical assessment, psychopathology, risk evaluation, or crisis response

  • Strong written communication and clinical judgment

  • Comfort reviewing sensitive mental health content and maintaining strict confidentiality

No PhD or prior AI experience is required.

Nice to Have

  • Experience with crisis intervention, suicide or self-harm assessment, or safety planning

  • Experience in adolescent mental health, psychosis, eating disorders, trauma, or other clinically complex areas

  • Experience developing clinical rating systems, assessment criteria, or coding frameworks

  • Familiarity with AI, digital mental health, trust & safety, or conversational systems

About us

Vals AI builds rigorous evaluations and benchmarks for frontier AI systems. Our work started from NLP evaluation research at Stanford, and today we work across technical and domain-specific areas, including healthcare and mental health. We raised a $5M seed and our team has backgrounds at Stanford, NVIDIA, Meta, Microsoft, Palantir, HRT, Jane Street, and Snorkel.

We recently announced our $40M Series A at a $400M valuation, led by Andreessen Horowitz, with participation from existing investors 8VC, Pear VC, and Bloomberg and new investors Hudson River Trading and NextLadder Ventures.

Vals in the Media:

Company at a glance

Vals AI builds the standard for evaluating how well large language models perform real-world tasks through high-quality benchmarks and large-scale evaluations used by major foundation model labs, enterprises, and research teams. The company's core methodology is grounded in NLP evaluation research conducted at Stanford, providing a rigorous foundation for its assessment platform. With a founding team that brings experience from leading technology companies including NVIDIA, Meta, Microsoft, Palantir, and HRT, along with over 300 collectively published research papers, Vals AI combines deep technical expertise with academic rigor. The company raised a $5M seed round from top institutional and angel investors in Silicon Valley, positioning it as a trusted partner for organizations seeking to understand and optimize their language model performance.

Founded2024
Team Size1-10
WorkspaceRemote ok
IndustryAI/ML
Location
San Francisco, California, United States
Websitevals.ai
LinkedInLinkedIn

Tired of cold applications?

Sign up with Clera and we'll reach out the moment a role actually fits you — no more spraying applications into the void.

Know someone who'd be great for this?