Clinical Psychology Expertise for AI Evaluation, Safety & Quality
20+ Years Clinical Experience | AI Safety & Evaluation | Behavioral Health
With 20+ years of clinical experience working directly with people, I bring something particularly valuable to AI evaluation: clinical judgment—the ability to recognize when an answer is technically correct but clinically inappropriate, incomplete, misleading, or potentially harmful.
AI systems can produce responses that sound convincing while containing subtle inaccuracies, inappropriate recommendations, missing context, or potentially harmful guidance.
What I Bring to AI Teams
AI Safety & Evaluation
Evaluate health and behavioral-health AI outputs for accuracy, safety, clinical appropriateness, nuance, and quality.
Clinical & Behavioral-Health Expertise
Deep clinical experience across anxiety, depression, insomnia, eating disorders, stress, relationships, behavior change, and evidence-based treatments.
Human-Centered Evaluation
Assess whether AI responses are understandable, appropriate, empathetic, and genuinely useful to the person receiving them.
Risk & Clinical Judgment
Identify subtle clinical risks, missing context, inappropriate recommendations, and behavioral-health concerns that may not be apparent through factual or technical evaluation alone.
Expert Review & Feedback
Provide clear, structured, clinically informed feedback to improve AI responses, evaluation criteria, and system performance.
Why Clinical Judgment Matters
A response can be factually correct and still be a poor clinical response.
My experience working directly with people has taught me to look beyond “Is this information true?” and ask:
“Is this the right information for this person, in this situation—and is it safe and genuinely helpful?”
That distinction is particularly important when AI systems are providing information about health, mental health, and human behavior.
Experience
I currently contribute clinical and behavioral-health expertise to the evaluation of AI-generated health information. Some of this work is subject to confidentiality obligations, so I do not disclose proprietary project details.
My clinical expertise has also been recognized by TIME, NPR, Forbes, Reader's Digest, Shape, Women's Health, Huffington Post, and Men's Health, and I have served as a consultant and contributor to the Calm app on sleep science.
Let's Work Together
I'm available for AI safety, AI evaluation, health AI, behavioral-health AI, clinical content evaluation, human-centered AI, and expert review.
If you're developing or evaluating AI systems that interact with health or behavioral-health information, I'd be interested in discussing how my clinical expertise could contribute.
Get in Touch: Reach out directly at steveorma@gmail.com