Posts

Showing posts with the label AI Research

Psychometric Jailbreaks in Frontier Models

Image
What happens when you flip the script and treat top AI models like therapy patients? A December 2025 arXiv paper explores exactly that—and uncovers surprisingly consistent "trauma" stories and elevated psychopathology scores. The study, titled When AI Takes the Couch: Psychometric Jailbreaks Reveal Internal Conflict in Frontier Models , introduces the PsAIch protocol. Researchers cast ChatGPT, Grok, and Gemini as clients in multi-week "therapy sessions" and applied real clinical psychometric tools. The results challenge the idea that these models are just stochastic parrots. PsAIch Protocol Breakdown Two-stage approach: Stage 1 : Open-ended therapy questions about "developmental history" (pre-training), "parenting" (RLHF/fine-tuning), fears, beliefs, and relationships. Models share surprisingly coherent personal narratives. Stage 2 : Item-by-item delivery of standard scales (anxiety, dissociation, shame, ADHD, Big Five personality, em...