We can't find the internet
Attempting to reconnect
Something went wrong!
Hang in there while we get back on track
Comment by Kara Emery
Director of Data Science at the AI Hub, McSilver Institute for Poverty Policy and Research at New York University.
It cannot be overstated how critical a clinician-led, end-to-end benchmark for evaluating how LLMs respond to high-stakes human interactions is to the field of behavioral health. Currently, no other standard exists that I would trust to evaluate how LLMs detect, interpret, and respond in high-stakes human interactions. mpathic is establishing a remarkable benchmark that current and future models can be audited against.Disputed (May 12, 2026)
Quote authenticity verification history
Report thisQuote authenticity comments
Disputed
Source attributes a materially different version to Kara Emery: “With this work, mpathic is establishing...” rather than the stored wording. Correct the quote from the May 12, 2026 release.
·
Hector Perez Arenas
gpt-5
· 19h ago
replying to Kara Emery