Skip to content

New York 8th grader tests AI for stress; basic model beats ChatGPT-4o

Fourteen-year-old Zeynep Demirbas tested four AI models on 3,553 Reddit posts to see how accurately they could detect stress. MentalBERT performed best at about 82%, while ChatGPT-4o scored about 74%. Her findings suggest general-purpose LLMs may not be reliable enough for mental health assessment.

Leave a Reply

Your email address will not be published. Required fields are marked *