Fourteen-year-old Zeynep Demirbas tested four AI models on 3,553 Reddit posts to see how accurately they could detect stress. MentalBERT performed best at about 82%, while ChatGPT-4o scored about 74%. Her findings suggest general-purpose LLMs may not be reliable enough for mental health assessment.
Read more at the source
Disclaimer: The content of this post is sourced from external sites and is for informational purposes only. All rights and credits belong to the original authors and publishers.
