Fourteen-year-old Zeynep Demirbas tested four AI models on 3,553 Reddit posts to see how accurately they could detect stress. MentalBERT performed best at about 82%, while ChatGPT-4o scored about 74%. Her findings suggest general-purpose LLMs may not be reliable enough for mental health assessment. Fourteen-year-old Zeynep Demirbas tested four AI models on 3,553 Reddit posts to see how accurately they could detect stress. MentalBERT performed best at about 82%, while ChatGPT-4o scored about 74%. Her findings suggest general-purpose LLMs may not be reliable enough for mental health assessment.
Trending
- PM Modi meets BJP’s new 65-member national team at party headquarter
- This US island was set to become a resort until 2 storms changed everything
- ‘You don’t belong to anyone else, but yourself’: Rahul Gandhi’s message to young women
- ‘Deal with China confidently’: Jaishankar on India strategy; lays out 3Cs, 4Ds global outlook
- ‘Modi Chalisa’ controversy: Press Club of India pesident says event violated venue rules
- Amrit Bharat 3.0: AC coaches with Vande Bharat-style features for common man soon
- One is not born, but becomes: India’s drag story, from myth to modern stage
- ‘Mind your tongue’: Congress vs Congress erupts over ‘role model’ praise for Sajjan Kumar