AI Understands Human Emotions Surprisingly Well, New Research Shows

AI Understands Human Emotions Surprisingly Well, New Research Shows

Ondřej Barták
Ondřej Barták
Entrepreneur and Programmer
27. 6. 2025
3 minutes reading · 4 views
AI Understands Human Emotions Surprisingly Well, New Research Shows

AI Understands Human Emotions Surprisingly Well, New Research Shows

A new study published in the prestigious journal Nature Human Behaviour has revealed a remarkable finding: large language models (LLMs) such as ChatGPT-4 and other artificial intelligence systems can perform better than average humans on emotional intelligence tests.

Results Exceeded Expectations

The research team tested six different AI models: ChatGPT-4, ChatGPT-o1, Gemini 1.5 flash, Copilot 365, Claude 3.5 Haiku, and DeepSeek V3. These systems achieved an average accuracy of 81% across five standard emotional intelligence tests, while people in the original validation studies achieved only 56% accuracy. The most successful models were ChatGPT-o1 and DeepSeek V3, which exceeded the human average by more than two standard deviations. All the AI systems tested significantly outperformed humans on all five tests focused on understanding and regulating emotions.

Tests Focused on Practical Situations

The researchers used five different tests measuring various aspects of emotional intelligence. For example, the Situational Test of Emotional Understanding (STEU) presented scenarios and asked participants to choose which emotion the person in the given situation was most likely to feel. One example was: "A supervisor who is unpleasant to work with leaves Alfonso's workplace. Alfonso is most likely to feel: a) joy, b) hope, c) regret, d) relief, e) sadness." The correct answer is relief. Other tests focused on emotion regulation—how to manage one's own emotions or help others with their emotional states in the workplace.

ChatGPT-4 Was Able to Create New Tests

In the second part of the research, the scientists asked ChatGPT-4 to create new versions of all five tests. The resulting tests were then administered to 467 participants from the United Kingdom and the United States. The results showed that the ChatGPT-created tests had a similar level of difficulty to the original versions. Participants rated the new items as clear and realistic, although they differed slightly from the original tests in some respects. The correlation between the original and AI-created tests reached r = 0.46, indicating a strong association.

Significance for the Future of AI

The study, led by a team from universities in Geneva and other institutions, suggests that current AI systems have a surprisingly sophisticated understanding of human emotions and their dynamics. This has significant implications for the development of chatbots, virtual assistants, and other applications where emotional intelligence is important. The researchers emphasize that the ability of AI systems to generate responses consistent with accurate knowledge of human emotions is a prerequisite for their successful deployment in healthcare, education, customer service, and other fields requiring a sensitive approach.

Questions for the Future

Although the results are promising, the researchers point to the ongoing debate over whether AI truly "understands" emotions or merely simulates understanding based on patterns in its training data. The key question remains whether these capabilities will lead to positive outcomes in real-world social interactions. The study also revealed a high level of agreement among different AI models—the correlation between their responses reached 0.88. Interestingly, AI systems tended to answer the same items correctly as humans—tasks that were more difficult for humans were also more difficult for artificial intelligence.

These findings open up new possibilities for using AI in psychology and other fields focused on human behavior, while also raising important questions about the nature of emotional intelligence and how it is measured.

Advertisement

Content created with help from UpTier.

SEO and GEO on autopilot. UpTier’s multi-agent systems write and optimize content for search engines and AI answers.

Discover UpTier ↗

Category:AI
Did you enjoy this article?
Discover more interesting posts on our blog
Back to blog

Related posts

OpenAI gives Codex reusable cloud workspaces accessible from any deviceOpenAI gives Codex reusable cloud workspaces accessible from any device
Codex gains reusable cloud development environments, alongside voice controls in its CLI, code reviews in the ChatGPT desktop app and cloud-based security tools.
2 min read
2. 10. 2026
Amazon releases Strands Decider 2B for AI workflow decisionsAmazon releases Strands Decider 2B for AI workflow decisions
Strands Decider 2B selects from predefined options and returns a confidence score. The fully open-source model is available now and small enough to run locally.
2 min read
1. 10. 2026
OpenAI says it disrupted a campaign to extract hidden model reasoningOpenAI says it disrupted a campaign to extract hidden model reasoning
OpenAI reported a coordinated effort to extract protected model reasoning and said it closed an extraction pathway. It attributed the main cluster of activity to individuals associated with Moonshot AI, the developer of Kimi.
3 min read
1. 10. 2026
Přihlaste se k odběru našeho newsletteru
Zůstaňte informováni o nejnovějších příspěvcích, exkluzivních nabídkách, a aktualizacích.
CodedTrip

Operated by CodedTrip LLC, USA.

YouTube
TikTok