MARATTO

article

Do LLMs Grasp Emotion or Just Keywords? A Benchmark for Implicit Emotion Recognition

Abstract

Recent advances in Large Language Models (LLMs) have shown strong performance in natural language understanding tasks, including emotion recognition. Implicit Emotion Recognition (IER) focuses on identifying emotions in text without explicit emotional cues, posing a significant challenge in Natural Language Processing. This study presents a systematic evaluation of state-of-the-art LLMs, namely, ChatGPT-5, Phi-4, LLaMA-3, and Gemma-3, on IER under zero-shot and few-shot prompting strategies. The models were tested on two complementary datasets: ISEAR, containing naturally implicit emotions in coherent narratives, and TEST 2018, where explicit emotion words were masked to simulate context-only inference. Experimental results show that LLMs, particularly ChatGPT-5, can effectively capture implicit emotions, with few-shot prompting consistently improving performance. Performance differences between datasets highlight the role of context quality, with narrative texts facilitating inference and social media content introducing challenges.

Research topics

  • Emotion and Mood Recognition
  • Sentiment Analysis and Opinion Mining
  • Emotions and Moral Behavior

Read the original research

This page summarises published work. The authoritative version sits with the publisher.

DOI: 10.1109/aicps66617.2025.11513483

Is something wrong with this record? Report it or request removal.

Discussion

Discuss this research

Have you built on this work, tried to replicate it, or seen it applied in practice? Share what you know. Verified researchers and MARATTO™ domain experts can open a discussion, and any member can reply. Contributions are reviewed before they appear.

No discussion yet. Open the first thread.