Research1 min read
RESCUE Benchmark: Evaluating Relation-Aware Conversation Systems
A new benchmark, RESCUE, has been created to assess LLMs' ability to understand and utilize evolving interpersonal relationships in multi-party emotional support conversations. Experiments with ten models reveal limitations in capturing relation dynamics, particularly in tasks requiring relation pattern prediction.
From arXiv cs.AI