LLM assistants are widely used for social advice, yet evaluating their social reasoning is challenging because it depends on subjective user narratives and lacks verifiable ground truth for properties such as others' intentions. To address this, the authors introduce Fuse, a multi‑agent simulation framework for studying user‑mediated social reasoning. Fuse enables a target agent to interact with simulated users, providing a controllable environment for assessing social reasoning capabilities.
Read original
huggingface/daily-papers