💻 computer science

Evaluating Large language models on Understanding Korean indirect Speech acts

本研究通过构建专门的数据集并采用自动化与人工评估相结合的方法,评估了各种大语言模型理解韩语间接言语行为的能力,结果表明,尽管像 Claude3-Opus 这样的闭源模型优于开源替代方案,但目前尚无模型能在解释依赖于语境的意图方面达到人类水平。

Youngeun Koo, Jiwoo Lee, Dojun Park, Seohyun Park, Sungeun Lee2026-08-10