Skip to content

Towards Benchmarking Situational Awareness of Large Language Models: Comprehensive Benchmark, Evaluation and Analysis.

Guo Tang, Zheng Chu, Wenxiang Zheng, Ming Liu, Bing Qin

VenueA*EMNLP
Year2024
ProceedingsEMNLP (Findings)

Browse the full EMNLP paper archive.