Frontier models communicate inefficiently with themselves across information asymmetry, extracting only ~0.93 bits per question instead of the theoretical 1 bit, with failures driven equally by answer errors and inability to discriminate between candidates.
This paper evaluates six frontier language models playing a communication game where one model asks yes/no questions to identify a target Wikipedia article from a set, while another answers with single words.