Language models produce incoherent probability forecasts that violate basic logical consistency rules, meaning you shouldn't rely on them for probabilistic predictions about real-world events without additional safeguards.
This paper tests whether language models produce coherent probability forecasts by using a mathematical technique called Dutch books—finding profitable bets against the model's predictions.