The field shifted from asking 'why did the model do that?' to 'how do we control what it does?'—and truthfulness is now the primary concern for LLM research, growing from absent to 37% of papers in just four years.
This paper analyzes six years of the TrustNLP workshop (2021-2026), tracking how AI safety research evolved from explaining static models to controlling generative systems.