Directions in a neural network's internal activation space that represent human-interpretable concepts like accent or age.