An average rating of perceived quality from human listeners, commonly used to evaluate speech synthesis systems.