The dataset includes 237 speakers, covering a variety of voice qualities such as mature female, middle-aged male, bass, falsetto, etc., and spans across young, middle-aged, and elderly ages. The audio is clear and natural, which can greatly enhance the naturalness and expressiveness of the model.