A request to generate speech in dataset.
Attributes:
language: the language of the text to be generated
text: the text for which audio is to be generated
voice: the voice to use to generate the audio
model: the model to use to generate the audio
format: the output format of the audio file