Generate speech, music, and sound effects from text prompts or video using state-of-the-art audio models. These models support text-to-audio, voice cloning, and video-to-music workflows.Browse the models in this section to see what each partner offers, or check Pricing for per-model rates.