NVIDIA Fugatto 1Research model for free-form audio synthesis and transformation from text and optional audio.
Qinyun琴韵音乐大模型TME music model applied to AI singing and voice workflows.Shenzhen, Mainland China, China
RAVERealtime Audio Variational autoEncoder for fast neural waveform synthesis and timbre transfer.Paris, France
RIMA AI리마 AIEmotionwave engine for music analysis, generation, recommendation, and automated performance systems.South Korea
Seed-MusicByteDance suite for controllable song generation, score-to-audio editing and zero-shot singing voice conversion.Beijing, Mainland China, China
SongGeneration 2 Large 4BLeVo 2 open checkpoint for multilingual song generation from text and audio prompts.
Stability AI Stable Audio 3 APIAsynchronous hosted API for Stable Audio 3 text-to-audio, audio-to-audio and inpainting.
Stable Audio 3.0 LargeStable Audio 3 exact tier for low-latency high-volume generation, text-to-audio, audio-to-audio, inpainting.