Cacophony
Contrastive Audio-Text Model
Generate audio variations and extensions with controllable masks
Generate spoken audio from text with optional voice cloning
Piano Cover Generation
Demucs stem separator wrapped in pyharp
Upscale audio to high‑quality 48 kHz