Glossary → Voice & AI audio
Voice & AI audio
Source separation
Source separation splits mixed audio into components — vocals from music, one speaker from another — using neural models trained on the task.
'Music/vocal split' became consumer-grade, and speech separation improves transcription of cross-talk. Separation quality degrades with similarity: two similar voices are far harder than voice-versus-music.