papersTODAY 04:00 UTC
Speech Generation Speaker Poisoning Aims to Block Zero-Shot Voice Cloning
A new arXiv paper introduces Speech Generation Speaker Poisoning (SGSP), a task focused on stopping zero-shot text-to-speech models from reproducing specific target voices. The work targets systems that can clone unseen speakers from just a few seconds of audio, framing speaker-level capability erasure as a defensive goal.