This page can copy a voice from about ten seconds of speech. That is a genuinely useful tool and a genuinely abusable one. The rules below are the deal for using it.
-
Everything runs on this device.
Your microphone audio, your reference clip and the speech you generate stay in this browser tab. They are never uploaded. The model files are downloaded from Hugging Face and cached locally; after that the page works offline.
-
Clone your own voice.
The only voice you may clone here is your own — or one where you hold explicit, documented permission from the person it belongs to, for this specific use. There is no third category.
-
Do not use it to deceive.
No impersonating real people. No putting words in someone's mouth. No fraud, no harassment, no defeating voice-based identity checks at a bank, a helpdesk or anywhere else.
-
Label what you publish.
If synthetic speech from this page leaves your machine, say that it is AI-generated. Downloads from here are named accordingly.
-
The law on this is real, and it varies.
Voice likeness is protected by right-of-publicity and anti-impersonation statutes in many US states, by biometric-privacy laws such as Illinois BIPA, and by transparency duties under the EU AI Act. Working out what applies to you is your job. This is a technical demo, not legal advice, and its author is not your lawyer.
-
Responsibility sits with you.
If you clone somebody else's voice without their permission, that is on you — not on whoever deployed this demo, not on Kyutai who trained the model, not on the authors of the open-source libraries it runs on. The page is provided as is, with no warranty of any kind, and its operator accepts no liability for anything you generate with it.