Generate without counting
The voice runs on your GPU: unlimited on every plan. Elsewhere, every sentence has a price.
ClonyVoice runs voice synthesis on your own GPU: unlimited speech, at no extra cost, and your voice never leaves your computer. Then paste a URL — the Studio turns it into a branded video with your cloned voiceover, in about three minutes.
Signups open soon — leave your email and be notified the day it does.
Windows · NVIDIA GPU (CUDA) required · macOS on the roadmap
The voice runs on your GPU: unlimited on every plan. Elsewhere, every sentence has a price.
A voice is biometric data. Voice generation is 100% local — nothing to upload, nothing to leak.
Paste a URL: AI script, your images, your logo, your voice. MP4 rendered on your machine.
Cloning + TTS + videos + dubbing in 10 languages. The equivalent stack costs $50–120/month elsewhere.
A few minutes of audio are enough — recorded or imported, with the speaker's consent.
Text to speech, or URL to Studio video. Listen sentence by sentence, regenerate what you want.
Rendered on your machine, with subtitles and sound design for videos.
Steps 1 and 3 are 100% local. The Studio and translation use our online AI, covered by your AI budget.
No editing skills, no stock-footage collage. The Studio reads your site and builds a video that looks like your brand — because it is your brand.
ClonyVoice analyzes the page: your logo, your images, your colors, your message.
Script, scenes, rhythm and design are composed on our servers — this is what your AI budget pays for.
The voiceover is generated locally with your cloned voice, and the MP4 is rendered on your GPU. Revise the script until it's right.
About 3 minutes from URL to MP4
Cloud voice tools pay for GPU time on every sentence you generate. So they sell credits, meter every character, and bill overages. That's not greed — it's their architecture.
ClonyVoice generates speech on your GPU. Once the app runs, a sentence costs us nothing — so we don't count it. Voice generation is unlimited on every plan, including Free.
The only thing we meter is the only thing that costs us: the server-side AI that writes scripts, directs videos and translates dubbing. That's what your AI budget is — a transparent meter on our costs, never on your voice.
A voice is biometric data. With ClonyVoice, your recordings, your voice models and every audio file you generate live on your computer — nothing is uploaded to clone or to speak. Publishing a voice to the VoiceStore is a separate, explicit choice.
Studio and dubbing use AI on our servers: the text of your scripts, the analysis of the URL you submit, and translation requests. These requests count against your AI budget and are logged for billing and abuse prevention. Your audio is not part of them.
Working offline? Voice generation keeps running without a connection for up to 48 hours between license checks.
ClonyVoice translates your script (on your AI budget), then your cloned voice speaks each language locally, synchronized to the timing of the original audio. English, French, German, Spanish, Italian, Portuguese, Russian, Japanese, Korean, Chinese.
To be precise: we synchronize the audio to the original timing. We do not alter the image or the lips.
Need a voice you don't have? Browse the VoiceStore and download ready-to-use voices, included from the Creator plan.
Voice artist? Publish your voice on your terms, with consent built into the process, and earn when it is downloaded. Your voice becomes an asset — one you control.
복제에서 생성까지, 하나의 플랫폼으로.
단 3초의 오디오로 어떤 목소리의 본질도 포착합니다. 1~5개 샘플로 더 높은 충실도를 달성하세요. 패스트 모드로 즉시 결과 또는 프리사이스 모드로 스튜디오 품질 복제.
원하는 목소리를 설명하면 AI가 구현합니다. 독특한 캐릭터, 브랜드 보이스, 또는 이전에 존재하지 않았던 가상의 페르소나 제작에 완벽합니다.
모든 내장 언어로 말할 수 있는 고품질 음성 라이브러리를 이용하세요. 따뜻한 내레이터부터 에너지 있는 발표자까지, 어떤 프로젝트에도 맞는 음성을 찾을 수 있습니다.
이미 음성 모델이 있으신가요? 바로 가져올 수 있습니다. ClonyVoice는 XTTS, Coqui 등 프레임워크의 주요 형식을 지원합니다.
Assign different voices to each sentence for realistic dialogues. Import scripts from .txt, .srt or .vtt files. Export as audio or as subtitled MP4 video.
생성되는 동안 실시간으로 각 문장을 들을 수 있습니다. 전체를 다시 하지 않고 개별 문장만 재생성. 멀티트랙 타임라인이 있는 비디오 편집기 내장.
실시간 VU 미터로 마이크에서 직접 녹음하세요. 모든 형식의 오디오 파일을 업로드하세요. 또는 YouTube URL을 붙여넣어 자동으로 음성을 추출하세요.
생성한 음성을 암호화된 .clonyvoice 패키지로 저장하세요. 기기 간 안전한 가져오기/내보내기. 테이크 히스토리가 있는 프로젝트 관리.
포괄적인 로컬 API로 ClonyVoice를 워크플로우에 통합하세요. 음성 생성, 보이스 관리, 모든 것을 프로그래밍으로 제어 — 클라우드 의존 없음.
We compare our Creator plan with same-tier monthly plans, no commitment.
| ClonyVoiceCreator — $12/월 | ElevenLabsCreator — $22/월 | FlikiStandard — $28/월 | |
|---|---|---|---|
| Voice generation | Unlimited — runs on your GPU | 121,000 credits/month, overages billed | Credit-based (cloud rendering) |
| Voice cloning | Unlimited cloned voices | Included | Premium plan only ($88/mo) |
| Videos with YOUR images and logo (site analysis) | URL → MP4, ≈ 30 videos/month | — | Generic stock media |
| Dubbing / translation | 10 languages, audio sync | Separate Dubbing product | Not offered |
| Where your voice goes | Never leaves your machine | Sent to and processed in the cloud | Sent to and processed in the cloud |
| Commercial use | ✓ | ✓ | ✓ |
| Runs in the browser | — (Windows app, NVIDIA GPU required) | ✓ | ✓ |
| ¹ Prices and allowances as listed on elevenlabs.io/pricing and fliki.ai/pricing on 6 July 2026 (monthly plans, USD, before tax). Offers change: check the vendors' sites. ² "Unlimited": voice synthesis runs locally on your GPU; Studio video and translation usage is covered by your monthly AI budget (≈ 30 videos/month on Creator at current rates). ³ ClonyVoice requires Windows 10/11 and a CUDA-compatible NVIDIA GPU. macOS is on the roadmap. |
|||
Creators typically stack $50–120 a month of separate tools to do what ClonyVoice does from $12/month.
| Voice cloning & text-to-speech | $6–99 per month, metered by credits |
| Faceless & marketing videos | $25–88 per month, minutes capped |
| Dubbing & translation | $25–60 per month |
| Open voice marketplace | No mainstream equivalent |
| ClonyVoice: all four, from $12/month — with unlimited local voice generation. | |
Published prices of leading tools in each category, observed July 2026.
From August 2, 2026, the AI Act (art. 50) requires synthetic audio and video to be disclosed. ClonyVoice was built for that world: consent-first cloning, content labeling built in, by a French company. Create with a tool designed for the rules — before they apply.
Clone your voice today, free. Two voices, about 3 Studio videos a month, and voice generation you'll never have to count.
One email at launch. Nothing else, no sharing.