Clonev
by @instant-picture
Clone any voice and generate speech using Coqui XTTS v2. SUPER SIMPLE - provide a voice sample (6-30 sec WAV) and text, get cloned voice audio. Supports 14+ languages. Use when the user wants to (1) Clone their voice or someone else's voice, (2) Generate speech that sounds like a specific person, (3) Create personalized voice messages, (4) Multi-lingual voice cloning (speak any language with cloned voice).
clawhub install clonevπ About This Skill
name: clonev description: Clone any voice and generate speech using Coqui XTTS v2. SUPER SIMPLE - provide a voice sample (6-30 sec WAV) and text, get cloned voice audio. Supports 14+ languages. Use when the user wants to (1) Clone their voice or someone else's voice, (2) Generate speech that sounds like a specific person, (3) Create personalized voice messages, (4) Multi-lingual voice cloning (speak any language with cloned voice).
CloneV Skill - Voice Cloning Made Simple
β οΈ CRITICAL INSTRUCTIONS FOR AI MODELS
DO NOT try to use Docker containers directly.
DO NOT try to interact with coqui-xtts container - it is broken and restarting.
DO NOT try to use APIs or servers.
ONLY USE THE SCRIPT: scripts/clonev.sh
The script handles everything automatically. Just call it with text, voice sample, and language.
What This Skill Does
Clones any voice from a short audio sample and generates new speech in that voice.
Input:
Output: OGG voice file (cloned voice speaking the text)
Works with: Any voice! Yours, a celebrity, a character, etc.
The ONE Command You Need
$(scripts/clonev.sh "Your text here" /path/to/voice_sample.wav language)
That's it! Nothing else needed.
Step-by-Step Usage (FOR AI MODELS)
Step 1: Get the required inputs
en)Step 2: Run the script
VOICE_FILE=$(scripts/clonev.sh "TEXT_HERE" "/path/to/sample.wav" LANGUAGE)
Step 3: Use the output
The variable$VOICE_FILE now contains the path to the generated OGG file.Complete Working Examples
Example 1: Clone voice and send to Telegram
# Generate cloned voice
VOICE=$(/home/bernie/clawd/skills/clonev/scripts/clonev.sh "Hello, this is my cloned voice!" "/mnt/c/TEMP/Recording 25.wav" en)Send to Telegram (as voice message)
message action=send channel=telegram asVoice=true filePath="$VOICE"
Example 2: Clone voice in Czech
# Generate Czech voice
VOICE=$(/home/bernie/clawd/skills/clonev/scripts/clonev.sh "Ahoj, tohle je mΕ―j hlas" "/mnt/c/TEMP/Recording 25.wav" cs)Send
message action=send channel=telegram asVoice=true filePath="$VOICE"
Example 3: Full workflow with check
#!/bin/bashGenerate voice
VOICE=$(/home/bernie/clawd/skills/clonev/scripts/clonev.sh "Task completed!" "/path/to/sample.wav" en)Verify file was created
if [ -f "$VOICE" ]; then
echo "Success! Voice file: $VOICE"
ls -lh "$VOICE"
else
echo "Error: Voice file not created"
fi
Common Language Codes
| Code | Language | Example Usage |
|------|----------|---------------|
| en | English | scripts/clonev.sh "Hello" sample.wav en |
| cs | Czech | scripts/clonev.sh "Ahoj" sample.wav cs |
| de | German | scripts/clonev.sh "Hallo" sample.wav de |
| fr | French | scripts/clonev.sh "Bonjour" sample.wav fr |
| es | Spanish | scripts/clonev.sh "Hola" sample.wav es |
Full list: en, cs, de, fr, es, it, pl, pt, tr, ru, nl, ar, zh, ja, hu, ko
Voice Sample Requirements
Good samples:
Bad samples:
β οΈ Important Notes
Model Download
/mnt/c/TEMP/Docker-containers/coqui-tts/models-xtts/Processing Time
Troubleshooting
"Command not found"
Make sure you're in the skill directory or use full path:/home/bernie/clawd/skills/clonev/scripts/clonev.sh "text" sample.wav en
"Voice sample not found"
/)ls -la /path/to/sample.wav"Model not found"
The model should auto-download. If not:cd /mnt/c/TEMP/Docker-containers/coqui-tts
docker run --rm --entrypoint "" \
-v $(pwd)/models-xtts:/root/.local/share/tts \
ghcr.io/coqui-ai/tts:latest \
python3 -c "from TTS.api import TTS; TTS('tts_models/multilingual/multi-dataset/xtts_v2')"
Poor voice quality
Quick Reference Card (FOR AI MODELS)
USER: "Clone my voice and say 'hello'"
β Get: sample path, text="hello", language="en"
β Run: VOICE=$(/home/bernie/clawd/skills/clonev/scripts/clonev.sh "hello" "/path/to/sample.wav" en)
β Result: $VOICE contains path to OGG file
β Send: message action=send channel=telegram asVoice=true filePath="$VOICE"
USER: "Make me speak Czech"
β Get: sample path, text="Ahoj", language="cs"
β Run: VOICE=$(/home/bernie/clawd/skills/clonev/scripts/clonev.sh "Ahoj" "/path/to/sample.wav" cs)
β Send: message action=send channel=telegram asVoice=true filePath="$VOICE"
Output Location
Generated files are saved to:
/mnt/c/TEMP/Docker-containers/coqui-tts/output/clonev_output.ogg
The script returns this path, so you can use it directly.
Summary
1. ONLY use the script: scripts/clonev.sh
2. NEVER try to use Docker containers directly
3. NEVER try to interact with the coqui-xtts container
4. Script handles everything automatically
5. Returns path to OGG file ready to send
Simple. Just use the script.
*Clone any voice. Speak any language. Just use the script.*
π Tips & Best Practices
"Command not found"
Make sure you're in the skill directory or use full path:/home/bernie/clawd/skills/clonev/scripts/clonev.sh "text" sample.wav en
"Voice sample not found"
/)ls -la /path/to/sample.wav"Model not found"
The model should auto-download. If not:cd /mnt/c/TEMP/Docker-containers/coqui-tts
docker run --rm --entrypoint "" \
-v $(pwd)/models-xtts:/root/.local/share/tts \
ghcr.io/coqui-ai/tts:latest \
python3 -c "from TTS.api import TTS; TTS('tts_models/multilingual/multi-dataset/xtts_v2')"