Clonación de voz

10 a 30 segundos Multilingüe En tiempo real
Arrastre y suelte su archivo de audio aquí

o haga clic para navegar

Soporta MP3, WAV, FLAC, M4A, OGG (máx. 100MB)

Click to start recording

Record 10-30 seconds of clear speech


Audio Preview
Duration: --
Voice Settings
Your Credits
0
Get more credits
Tips for Best Results
  • Use 10-30 seconds of clear audio
  • Minimal background noise
  • Natural speaking pace
  • Varied intonation helps quality
  • Single speaker only
My Cloned Voices
0

No voices cloned yet.
Upload audio to create your first voice.

How Voice Cloning Works

1
Upload Sample

Provide 10-30 seconds of clear speech from the voice you want to clone.

2
AI Analysis

Our AI extracts voice characteristics, tone, accent, and speaking patterns.

3
Voice Model

A personalized voice model is created and saved to your account.

4
Generate Speech

Use your cloned voice to generate speech from any text, in any language.

Available Models

Chatterbox
Recommended
  • Zero-shot cloning
  • 23 languages
  • 6 to 60 second samples
  • Commercial use allowed
OpenVoice v2
Fast
  • Fastest processing
  • Tone control
  • Style transfer
  • Lightweight
GPT-SoVITS
Detailed
  • Voice cloning
  • Multilingual
  • Slower, more detailed
  • Open weights
Tortoise TTS
Highest detail
  • Voice cloning
  • English
  • Slow, high detail
  • Open weights

Clone Voices via API

Integrate voice cloning into your applications with our simple REST API.

API Documentation
curl -X POST https://api.speechtospeechai.com/v1/voices/clone \
  -H "Authorization: Bearer YOUR_API_KEY" \
  -F "name=My Voice" \
  -F "file=@sample.mp3" \
  -F "model=chatterbox"

# Response
{
  "id": "voice_abc123",
  "name": "My Voice",
  "model": "chatterbox",
  "status": "ready",
  "created_at": "2026-02-03T12:00:00Z"
}