Integrate ultra-natural Text-to-Speech (TTS), Instant Voice Cloning, Speech-to-Text (STT), Subtitle SRT sync, and AI Dubbing into your applications with just a few lines of code.
Select a capability and language to view production-ready request and response payloads.
Synthesize speech from text using 3,000+ studio-grade 48kHz voices with Natural, Narration, or Dramatic style presets.
curl -X POST "https://api.yupvox.com/v1/tts" \
-H "Authorization: Bearer sk-yupvox-YOUR_API_KEY" \
-H "Content-Type: application/json" \
-d '{
"voiceId": "DFE237",
"text": "Hello, this is a high-quality AI voice generated by the YupVox API.",
"speed": 1.0,
"pitch": 0
}'{
"success": true,
"data": {
"jobId": "tts_82910_1725184900000",
"historyId": 82910,
"status": "processing",
"voiceId": "DFE237",
"charCount": 62,
"estimatedTimeSec": 1.2
}
}All cutting-edge voice and audio AI algorithms unified under a single, robust REST API.
3,000+ studio voices with natural pacing, audiobooks narration, and emotional drama styles.
Instant zero-shot voice cloning. Dual-engine architecture with Omega (multilingual) & Alpha (ultra-fast VN).
High-accuracy automatic speech recognition with timestamps, punctuation, and multi-speaker support.
Convert SRT subtitle text into perfectly timed audio segments aligned with original video frames.
Transform speaker timbre and tone into target AI voices while retaining nuance and cadence.
End-to-end automated pipeline: Transcribe → Translate → Synthesize localized voice audio.
Start generating high-quality speech in under 3 minutes.
Sign up for free and get your secret API key (`sk-yupvox-...`) with 50,000 complimentary credits.
Call standard REST endpoints using any language (cURL, Python, Node.js, Go) with Bearer token authentication.
Instantly download high-bitrate MP3/WAV audio via high-speed CDN ready for streaming or download.
Trusted by creators, SaaS platforms, media publishers, and conversational AI developers.
Generate thousands of daily automated videos with captivating, viral-ready AI voiceovers.
Sub-second latency enabling natural conversational AI agents, smart IVR, and customer support bots.
Convert entire article repositories and e-books into studio-grade audiobooks with narrative pacing.
Automatically dub educational courses into dozens of global languages for international students.
Your credit balance is shared seamlessly between the web studio and API calls. No setup fees, no per-request surcharge.
Simple 1:1 transparent ratio: 1 Credit = 1 Character generated (TTS) or transcribed (STT). API requests draw from your account credit balance with no hidden API fees.
YupVox GPU cloud infrastructure uses distributed worker queues, allowing high concurrent requests without artificial bottlenecks.
Demo preview mode responds in 0.5s - 1.5s. Full production conversions process via asynchronous background queues within a few seconds depending on text length.
Yes! Upon creating an account, you receive 50,000 free trial credits to test all voice models and studio tools.
100% data security. Cloned voice models are strictly encrypted and tied to your account ID. You can also permanently delete history records and files via API.
Create your account today to receive 50,000 free credits and integrate studio-grade AI voices into your pipeline.