XiaoMin β Female Chinese AI voice
XiaoMin is a female AI voice for Chinese (Cantonese, Simplified). Use XiaoMin for text-to-speech narration, voice cloning, UGC video creation, AI avatar videos, music generation, sound effects, and multilingual content creation. Integrate via API with voice ID xiaomin-yue-cn.
Voice sample
Listen to XiaoMin's female Chinese voice. This is the same quality you get with our text-to-speech, voice cloning, and video generation tools.
Voice details
- Voice name
- XiaoMin
- Gender
- Female
- Language
- Chinese (Cantonese, Simplified)
- API Voice ID
xiaomin-yue-cn- Sample URL
- Download MP3
What you can create with XiaoMin
XiaoMin's female Chinese voice works across every Verbatik AI tool. From text-to-speech to video generation, music creation to voice cloning β one voice, unlimited possibilities.
Social media content
Generate TikTok, Instagram Reels, and Shorts with XiaoMin's Chinese voiceover. Create viral-ready content with authentic Cantonese, Simplified Chinese narration that resonates with native speakers and boosts your social media presence.
E-learning with XiaoMin
Build online courses, training modules, and educational content narrated by XiaoMin in Chinese. XiaoMin's clear female voice ensures learners in Cantonese, Simplified Chinese can follow along easily, improving comprehension and retention.
Audiobooks in Chinese
Produce full-length audiobooks narrated by XiaoMin in Chinese. XiaoMin's female voice brings stories to life with natural pacing, emotional range, and authentic Cantonese, Simplified Chinese pronunciation for an engaging listening experience.
Presentations & demos
Create professional presentations and product demos with XiaoMin's female Chinese narration. Add voiceovers to slides, screen recordings, and walkthrough videos for Cantonese, Simplified Chinese business audiences.
Use XiaoMin via API
Integrate XiaoMin's female Chinese voice into your application with a single API call. Use voice ID xiaomin-yue-cn for text-to-speech, voice cloning, video generation, and more.
Our API supports real-time streaming, batch processing, SSML markup, and multiple output formats (MP3, WAV, OGG, FLAC). Average latency is under 75ms for the first byte, making it suitable for interactive applications, chatbots, and IVR systems.
Supported operations:
- Text-to-Speech
- Voice Cloning
- Video Generation
- Avatar Creation
- Real-time Streaming
- Batch Processing
curl -X POST https://api.verbatik.com/v1/tts \
-H "Authorization: Bearer YOUR_API_KEY" \
-H "Content-Type: application/json" \
-d '{
"voice_id": "xiaomin-yue-cn",
"text": "Hello, this is XiaoMin speaking in Chinese.",
"output_format": "mp3"
}'XiaoMin voice β FAQ
Ready to bring your content to life?
Join 150,000+ creators, developers, and businesses using Verbatik AI to produce studio-quality voiceovers, clone voices, and generate music and sound effects.
- 1,500+ neural voices in 150+ languages
- Voice cloning from a single audio sample
- Music, sound effects, and video generation
- Full API access with real-time streaming
- Commercial license included
- 14-day money-back guarantee
150K+
creators
150+
languages
75ms
latency
Trusted by teams at leading companies worldwide