A RESTful service for high-quality text-to-speech using Qwen3 and specialized voice cloning. Optimized for reusing a specific voice prompt to avoid re-computation.
Scanned 9/7/2026
Install to Claude Code
npx -y skills add modbender/skill-library-mcp --skill chichi-speech --agent claude-codeInstalls into .claude/skills of the current project.
Are you the author of Chichi Speech?
Add the live security badge to your README — it updates automatically with every re-scan.
[](https://www.skillsdirectory.com/skills/modbender-chichi-speech)More formats (shields.io, HTML) on the badges page.
---
name: chichi-speech
description: A RESTful service for high-quality text-to-speech using Qwen3 and specialized voice cloning. Optimized for reusing a specific voice prompt to avoid re-computation.
---
# Chichi Speech Service
This skill provides a FastAPI-based REST service for Qwen3 TTS, specifically configured for reusing a high-quality reference audio prompt for efficient and consistent voice cloning. This service is packaged as an installable CLI.
## Installation
Prerequisites: `python >= 3.10`.
```bash
pip install -e .
```
## Usage
### 1. Start the Service
The service runs on port **9090** by default.
```bash
# Start the server (runs in foreground, use & for background or a separate terminal)
# Optional: Uudate to your own reference audio and text for voice cloning
chichi-speech --port 9090 --host 127.0.0.1 --ref-audio "https://qianwen-res.oss-cn-beijing.aliyuncs.com/Qwen3-TTS-Repo/clone_2.wav" --ref-text "Okay. Yeah. I resent you. I love you. I respect you. But you know what? You blew it! And thanks to you."
```
### 2. Verify Service is Running
Check the health/docs:
```bash
curl http://localhost:9090/docs
```
### 3. Generate Speech
Use cURL:
```bash
curl -X POST "http://localhost:9090/synthesize" \
-H "Content-Type: application/json" \
-d '{
"text": "Nice to meet you",
"language": "English"
}' \
--output output/nice_to_meet.wav
```
## Functionality
- **Endpoint**: `POST /synthesize`
- **Default Port**: 9090
- **Voice Cloning**: Uses a pre-computed voice prompt from reference files to ensure the cloned voice is consistent and generation is fast.
## Requirements
- Python 3.10+
- `qwen-tts` (Qwen3 model library)
- Access to a reference audio file for voice cloning.
- By default, it uses public sample audio from Qwen3.
- **CRITICAL**: You can provide your own reference audio using the `--ref-audio` and `--ref-text` flags.
Is this your skill, or is something wrong with this listing? Request removal or report an issue. Author removals are honored within 72 hours.
No comments yet. Be the first to comment!