connecting…
RTX 3090
▸ Voice Design — Text to Speech
male female child teenager young adult middle-aged elderly low pitch moderate pitch high pitch very low pitch very high pitch whisper american accent british accent australian accent canadian accent indian accent chinese accent
Tip: Click pills above to add attributes, or type freely and press Enter/comma to confirm each tag. Language is auto-detected from your text — just write in any language.
32
1.00
✓   Audio Generated
▸ Zero-Shot Voice Cloning
🎵
Drag & drop or click to upload reference audio
32
1.00
✓   Voice Cloned
▸ System Health
API
—
CUDA
—
GPU
—
VRAM Total
—
VRAM Free
—
Model
—
▸ API Reference
MethodEndpointBodyReturns
GET/health—JSON — GPU + model status
POST/synthesizeJSONaudio/wav — voice design TTS
POST/clonemultipart/form-data + fileaudio/wav — voice-cloned speech
GET/docs—Swagger UI (auto-generated)
▸ Voice Attributes Reference
# Gender male, female
# Age child, young, middle-aged, elderly
# Pitch very low, low, normal, high, very high
# Style whisper
# Accent american accent, british accent, australian accent

# Example combinations:
instruct = "female, low pitch, british accent"
instruct = "male, elderly, american accent"
instruct = "female, child, high pitch, whisper"
▸ cURL Examples
# Synthesize
curl -X POST http://164.52.194.46:8007/synthesize \
  -H "Content-Type: application/json" \
  -d '{"text":"Hello world","instruct":"female, british accent","num_step":32}' \
  --output output.wav

# Voice clone
curl -X POST http://164.52.194.46:8007/clone \
  -F "text=Hello from OmniVoice" \
  -F "reference_audio=@/path/to/ref.wav" \
  --output cloned.wav
▸ Python Example
import requests

# Synthesize with voice attributes
r = requests.post(
  "http://164.52.194.46:8007/synthesize",
  json={"text": "Hello", "instruct": "female, british accent", "num_step": 32})
open("out.wav", "wb").write(r.content)

# Voice clone
r = requests.post(
  "http://164.52.194.46:8007/clone",
  data={"text": "Hi there"},
  files={"reference_audio": open("ref.wav", "rb")})
open("cloned.wav", "wb").write(r.content)