Developer API v1Flat 30 Credits / CallRate Limit: 60 req/min

NouAI Voice Synthesis API

Integrate ultra-realistic Lao and Thai speech directly into your apps, bots, backend services, and workflows with simple HTTP REST requests.

POSThttps://nouai.app/api/v1/tts/speak
Top Up Credits

AI-Ready DocumentationLLM Friendly

Need an AI agent to write the integration code for you? Click the button to copy the full Markdown specification.

TTS Models & Character Limits

Choose between NouAI Speech v2 (fast & expressive) or NouAI Speech v1 (extended 4,000 characters long-form). Both cost flat 30 credits.

Default Model

NouAI Speech v2nouai-speech-v2

Breeze TTS 2 Engine
1,500
Chars / Call

High-fidelity conversational voice synthesis with 20 studio voices and expressive vocal event tags (like laughing, sighing).

Character Limit:1,500 characters
Voices:20 Studio Voices (Alice, Daniel, etc.)
Voice Cloning:Supported (Zero-Shot)
Cost:30 Credits flat
Click to view snippets ↓
Extended Long Context

NouAI Speech v1nouai-speech-v1

OmniVoice Engine
4,000
Chars / Call

Engineered for long-form scripts, news broadcasts, storytelling, and educational content. Synthesizes up to 4,000 characters in a single request!

Character Limit:4,000 characters
Voices:default_voice, female, female_2, teenneger, teenneger_2
Long Content:Ideal for long paragraphs & articles
Cost:30 Credits flat
Click to view snippets ↓
Code Snippets (NouAI Speech v2 1,500 Chars)
curl -X POST https://nouai.app/api/v1/tts/speak \
  -H "Authorization: Bearer YOUR_API_KEY" \
  -H "Content-Type: application/json" \
  -d '{
    "text": "Nyob zoo sawv daws, zoo siab txais tos nej tuaj rau NouAI Studio.",
    "model": "nouai-speech-v2",
    "voice": "alice",
    "speed": 1.0
  }' \
  --output speech.wav

Request Parameterslanguage is automatic

Endpoint: POST /api/v1/tts/speak • Flat 30 credits per request

Default EngineMax 1,500 chars

NouAI Speech v2nouai-speech-v2

Breeze TTS 2 Engine
Language: Auto

Optimized for natural conversational speech, voicebots, and dialogue. Supports 20 studio voices and inline vocal tags like (laugh), (sigh).

FieldTypeReqDescription
textstringYesText to synthesize. Max 1,500 chars.
modelstringNo"nouai-speech-v2" (default if omitted).
voicestringNoPreset voice: alice, aria, sarah, daniel, charlie, etc. (20 studio voices).
speednumberNoPlayback speed multiplier (0.5 to 2.0). Default 1.0.
Billing: 30 credits flat
Long ContextMax 4,000 chars

NouAI Speech v1nouai-speech-v1

OmniVoice Engine
Language: Auto

Engineered for long articles, news reading, audiobooks, and storytelling. Allows up to 4,000 characters per single synthesis request without cutting.

FieldTypeReqDescription
textstringYesText to synthesize. Max 4,000 chars for long-form context.
modelstringYesMust specify "nouai-speech-v1" to enable 4,000 chars.
voicestringNoPreset voice: default_voice, female, female_2, teenneger, teenneger_2.
speednumberNoPlayback speed multiplier (0.5 to 2.0). Default 1.0.
Billing: 30 credits flat
Voice Cloning • Zero-ShotSupported on Both Models

Audio Reference Parameters (ref_audio)

Provide 3–15 seconds of clean speech audio to dynamically clone and mimic any speaker's voice. When reference audio is provided, it automatically overrides the preset voice parameter.

FieldTypeReqDescription
ref_audio_base64stringNo*Base64-encoded audio bytes of the target speaker (WAV, MP3, FLAC, M4A, OGG). Max 10MB. 3–15 seconds of clear speech recommended.
ref_audio_suffixstringNoAudio file extension: ".wav" (default), ".mp3", ".flac", ".m4a", ".ogg".
ref_audiostringNo*Direct public URL (https://...) or Data URI (data:audio/wav;base64,...) pointing to reference audio. Alternative to ref_audio_base64.
ref_textstringNoExact transcript text of words spoken in the reference audio. (Highly recommended — improves voice similarity and pronunciation).
* Provide either ref_audio_base64 or ref_audio to activate voice cloning.Billing: 30 credits flat

Response Headers & Status Codes

HTTP Status 200 = OK
HeaderDescription
Content-Typeaudio/wav binary audio stream.
X-Audio-Duration-SecondsTotal audio length in seconds.
X-Credits-ChargedCredits deducted (always 30).
X-Credits-RemainingRemaining account balance.
X-RateLimit-RemainingCalls remaining in current 1-min window.
HTTP Status Codes
200 OKSuccessful synthesis. Returns binary WAV audio stream.
400Bad Request or text limit exceeded (1,500 chars for nouai-speech-v2, 4,000 for nouai-speech-v1).
401Unauthorized. Missing, invalid, or revoked API key.
402Insufficient Credits (< 30 credits balance remaining).
429Rate limit exceeded (> 60 requests per minute).
502Synthesis service error. Charged credits are automatically refunded.

Live API Tester

Test synthesis directly from your browser. Flat 30 credits deducted per call.

65 / 1,500 chars