
Inworld Text To Speech
by InworldUltra-realistic, low-latency voice cloning supports 11 languages, instant & professional cloning, 48 kHz audio, fine emotional control, API access—ideal for dynamic, expressive AI interactions.
inworld-tts-1Input
Output
Unknown content type
Looking for something different?
Models that trade off differently against the one you are viewing
Lower cost per run
Open-source models we host, unlimited on the $149/month plan.
About Inworld Text To Speech
Ultra-realistic, low-latency voice cloning supports 11 languages, instant & professional cloning, 48 kHz audio, fine emotional control, API access—ideal for dynamic, expressive AI interactions.
Technical Specifications
- Model ID
- inworld-tts-1
- Provider
- Inworld
- Category
- Audio Models
- Task
- Audio Generation
- Price
- $6 per million characters
- Added
- August 7, 2025
Key Features
- AI voice synthesis and text-to-speech
- Multiple language and accent support
- Voice cloning from short audio samples
- Real-time audio processing via API
- Customizable speech parameters
Quick Start
Integrate Inworld Text To Speech into your application with a single API call. Get your API key from the pricing page to get started.
import requestsimport jsonurl = "https://modelslab.com/api/v7/voice/text-to-speech"headers = {"Content-Type": "application/json"}data = {"model_id": "inworld-tts-1","prompt": "your prompt here","key": "YOUR_API_KEY"}try:response = requests.post(url, headers=headers, json=data)response.raise_for_status() # Raises an HTTPError for bad responses (4XX or 5XX)result = response.json()print("API Response:")print(json.dumps(result, indent=2))except requests.exceptions.HTTPError as http_err:print(f"HTTP error occurred: {http_err} - {response.text}")except Exception as err:print(f"Other error occurred: {err}")
Pricing
Inworld Text To Speech API costs $6 per million characters. Plans start at $21/month, and open-source models are unlimited on the $149/month plan. View pricing plans
Use Cases
- Voice-over production for video content
- Podcast and audiobook narration
- Multilingual customer support automation
- Interactive voice response (IVR) systems


