DeepInfra raises $107M Series B to scale the inference cloud — read the announcement

canopylabs/
$7.00
/ 1M characters
Orpheus TTS is a state-of-the-art, Llama-based Speech-LLM designed for high-quality, empathetic text-to-speech generation. This model has been finetuned to deliver human-level speech synthesis, achieving exceptional clarity, expressiveness, and real-time streaming performances.

Input text
Text to convert to speech
You need to log in to use this model
Log InSettings
ServiceTier
The service tier used for processing the request. 'priority' processes the request with higher priority (premium rate); 'flex' processes it at lower priority for a discount, served only when spare capacity exists and may be retried/timed out under load. Both apply only to models that support the respective tier. For compatibility, 'auto' is treated as 'priority' and 'standard_only' as 'default'.
Fail Fast
If true, the request is rejected immediately with HTTP 429 when the model has no spare capacity, instead of waiting in the queue. Opt-in; the default (false) keeps standard queueing behavior.
OrpheusTtsVoice
Select the desired voice for the speech output.
TtsResponseFormat
Select the desired format for the speech output. Supported formats include mp3, opus, flac, wav, and pcm.
Temperature
Temperature of the generation (Default: 0.4, 0 ≤ temperature ≤ 2)
Top P
Top p value for the generation (Default: 0.9, 0 ≤ top_p ≤ 1)
Max Tokens
Maximum number of tokens for the generation (Default: 2000, 0 < max_tokens ≤ 4096)
Repetition Penalty
Repetition penalty for the generation (Default: 1.1, 0 ≤ repetition_penalty)
Stream
Whether to stream audio bytes in chunks
Waiting for audio data... Submit request to start streaming.
© 2026 DeepInfra. All rights reserved.