We use essential cookies to make our site work. With your consent, we may also use non-essential cookies to improve user experience and analyze website traffic…

DeepInfra raises $107M Series B to scale the inference cloud — read the announcement

canopylabs logo

canopylabs/

orpheus-3b-0.1-ft

$7.00

/ 1M characters

Orpheus TTS is a state-of-the-art, Llama-based Speech-LLM designed for high-quality, empathetic text-to-speech generation. This model has been finetuned to deliver human-level speech synthesis, achieving exceptional clarity, expressiveness, and real-time streaming performances.

canopylabs/orpheus-3b-0.1-ft cover image

Input

Input text

Text to convert to speech

You need to log in to use this model

Log In

Settings

ServiceTier

The service tier used for processing the request. 'priority' processes the request with higher priority (premium rate); 'flex' processes it at lower priority for a discount, served only when spare capacity exists and may be retried/timed out under load. Both apply only to models that support the respective tier. For compatibility, 'auto' is treated as 'priority' and 'standard_only' as 'default'.

Fail Fast

If true, the request is rejected immediately with HTTP 429 when the model has no spare capacity, instead of waiting in the queue. Opt-in; the default (false) keeps standard queueing behavior.

OrpheusTtsVoice

Select the desired voice for the speech output.

TtsResponseFormat

Select the desired format for the speech output. Supported formats include mp3, opus, flac, wav, and pcm.

Temperature

Temperature of the generation (Default: 0.4, 0 ≤ temperature ≤ 2)

Top P

Top p value for the generation (Default: 0.9, 0 ≤ top_p ≤ 1)

Max Tokens

Maximum number of tokens for the generation (Default: 2000, 0 < max_tokens ≤ 4096)

Repetition Penalty

Repetition penalty for the generation (Default: 1.1, 0 ≤ repetition_penalty)

Stream

Whether to stream audio bytes in chunks

Output

Waiting for audio data... Submit request to start streaming.