DeepInfra raises $107M Series B to scale the inference cloud — read the announcement
By Category
Automatic Speech Recognition
Embeddings
Reranker
Text Generation
Text To Image
Text To Music
Text To Speech
Text To Video
World Model
Zero Shot Image Classification
By Family
/Claude
/DeepSeek
/Flux
/Gemini
/Kimi
/Llama
/Mistral
/Nemotron
/Qwen
Models
Bria/
$0.1400
/ second
Identify and segment objects across video frames using specific coordinate points. Just point in the right direction and the model will figure out by itself which object should be masked.
Have questions or need a custom solution?
Company
Latest Models
Featured Models
© 2026 DeepInfra. All rights reserved.