Live video experiences
Connect video generation to a player and live controls, so a scene responds while it runs. Add speech and it becomes an avatar.
Real-time model APIs that run together
Speech, language, and video models available as real-time model APIs for voice agents, live video, and avatars. Chain several in one Session on one machine, with no network trip between them.
Explore the models
Models tuned for latency-sensitive work. Point the client you run today at a uRun model API. Decreases in latency translate to shorter wait times for users and cheaper infrastructure costs for your business.
99-language speech recognition with native translation into English text
Request accessStreaming text-to-speech in nine languages with cancelable utterances
Request accessStreaming transcription in English and French with turn detection built in
Request accessFull-duplex speech-to-speech that listens and speaks at once, with barge-in
Request accessContinuous prompt-steered video that keeps rolling, steered as it plays
Request accessChunked realtime video with live prompt controls and cuts between chunks
Request accessA compact 7B coder for everyday code assistance, completions to edits
Request accessLong-context chat and reasoning across 262K tokens, thinking on by default
Request accessLarge-scale coding on a 480B mixture-of-experts with 35B active parameters
Request accessVoice pipeline
Speech recognition. A 27B language model. Speech synthesis.
Whisper large-v3Qwen3.8-27BKokoro 82M run together on one machine, with no queue or network trip between them. You connect to the pipeline the same way you connect to a single model.
Use cases
Two applications built from the models, one for live video and one for speech across languages. Each keeps the models separate from your application logic.
Connect video generation to a player and live controls, so a scene responds while it runs. Add speech and it becomes an avatar.
Turn incoming speech into English text, with optional spoken output. See the models and stages you need.
Labs
The platform is open for model labs. Work with us to bring your models to developers, or have our engineers tune and serve your own.
Make your model available to developers on uRun. Your weights stay under your control, and distribution is opt-in.
Discuss a model partnershipHave your own model or pipeline? Run it on uRun and work with our engineers on the configuration your application needs.
Discuss custom model servingBuild with uRun
Explore the catalog and request access to the models your application needs, across speech, language, and video.