Voice and audio models forevery soundtrack
ElevenLabs narrates, transcribes, dubs, generates sound effects and cleans up noisy voices — all inside the editor, all on one balance.
ElevenLabs v4
ElevenLabs · Voice & Audio
ElevenLabs' flagship text-to-speech: 90+ languages, audio tags, 10,000 characters a request.
ElevenLabs Scribe v2
ElevenLabs · Voice & Audio
ElevenLabs' speech-to-text — word-timed transcripts behind captions, clips and B-roll.
ElevenLabs Dubbing v2
ElevenLabs · Voice & Audio
ElevenLabs' dubbing model — re-voice a video into around 108 languages with re-synced captions.
ElevenLabs Sound Effects
ElevenLabs · Voice & Audio
Describe a sound, get a sound — 0.5 to 22 second effects for any edit.
ElevenLabs Voice Isolator
ElevenLabs · Voice & Audio
Strip background noise and keep the voice — studio-clean audio from a noisy recording.
Compare voice & audio models
Credit prices are read live from the same catalog the app bills from (1 credit = €0.01).
| Model | Best at | Does | Billed per | Credits |
|---|---|---|---|---|
| ElevenLabs' flagship text-to-speech: 90+ languages, audio tags, 10,000 characters a request. | Text → speech | 1k characters | 13 per 1k characters | |
| ElevenLabs' speech-to-text — word-timed transcripts behind captions, clips and B-roll. | Speech → text | Minute | 1 per minute | |
| ElevenLabs' dubbing model — re-voice a video into around 108 languages with re-synced captions. | Dub → other language | Minute | 286 per minute | |
| Describe a sound, get a sound — 0.5 to 22 second effects for any edit. | Text → SFX | Generation | 16 per generation | |
| Strip background noise and keep the voice — studio-clean audio from a noisy recording. | Noise → clean voice | Minute | 16 per minute |
Every model, one app
Shorz isn't a model playground. The models generate the parts; the editor around them turns those parts into a finished, captioned, published video.
- Every model on one prepaid balance — no per-provider subscriptions
- No watermark on anything you generate or export
- Full commercial rights on what you create
- No API keys to manage: the models run on your Shorz account
- New model versions appear in the picker without an app update
- Drivable from Claude, Cursor, Codex or Gemini over MCP
AI voice & audio models questions
No. Preset voices, transcription, dubbing, sound effects and voice isolation all run on your Shorz account. An ElevenLabs key is only needed to narrate with your own cloned voice.
Scribe v2 transcribes the audio with word-level timing, so subtitles land on the spoken word — including after dubbing into another language.
Still have questions? Ask founder
More model families
AI video models
Text-to-video and image-to-video clips for scenes, B-roll and ads.
AI image models
Scene stills, B-roll images and thumbnails, up to 4K.
AI avatar models
Lip-synced talking presenters from one photo and a voice track.
AI language models
The brain of every project: scripts, edit plans and the agent that runs them.
AI music models
Royalty-free background tracks generated from a description.
Every voice & audio model, one picker
Download Shorz for Windows or Mac — every model on this page is already installed and billed from one prepaid balance.
Publishes straight to

