Audio module
AI voice and transcription, every model, one subscription.
Omnibo gives you 29 AI audio models from 8 providers — MiniMax Speech, Seed TTS, Gemini TTS, GPT-4o mini TTS, Grok Voice and transcription — on a single $29.99/month subscription. Speech generation is metered per million characters of input and transcription per minute of audio, and both figures are shown before the run. A standalone voice subscription starts at about $22/month and covers voice alone.
What you can do in the Omnibo audio module
The Omnibo audio module covers both directions: text into speech, and recorded audio into text. Both are metered in the same balance as everything else you make.
for 52 characters
Rebuilt in HTML from the live components — “MiniMax Speech 2.8 HD, 120,000 cr / M characters” is real text on this page, not an image.
Speech
Turn a script into narration on any voice model your plan reaches, up to 9,999 characters a run.
Voices
Pick a voice per run, so the same script can be auditioned in several deliveries before you commit to one.
Transcription
Turn recorded audio into text, metered per minute rather than per character.
Carry forward
Use a generated voice track as the reference audio on a video run without leaving Omnibo.
Every audio model in Omnibo, and what a minute of speech costs
Omnibo carries 29 audio models across 8 providers. Speech models are metered per million characters of input text; transcription models are metered per minute of audio. Both units are shown per row rather than averaged into one number. ACE-Step costs 0.2 cr / second; Eleven Multilingual v2 costs 120,000 cr / M characters. A Pro plan includes 9,000 credits a month.
Catalogue last updated 14 September 2026 · rates read from the live catalogue on every revalidation
Required plan for these audio models, cheapest first:Hobby · 17Pro · 12
| Model | Provider | Cost | Plan | Added |
|---|---|---|---|---|
| ACE-StepNew | fal.ai | 0.2 cr / second | Hobby | 2026-09-14 |
| CassetteAI Sound EffectsNew | fal.ai | 12 cr / clip | Hobby | 2026-09-14 |
| Eleven Flash v2.5New | ElevenLabs | 60,000 cr / M characters | Hobby | 2026-09-14 |
| Eleven Multilingual v2New | ElevenLabs | 120,000 cr / M characters | Pro | 2026-09-14 |
| Eleven MusicNew | fal.ai | 720 cr / minute | Pro | 2026-09-14 |
| Eleven Music v2.5New | ElevenLabs | 3 cr / second | Pro | 2026-09-14 |
| ElevenLabs Sound Effects v2New | ElevenLabs | 2.4 cr / second | Hobby | 2026-09-14 |
| Eleven v3New | ElevenLabs | 120,000 cr / M characters | Pro | 2026-09-14 |
| ElevenLabs Sound Effects v2 (fal.ai)New | fal.ai | 2.4 cr / second | Hobby | 2026-09-14 |
| MiniMax Music 2.6New | fal.ai | 180 cr / clip | Pro | 2026-09-14 |
| Scribe v2New | ElevenLabs | 4.4 cr / minute | Hobby | 2026-09-14 |
| Stable Audio 2.5New | fal.ai | 240 cr / clip | Pro | 2026-09-14 |
| Kling Text to Audio | Kling AI | 42 cr / clip | Hobby | 2026-09-13 |
| Lyria 3 Clip | 48 cr / clip | Hobby | 2026-09-13 | |
| Lyria 3.5 | 96 cr / clip | Pro | 2026-09-13 | |
| Seed Audio 1.0 | ByteDance | 3 cr / second | Hobby | 2026-09-13 |
| Gemini 3.5 Transcribe | 6 cr / minute | Hobby | 2026-09-08 | |
| MiniMax Speech 2.8 HD | MiniMax | 120,000 cr / M characters | Pro | 2026-08-30 |
| MiniMax Speech 2.8 Turbo | MiniMax | 72,000 cr / M characters | Pro | 2026-08-30 |
| Gemini 3.1 Flash TTS | 55,200 cr / M characters | Pro | 2026-08-08 | |
| GPT Transcribe | OpenAI | 5.4 cr / minute | Hobby | 2026-08-08 |
| Gemini 2.5 Flash TTS | 27,600 cr / M characters | Hobby | 2026-07-05 | |
| Gemini 2.5 Pro TTS | 55,200 cr / M characters | Pro | 2026-07-05 | |
| GPT-4o mini TTS | OpenAI | 22,800 cr / M characters | Hobby | 2026-07-05 |
| Grok Transcribe | xAI | 2 cr / minute | Hobby | 2026-07-05 |
| Grok Voice | xAI | 18,000 cr / M characters | Hobby | 2026-07-05 |
| Seed TTS 1.0 | ByteDance | 18,000 cr / M characters | Hobby | 2026-07-05 |
| TTS-1 | OpenAI | 18,000 cr / M characters | Hobby | 2026-07-05 |
| TTS-1 HD | OpenAI | 36,000 cr / M characters | Pro | 2026-07-05 |
Scroll the table sideways for cost, plan and date
How much does AI voice generation cost per month?
The voice tools sell character allowances that expire monthly, and none of them writes the script you are about to narrate.
| Service | Monthly | What it covers | Effective |
|---|---|---|---|
| ElevenLabs Creator | $22 | 100,000 characters | Voice only — no chat, image or video |
| Omnibo Pro | $29.99 | ~0 runs on MiniMax Speech 2.8 HD | Plus 57 chat, 25 image, 20 video models |
Scroll the table sideways for the full comparison
The Omnibo row is arithmetic you can check: 9,000 credits a month ÷ 120,000 credits per run on MiniMax Speech 2.8 HD = 0 runs on MiniMax Speech 2.8 HD. Credits do not roll over on any of these, Omnibo included. The difference is that the others each want their own subscription, and none of them will write your script or voice it. September 2026 list prices for the cheapest tier of each service that permits commercial, watermark-free output.
Work made in Omnibo
Every image on this site was generated in Omnibo, tagged with the model that made it and what that run cost. No stock, no other tools.
Questions about AI voice and transcription on Omnibo
How is Omnibo speech generation metered?
Per million characters of the text you submit, which is how the voice providers meter it themselves. A run is capped at 9,999 characters, and the composer shows the credit cost of the exact text in the box before you generate.
How is Omnibo transcription metered?
Per minute of audio, not per character — transcription models take sound in rather than text, so the table below quotes them in credits per minute.
Can I use Omnibo voice output commercially?
Yes on every paid Omnibo plan, subject to each provider’s own terms, which are linked per model on the catalogue page. Free-plan output is for evaluation.
Does the audio balance come out of the same credits as everything else?
Yes. Omnibo has one balance for chat, image, video and audio, so an unused voice allowance is not stranded in a separate subscription at the end of the month.
Write it, film it and voice it on one balance.
Start free. No card. All 29 audio models are on the same account as the 57 chat, 25 image, 20 video models.
Start free






