
What's New on DubVoice.ai: 3 New Voice Providers, Meta AI Video, YouTube Videos & UGC
A lot has shipped recently. Here is the short version, grouped by what it changes for you.
Text-to-speech: three new providers, 17,800+ voices
The voice library went from 10,500 to 17,800+ voices across six providers. The three new arrivals:
Vbee — 1,649 voices, 52 languages
Vietnamese-heavy (1,212 of the voices) but far broader than that: Arabic, Hindi, Japanese, French, Indonesian, Filipino and more. If you publish for South-East Asian audiences, this is the single biggest addition.
Fish Audio — 275 community voices
A multilingual community library spanning 15 languages, with 96 English voices plus Chinese, Spanish, Portuguese, Arabic and Turkish. Strong on character and personality voices rather than corporate neutrality.
Kokoro — open source, 70% off
Hexgrad's open-source Kokoro-82M, self-hosted. 1 character = 0.3 credits — the cheapest option on the platform after Edge TTS, with no rate limit and no text-length cap. 54 voices across 9 languages, and it supports weighted voice mixing (af_bella+af_sky(2)) which nothing else here does.
All three sit alongside ElevenLabs (15,190 voices), Minimax (701) and Edge TTS (90% off). Every one of them is reachable from the same dashboard and the same [public API](/dashboard/api-docs) — Vbee and Fish Audio synthesise through the standard TTS endpoint using a prefixed voice id.
The voice library itself was rebuilt too: proper cards with avatars, full metadata tags, per-provider filters (language, gender, age, region, category) and working previews for every provider.
Video: Meta AI at 2,000 credits
Video went from 3 models to 8. Two are new:
- Meta AI — 2,000 credits per clip. By a wide margin the cheapest way to generate video here. Supports square 1:1 alongside portrait and landscape, and does image-to-video with a start frame plus an optional end frame, interpolating between the two.
- Omni Flash — from 4,688 credits. The fastest model (30-90s), priced by duration, and it accepts up to 7 reference images — more visual context than anything else on the platform.
Prices moved on the existing family too: Veo 3.1 Fast halved to 7,500 credits. It is now the best value in the Veo range and the one to reach for by default — and the Veo family remains the only one that generates native audio.
Images: Nano Banana 2 Lite and Meta AI
The image picker is now 8 models. New:
- Nano Banana 2 Lite — 500 credits. The fastest and cheapest image model, for iterating on a prompt before committing.
- Meta AI — 1,500 credits. Uses named reference components rather than a flat list:
character,sceneandstyleimages that bind to specific roles in the composition.
Nano Banana Pro remains the quality pick with 4K upscale, and Nano Banana 2 sits in the middle at 1,000.
YouTube Videos
The YouTube Videos tool takes a topic and produces a finished narrated video — script, voiceover and visuals — without hopping between tools. Useful for faceless channels and for turning an existing article into a video without rewriting it yourself.
UGC Video — coming soon
UGC Video is the next thing landing: user-generated-content style clips, the format that carries most short-form ad creative right now. The section is live in the dashboard with a preview of what is coming.
Also worth knowing
- Voice cloning now handles the full 10MB / 5-minute sample the interface always advertised. Larger uploads used to fail; they go straight to storage now instead of through the API.
- Kokoro discount deepened to 70% (was 50%).
- The public API covers Vbee and Fish Audio, with the full model line-up documented at [/dashboard/api-docs](/dashboard/api-docs) and in machine-readable form at [/llms.txt](/llms.txt).
Everything above runs off the same credit balance, starting at $4.99. No subscription, and credits never expire.
Try DubVoice.ai Today
10500+ AI voices, 6 video providers, 10 image models, AI music, translation & more — all in one platform. No subscription required.