K2-FSA logo

OmniVoice

OmniVoicev2026Current
byK2-FSAK2-FSA(open source)
Released March 31, 2026
Context--
Price (In / Out)Free / Free
Training CutoffNot publicly specified
CategoryAudio Model

About OmniVoice

OmniVoice is an open-source massively multilingual zero-shot text-to-speech and voice-cloning model from K2-FSA. The project README says it supports more than 600 languages, voice design controls, and fast inference with real-time factor as low as 0.025.

Capabilities

text to speechvoice cloningmultilingual audiovoice design

Input Modalities

textaudio

Output Modalities

audio

Technical Details

API Identifier
k2-fsa/OmniVoice
Category
Audio Model

Tags

text-to-speechvoice-cloningmultilingualspeech-aiopen-sourcediffusion-model

Benchmarks

Performance scores for OmniVoice across standard benchmarks.

Real-time factorofficial-github-readme · Sep 2026
0.025%

Pricing

Token pricing for OmniVoice API usage.

Input Tokens

Free

per million tokens

Output Tokens

Free

per million tokens

Pricing Calculator

Input cost$0.00
Output cost$0.00
Estimated monthly cost$0.00

Open-source model repository; hosted API pricing is not listed in the GitHub README.

Competing Models

Same pricing tier — direct alternatives to OmniVoice

specialty
K2-FSA logo
K2-FSAopen source

Open-source speech AI and finite-state automata projects

1 ModelsFounded 2020Not publicly listed
View full profile