Voice & audio, with a verdict on each.

Turning speech into text, text into speech, and cloning a voice from a short sample.

Consent is not optional. Cloning a voice without documented permission is a legal problem in several jurisdictions before it is a technical one.

16 tracked1 we would adopt7 with a licence to check

Moving fastest this week: MoneyPrinterTurbo (+17k), OpenMontage (+8.0k), unsloth (+5.9k).

Adopt now

1

Mature enough to put in production this quarter.

Pilot it

9

Worth a timeboxed spike before you bet on it.

MoneyPrinterTurbo117k
harry0703
Pilot it+17k this week

AI-powered short video generator

Instead of: manual video editing tools

MITUpdated today
Pilot it
VoxCPM36k
OpenBMB
Pilot it+1.9k this week

Open-source text-to-speech synthesis

Instead of: Paid TTS services or manual voiceovers

Apache-2.0Updated today
Pilot it
sherpa-onnx14k
k2-fsa
Pilot it+611 this week

Offline speech-to-text and text-to-speech engine

Instead of: Cloud-based speech services like Google Cloud Speech-to-Text

Apache-2.0Updated today
Pilot it
GPT-SoVITS61k
RVC-Boss
Pilot it+1.1k this week

Clones a voice from roughly a minute of audio and reads new text in it.

Instead of: Per-character pricing at a commercial voice vendor.

MITLast push 8 days ago
Pilot it
OmniVoice-Studio9.4k
debpalash
Pilot it+293 this week

Local voice cloning and audio editing tool

Instead of: ElevenLabs or manual voice editing

AGPL-3.0Last push 26 days ago
Pilot it
pyvideotrans19k
jianchang512
Pilot it+362 this week

Video translation with dubbing and subtitles

Instead of: Manual video translation and editing

GPL-3.0Updated today
Pilot it
mlx-audio7.8k
Blaizzy
Pilot it+171 this week

Apple Silicon speech analysis library

Instead of: Manual speech processing or paid APIs

MITUpdated 6 days ago
Pilot it
espeak-ng6.8k
espeak-ng
Pilot it+55 this week

Open-source speech synthesizer

Instead of: Paid text-to-speech services

GPL-3.0Updated 7 days ago
Pilot it
espnet9.9k
espnet
Pilot it+37 this week

Python toolkit for end-to-end speech processing

Instead of: Manual speech processing or Kaldi

Apache-2.0Updated today
Pilot it

Watch, or skip

6

Real software; just not where a small team's next hundred hours should go.

Voice & audio, in short

What are the best open source speech and voice AI right now?
We would put supertonic into production this quarter. supertonic — On-device text-to-speech engine The full list below ranks 16 by momentum, size and how recently they shipped.
Which of these can I use in a commercial product?
9 of the 16 carry a permissive licence (MIT, Apache-2.0, BSD). 7 do not — OpenMontage (AGPL-3.0), index-tts (Custom / check the LICENSE file), voice-pro (GPL-3.0) and others. Copyleft and source-available licences carry obligations when you ship them inside something you sell. Read the LICENSE file; this is a reading of the label, not legal advice.
How is this list ranked?
By real week-over-week star movement first, with a size floor so an established tool is not buried by a two-week-old project, and a hard penalty for anything that has stopped shipping. Verdicts are an editorial call for a team of 2-20 with limited engineering hours — not a judgement on the software, and not the same call we would make for a lab.

Which of these matters to your business?

Tell us what you sell in one sentence and we will hand you three specific moves — priced per month, with the arithmetic shown. Free, no signup.

Give me three moves

Stars, forks, licence and last-push data come from the public GitHub API. Verdicts are NoizeOff's editorial opinion for a team of 2–20, not advice from any project's maintainers, and not legal advice on licensing.

All categories · Adoption Radar · Home