Home Projects Ming-UniAudio
Ming-UniAudio
Python

Ming-UniAudio

Ming-UniAudio: Speech LLM for Joint Understanding, Generation and Editing with Unified Representation

by inclusionAI · GitHub
Stars
Forks
License
Created
Last commit
Category
Language
speechspeech-recognitionspeech-synthesisMITPython
View on GitHub
In plain words

Use a framework to understand, generate, and edit speech in one unified model.

You maintain this project?

Claim its page: indexed whatever its rank, translated into six languages, and enriched with what you write yourself.

Claim this page →
Ming-UniAudio — GitHub preview card
📈 Star history
451450
2026-07-202026-08-31
📈 Track Ming-UniAudio

Get an email alert on its next release or when it starts trending — never miss the moment.

Free · no card · unsubscribe anytime
Get email alerts →
📄 About

Ming-UniAudio: Speech LLM for Joint Understanding, Generation and Editing with Unified Representation

Ming-UniAudio has 450 stars on GitHub. It has been forked 30 times. Ming-UniAudio is written mainly in Python. It has been in active development since 2025. Ming-UniAudio is available under the MIT license. Its main topics are speech, speech-recognition, speech-synthesis, unified-model.

Frequently asked questions

What is Ming-UniAudio?

Ming-UniAudio: Speech LLM for Joint Understanding, Generation and Editing with Unified Representation

Is Ming-UniAudio open source?

Ming-UniAudio is an open-source project. It is released under the MIT license.

Is Ming-UniAudio free?

Yes. Ming-UniAudio is free and open source — you can use, modify and self-host it.

What license does Ming-UniAudio use?

Ming-UniAudio is available under the MIT license.

What language is Ming-UniAudio written in?

Ming-UniAudio is written mainly in Python.

🏅 Maintainer of this project?
olud.ai badge — Ming-UniAudio

Add this live badge to your README — your GitHub stars and directory rank, refreshed daily.

[![olud.ai](https://olud.ai/badge.php?tool=inclusionai-ming-uniaudio)](https://olud.ai/project/inclusionai-ming-uniaudio.html)
More badge options →
🧬 Shares DNA with🧬 View the DNA map →
speech-recognition-uk
🇺🇦 Speech Recognition & Synthesis for Ukrainian
439 · speech
sharesspeechspeech-synthesis
Freeze-Omni
✨✨Freeze-Omni: A Smart and Low Latency Speech-to-speech Dialogue Model with Frozen LLM
395 · large-language-models
sharesspeechspeech-synthesis
vits2
VITS2: Improving Quality and Efficiency of Single-Stage Text-to-Speech with Adversarial Learnin…
642 · deep-learning
sharesspeechspeech-synthesis
local-talking-llm
A talking LLM that runs on your own computer without needing the internet.
878 · chatbot
sharesspeech-synthesisspeech-recognition
Speech-Backbones
This is the main repository of open-sourced speech technology by Huawei Noah's Ark Lab.
604 · speech-processing
sharesspeech-synthesisspeech-recognition
Irene-Voice-Assistant
Ирина - русский голосовой ассистент для работы оффлайн. Поддерживает скиллы через плагины.
1.1k · python
sharesspeech-synthesisspeech-recognition
whisperX
WhisperX: Automatic Speech Recognition with Word-level Timestamps (& Diarization)
23.1k · asr
sharesspeechspeech-recognition
MARS5-TTS
MARS5 speech model (TTS) from CAMB.AI
2.8k · prosody
sharesspeechspeech-synthesis
MsEdgeTTS
A simple Azure Speech Service module that uses the Microsoft Edge Read Aloud API. https://www.n…
335 · free-tts
sharesspeechspeech-synthesis
stt
Voice Recognition to Text Tool / 一个离线运行的本地音视频转字幕工具,输出json、srt字幕、纯文字格式
4.7k · speech
sharesspeechspeech-recognition

Measured from GitHub topics shared by both projects, weighted by how rare each topic is.