Home Projects Freeze-Omni
Freeze-Omni
Python

Freeze-Omni

✨✨Freeze-Omni: A Smart and Low Latency Speech-to-speech Dialogue Model with Frozen LLM

by VITA-MLLM · GitHub
Stars
Forks
Trending
Created
Last commit
Language
large-language-modelsmultimodal-large-language-modelsspeechPython
View on GitHub
In plain words

Engage in real-time speech conversations with a smart model that translates spoken dialogue into speech.

You maintain this project?

Claim its page: indexed whatever its rank, translated into six languages, and enriched with what you write yourself.

Claim this page →
Freeze-Omni — GitHub preview card
📈 Star history
395387
2026-07-202026-08-31
📈 Track Freeze-Omni

Get an email alert on its next release or when it starts trending — never miss the moment.

Free · no card · unsubscribe anytime
Get email alerts →
📄 About

✨✨Freeze-Omni: A Smart and Low Latency Speech-to-speech Dialogue Model with Frozen LLM

Freeze-Omni has 395 stars on GitHub. It has been forked 29 times. Freeze-Omni is written mainly in Python. It has been in active development since 2024. Its main topics are large-language-models, multimodal-large-language-models, speech, speech-recognition.

Frequently asked questions

What is Freeze-Omni?

✨✨Freeze-Omni: A Smart and Low Latency Speech-to-speech Dialogue Model with Frozen LLM

Is Freeze-Omni open source?

Freeze-Omni is an open-source project.

Is Freeze-Omni free?

Yes. Freeze-Omni is free and open source — you can use, modify and self-host it.

What language is Freeze-Omni written in?

Freeze-Omni is written mainly in Python.

🏅 Maintainer of this project?
olud.ai badge — Freeze-Omni

Add this live badge to your README — your GitHub stars and directory rank, refreshed daily.

[![olud.ai](https://olud.ai/badge.php?tool=vita-mllm-freeze-omni)](https://olud.ai/project/vita-mllm-freeze-omni.html)
More badge options →
🧬 Shares DNA with🧬 View the DNA map →
speech-recognition-uk
🇺🇦 Speech Recognition & Synthesis for Ukrainian
439 · speech
sharesspeechspeech-synthesis
Ming-UniAudio
Ming-UniAudio: Speech LLM for Joint Understanding, Generation and Editing with Unified Represen…
450 · speech
sharesspeechspeech-synthesis
vits2
VITS2: Improving Quality and Efficiency of Single-Stage Text-to-Speech with Adversarial Learnin…
642 · deep-learning
sharesspeechspeech-synthesis
local-talking-llm
A talking LLM that runs on your own computer without needing the internet.
878 · chatbot
sharesspeech-synthesisspeech-recognition
Speech-Backbones
This is the main repository of open-sourced speech technology by Huawei Noah's Ark Lab.
604 · speech-processing
sharesspeech-synthesisspeech-recognition
Irene-Voice-Assistant
Ирина - русский голосовой ассистент для работы оффлайн. Поддерживает скиллы через плагины.
1.1k · python
sharesspeech-synthesisspeech-recognition
whisperX
WhisperX: Automatic Speech Recognition with Word-level Timestamps (& Diarization)
23.1k · asr
sharesspeechspeech-recognition
MARS5-TTS
MARS5 speech model (TTS) from CAMB.AI
2.8k · prosody
sharesspeechspeech-synthesis
MsEdgeTTS
A simple Azure Speech Service module that uses the Microsoft Edge Read Aloud API. https://www.n…
335 · free-tts
sharesspeechspeech-synthesis
stt
Voice Recognition to Text Tool / 一个离线运行的本地音视频转字幕工具,输出json、srt字幕、纯文字格式
4.7k · speech
sharesspeechspeech-recognition

Measured from GitHub topics shared by both projects, weighted by how rare each topic is.