Home Projects vosk
vosk
C

vosk

VOSK Speech Recognition Toolkit

by alphacep · GitHub
Stars
Forks
License
Created
Last commit
Language
lifelong-learningmultilingualpythonApache-2.0C
View on GitHub
In plain words

Convert spoken words into written text in multiple languages using a speech recognition system.

You maintain this project?

Claim its page: indexed whatever its rank, translated into six languages, and enriched with what you write yourself.

Claim this page →
vosk — GitHub preview card
📈 Star history
502501
2026-07-202026-08-31
📈 Track vosk

Get an email alert on its next release or when it starts trending — never miss the moment.

Free · no card · unsubscribe anytime
Get email alerts →
📄 About

VOSK Speech Recognition Toolkit

vosk has 501 stars on GitHub. It has been forked 56 times. vosk is written mainly in C. It has been in active development since 2019. vosk is available under the Apache-2.0 license. Its main topics are lifelong-learning, multilingual, python, semi-supervised-learning.

Frequently asked questions

What is vosk?

VOSK Speech Recognition Toolkit

Is vosk open source?

vosk is an open-source project. It is released under the Apache-2.0 license.

Is vosk free?

Yes. vosk is free and open source — you can use, modify and self-host it.

What license does vosk use?

vosk is available under the Apache-2.0 license.

What language is vosk written in?

vosk is written mainly in C.

🏅 Maintainer of this project?
olud.ai badge — vosk

Add this live badge to your README — your GitHub stars and directory rank, refreshed daily.

[![olud.ai](https://olud.ai/badge.php?tool=alphacep-vosk)](https://olud.ai/project/alphacep-vosk.html)
More badge options →
🧬 Shares DNA with🧬 View the DNA map →
data-preparation
Code used for sourcing and cleaning the BigScience ROOTS corpus
318 · dataset
sharesmultilingual
CCT
:page_facing_up: Semi-Supervised Semantic Segmentation with Cross-Consistency Training (CVPR 20…
412 · computer-vision
sharessemi-supervised-learning
LIBERO
Benchmarking Knowledge Transfer in Lifelong Robot Learning
2.1k · benchmark
shareslifelong-learning
mean-teacher
A state-of-the-art semi-supervised method for image recognition
1.7k · deep-learning
sharessemi-supervised-learning
self-supervised-speech-recognition
speech to text with self-supervised learning based on wav2vec 2.0 framework
380 · self-supervised-learning
sharessemi-supervised-learning
LaserMix
[CVPR 2023 Highlight] LaserMix for Semi-Supervised LiDAR Semantic Segmentation
320 · autonomous-driving
sharessemi-supervised-learning
xiaoniu
小牛视频翻译 是一款支持本地视频翻译、字幕翻译和 YouTube 视频翻译下载的 AI 工具,集成自动语音识别与多语言翻译功能,助力创作者高效完成视频翻译,应用于视频本地化与视频出海场景。
383 · ai-subtitles
sharesmultilingual
VoxCPM
VoxCPM2: Tokenizer-Free TTS for Multilingual Speech Generation, Creative Voice Design, and True…
33.5k · audio
sharesmultilingual
Foundation-Models-Framework-Lab
A practical lab for building, testing, and evaluating apps with Apple's Foundation Models frame…
1.2k · ai
Qwen3-TTS-EasyFinetuning
Easy fine-tuning for Qwen3-TTS: Fast voice cloning and high-quality multilingual speech synthes…
121 · ai

Measured from GitHub topics shared by both projects, weighted by how rare each topic is.