Home Projects GigaAM
GigaAM
Python

GigaAM

Foundational Model for Speech Recognition Tasks

by salute-developers · GitHub
Stars
Forks
License
Created
Last commit
Language
emotion-recognitionfoundation-modelsself-supervised-learningMITPython
View on GitHub
In plain words

Use a speech recognition model trained on multiple languages to transcribe audio with high accuracy.

You maintain this project?

Claim its page: indexed whatever its rank, translated into six languages, and enriched with what you write yourself.

Claim this page →
GigaAM — GitHub preview card
📈 Star history
689688
2026-07-202026-08-31
📈 Track GigaAM

Get an email alert on its next release or when it starts trending — never miss the moment.

Free · no card · unsubscribe anytime
Get email alerts →
📄 About

Foundational Model for Speech Recognition Tasks

GigaAM has 688 stars on GitHub. It has been forked 93 times. GigaAM is written mainly in Python. It has been in active development since 2024. GigaAM is available under the MIT license. Its main topics are emotion-recognition, foundation-models, self-supervised-learning, speech-recognition.

Frequently asked questions

What is GigaAM?

Foundational Model for Speech Recognition Tasks

Is GigaAM open source?

GigaAM is an open-source project. It is released under the MIT license.

Is GigaAM free?

Yes. GigaAM is free and open source — you can use, modify and self-host it.

What license does GigaAM use?

GigaAM is available under the MIT license.

What language is GigaAM written in?

GigaAM is written mainly in Python.

🏅 Maintainer of this project?
olud.ai badge — GigaAM

Add this live badge to your README — your GitHub stars and directory rank, refreshed daily.

[![olud.ai](https://olud.ai/badge.php?tool=salute-developers-gigaam)](https://olud.ai/project/salute-developers-gigaam.html)
More badge options →
🧬 Shares DNA with🧬 View the DNA map →
self-supervised-speech-recognition
speech to text with self-supervised learning based on wav2vec 2.0 framework
380 · self-supervised-learning
sharesself-supervised-learningspeech-recognition
byol-pytorch
Usable Implementation of "Bootstrap Your Own Latent" self-supervised learning, from Deepmind, i…
1.9k · artificial-intelligence
sharesself-supervised-learning
Emotion-detection
Real-time Facial Emotion Detection using deep learning
1.4k · computer-vision
sharesemotion-recognition
kaldi-gstreamer-server
Real-time full-duplex speech recognition server, based on the Kaldi toolkit and the GStreamer f…
1.1k · speech-recognition
sharesspeech-recognition
FastASR
这是一个用C++实现ASR推理的项目,它依赖很少,安装也很简单,推理速度很快,在树莓派4B等ARM平台也可以流畅的运行。 支持的模型是由Google的Transformer模型中优化而来,数…
553 · speech-recognition
sharesspeech-recognition
World-Models-Autonomous-Driving-Survey
A curated list of world models for autonomous driving.
524 · autonomous-driving
sharesself-supervised-learning
dragonfly
Speech recognition framework allowing powerful Python-based scripting and extension of Dragon N…
413 · python
sharesspeech-recognition
SfMLearner
An unsupervised learning framework for depth and ego-motion estimation from monocular videos
2k · deep-learning
sharesself-supervised-learning
Sound-of-Pixels
Codebase for ECCV18 "The Sound of Pixels"
393 · computer-vision
sharesself-supervised-learning
GRID-playground
Platform for General Robot Intelligence Development
345 · computer-vision
sharesfoundation-models

Measured from GitHub topics shared by both projects, weighted by how rare each topic is.