Home Projects OmniQuant
OmniQuant
Python

OmniQuant

[ICLR2024 spotlight] OmniQuant is a simple and powerful quantization technique for LLMs.

by OpenGVLab · GitHub
Stars
Forks
Trending
License
Created
Last commit
Language
large-language-modelsllmquantizationMITPython
View on GitHub
In plain words

Run a technique to reduce the size of large language models for better performance on your own hardware.

You maintain this project?

Claim its page: indexed whatever its rank, translated into six languages, and enriched with what you write yourself.

Claim this page →
OmniQuant — GitHub preview card
📈 Star history
911901
2026-07-202026-08-31
📈 Track OmniQuant

Get an email alert on its next release or when it starts trending — never miss the moment.

Free · no card · unsubscribe anytime
Get email alerts →
📄 About

[ICLR2024 spotlight] OmniQuant is a simple and powerful quantization technique for LLMs.

OmniQuant has 911 stars on GitHub. It has been forked 87 times. OmniQuant is written mainly in Python. It has been in active development since 2023. OmniQuant is available under the MIT license. Its main topics are large-language-models, llm, quantization.

Frequently asked questions

What is OmniQuant?

[ICLR2024 spotlight] OmniQuant is a simple and powerful quantization technique for LLMs.

Is OmniQuant open source?

OmniQuant is an open-source project. It is released under the MIT license.

Is OmniQuant free?

Yes. OmniQuant is free and open source — you can use, modify and self-host it.

What license does OmniQuant use?

OmniQuant is available under the MIT license.

What language is OmniQuant written in?

OmniQuant is written mainly in Python.

🏅 Maintainer of this project?
olud.ai badge — OmniQuant

Add this live badge to your README — your GitHub stars and directory rank, refreshed daily.

[![olud.ai](https://olud.ai/badge.php?tool=opengvlab-omniquant)](https://olud.ai/project/opengvlab-omniquant.html)
More badge options →
🧬 Shares DNA with🧬 View the DNA map →

Measured from GitHub topics shared by both projects, weighted by how rare each topic is.