Home Projects UniTok
UniTok
Python

UniTok

[NeurIPS 2025 Spotlight] A Unified Tokenizer for Visual Generation and Understanding

by FoundationVision · GitHub
Stars
Forks
License
Created
Last commit
Language
autoregressive-modelsgenerativegenerative-aiMITPython
View on GitHub
In plain words

Utilize a unified tool for processing images and text, suitable for generating and understanding visual content.

You maintain this project?

Claim its page: indexed whatever its rank, translated into six languages, and enriched with what you write yourself.

Claim this page →
UniTok — GitHub preview card
📈 Star history
530529
2026-07-202026-08-31
📈 Track UniTok

Get an email alert on its next release or when it starts trending — never miss the moment.

Free · no card · unsubscribe anytime
Get email alerts →
📄 About

[NeurIPS 2025 Spotlight] A Unified Tokenizer for Visual Generation and Understanding

UniTok has 530 stars on GitHub. It has been forked 14 times. UniTok is written mainly in Python. It has been in active development since 2025. UniTok is available under the MIT license. Its main topics are autoregressive-models, generative, generative-ai, generative-model.

Frequently asked questions

What is UniTok?

[NeurIPS 2025 Spotlight] A Unified Tokenizer for Visual Generation and Understanding

Is UniTok open source?

UniTok is an open-source project. It is released under the MIT license.

Is UniTok free?

Yes. UniTok is free and open source — you can use, modify and self-host it.

What license does UniTok use?

UniTok is available under the MIT license.

What language is UniTok written in?

UniTok is written mainly in Python.

🏅 Maintainer of this project?
olud.ai badge — UniTok

Add this live badge to your README — your GitHub stars and directory rank, refreshed daily.

[![olud.ai](https://olud.ai/badge.php?tool=foundationvision-unitok)](https://olud.ai/project/foundationvision-unitok.html)
More badge options →
🧬 Shares DNA with🧬 View the DNA map →
Liquid
(Accepted by IJCV) Liquid: Language Models are Scalable and Unified Multi-modal Generators
640 · autoregressive-models
sharesautoregressive-modelsgenerative
Infinity
[CVPR 2025 Oral]Infinity ∞ : Scaling Bitwise AutoRegressive Modeling for High-Resolution Image…
1.6k · auto-regressive-model
sharesautoregressive-modelsgenerative-model
NOVA
[ICLR 2025] Autoregressive Video Generation without Vector Quantization
655 · autoregressive-models
sharesautoregressive-modelsimage-generation
Ovis-Image
Ovis-Image is a 7B text-to-image model specifically optimized for high-quality text rendering,…
318 · image-generation
sharestext-to-imageimage-generation
LakonLab
Official implementation of AsymFlow, pi-Flow, GMFlow
464 · diffusion-models
sharesgenerative-modeltext-to-image
VAR
[NeurIPS 2024 Best Paper Award][GPT beats diffusion🔥] [scaling laws in visual generation📈] Offi…
8.7k · auto-regressive-model
sharesautoregressive-modelsgenerative-model
story-iter
[ICLR 2026] A Training-free Iterative Framework for Long Story Visualization
961 · diffusion-models
sharesgenerative-modelimage-generation
Text-to-Image-Synthesis
Pytorch implementation of Generative Adversarial Text-to-Image Synthesis paper
411 · gans
sharestext-to-imageimage-generation
awesome-text-to-video
A Survey on Text-to-Video Generation/Synthesis.
736 · aigc
sharestext-to-imageimage-generation
MultiDiffusion
Official Pytorch Implementation for "MultiDiffusion: Fusing Diffusion Paths for Controlled Imag…
1.1k · diffusion-models
sharesgenerative-modelimage-generation

Measured from GitHub topics shared by both projects, weighted by how rare each topic is.