Home Projects Ovis-U1
Ovis-U1
Python

Ovis-U1

An unified model that seamlessly integrates multimodal understanding, text-to-image generation, and image editing within a single powerful framework.

by ATH-MaaS · GitHub
Stars
Forks
License
Created
Last commit
Language
image-editingmultimodal-large-language-modelstext-to-imageApache-2.0Python
View on GitHub
In plain words

Generate images from text, edit images, and understand complex scenes using a single AI model.

You maintain this project?

Claim its page: indexed whatever its rank, translated into six languages, and enriched with what you write yourself.

Claim this page →
Ovis-U1 — GitHub preview card
📈 Star history
451450
2026-07-202026-08-31
📈 Track Ovis-U1

Get an email alert on its next release or when it starts trending — never miss the moment.

Free · no card · unsubscribe anytime
Get email alerts →
📄 About

An unified model that seamlessly integrates multimodal understanding, text-to-image generation, and image editing within a single powerful framework.

Ovis-U1 has 450 stars on GitHub. It has been forked 15 times. Ovis-U1 is written mainly in Python. It has been in active development since 2025. Ovis-U1 is available under the Apache-2.0 license. Its main topics are image-editing, multimodal-large-language-models, text-to-image.

Frequently asked questions

What is Ovis-U1?

An unified model that seamlessly integrates multimodal understanding, text-to-image generation, and image editing within a single powerful framework.

Is Ovis-U1 open source?

Ovis-U1 is an open-source project. It is released under the Apache-2.0 license.

Is Ovis-U1 free?

Yes. Ovis-U1 is free and open source — you can use, modify and self-host it.

What license does Ovis-U1 use?

Ovis-U1 is available under the Apache-2.0 license.

What language is Ovis-U1 written in?

Ovis-U1 is written mainly in Python.

🏅 Maintainer of this project?
olud.ai badge — Ovis-U1

Add this live badge to your README — your GitHub stars and directory rank, refreshed daily.

[![olud.ai](https://olud.ai/badge.php?tool=ath-maas-ovis-u1)](https://olud.ai/project/ath-maas-ovis-u1.html)
More badge options →
🧬 Shares DNA with🧬 View the DNA map →
RPG-DiffusionMaster
[ICML 2024] Mastering Text-to-Image Diffusion: Recaptioning, Planning, and Generating with Mult…
1.8k · image-editting
sharesmultimodal-large-language-modelstext-to-image
stable-flow
Official implementation for "Stable Flow: Vital Layers for Training-Free Image Editing" [CVPR 2…
409 · deep-learning
sharesimage-editingtext-to-image
Ovis-Image
Ovis-Image is a 7B text-to-image model specifically optimized for high-quality text rendering,…
318 · image-generation
sharestext-to-image
DALLE2-pytorch
Implementation of DALL-E 2, OpenAI's updated text-to-image synthesis neural network, in Pytorc…
11.3k · artificial-intelligence
sharestext-to-image
DF-GAN
[CVPR2022 oral] A Simple and Effective Baseline for Text-to-Image Synthesis
326 · generative-adversarial-network
sharestext-to-image
Paint-by-Example
Paint by Example: Exemplar-based Image Editing with Diffusion Models
1.3k · computer-vision
sharesimage-editing
Awesome-Medical-Large-Language-Models
Curated papers on Large Language Models in Healthcare and Medical domain
390 · large-language-models
sharesmultimodal-large-language-models
VQGAN-CLIP
Just playing with getting VQGAN+CLIP running locally, rather than having to use colab.
2.7k · text-to-image
sharestext-to-image
Spatial-MLLM
[NeurIPS 2025 Spotlight] Official implementation of Spatial-MLLM: Boosting MLLM Capabilities in…
477 · aigc
sharesmultimodal-large-language-models
Attend-and-Excite
Official Implementation for "Attend-and-Excite: Attention-Based Semantic Guidance for Text-to-I…
772 · diffusion-models
sharestext-to-image

Measured from GitHub topics shared by both projects, weighted by how rare each topic is.