text-to-image

65 projects share this GitHub topic

text-to-image — Open-Generative-AI ★27.5ktext-to-imageawesome-gpt-image-2-API-and-Prompts — ★17kDALLE2-pytorch — ★11.3kimagen-pytorch — ★8.4kDreambooth-Stable-Diffusion — ★7.7kGenerative-Media-Skills — ★4.2kawesome-generative-ai — ★3.5kmin-dalle — ★3.5kawesome-generative-ai-apps — ★3.1kStable-Diffusion — ★2.8kVQGAN-CLIP — ★2.7kpaperbanana — ★2.3kRPG-DiffusionMaster — ★1.8kCogView — ★1.8kTokenFlow — ★1.7kcog-face-to-many — ★1.4kWebAI2API — ★1.3kCogView4 — ★1.1ktext2room — ★1.1kCogView2 — ★955muse-maskgit-pytorch — ★918unfake.js — ★866TF-ICON — ★814aphantasia — ★790Attend-and-Excite — ★772awesome-text-to-image-studies — ★760stable-diffusion-aesthetic-gradients — ★741awesome-text-to-video — ★736cog-face-to-sticker — ★641Liquid — ★640text-to-image — ★599AI-Visual-Prompt-Cookbook — ★580Awesome-LLMs-meet-Multimodal-Generation — ★552parti-pytorch — ★538UniTok — ★530awesome-aiart-pics-prompts — ★473StyleShot — ★471LakonLab — ★464Ovis-U1 — ★450DALLE-mtf — ★431Text-to-Image-Synthesis — ★411awesome-gpt-image-2-API-…★ 17kDALLE2-pytorch★ 11.3kimagen-pytorch★ 8.4kDreambooth-Stable-Diffus…★ 7.7kGenerative-Media-Skills★ 4.2kawesome-generative-ai★ 3.5kmin-dalle★ 3.5kawesome-generative-ai-ap…★ 3.1kStable-Diffusion★ 2.8kVQGAN-CLIP★ 2.7kpaperbanana★ 2.3kRPG-DiffusionMaster★ 1.8kCogView★ 1.8kTokenFlow★ 1.7kcog-face-to-many★ 1.4kWebAI2API★ 1.3kCogView4★ 1.1ktext2room★ 1.1kCogView2★ 955muse-maskgit-pytorch★ 918unfake.js★ 866TF-ICON★ 814aphantasia★ 790Attend-and-Excite★ 772awesome-text-to-image-st…★ 760stable-diffusion-aesthet…★ 741awesome-text-to-video★ 736cog-face-to-sticker★ 641Liquid★ 640text-to-image★ 599AI-Visual-Prompt-Cookboo…★ 580Awesome-LLMs-meet-Multim…★ 552parti-pytorch★ 538UniTok★ 530awesome-aiart-pics-promp…★ 473StyleShot★ 471LakonLab★ 464Ovis-U1★ 450DALLE-mtf★ 431Text-to-Image-Synthesis★ 411

Lines connect members that are measurably related to each other. Dot size reflects stars.

🧬 Members
Open-Generative-AI
Unrestricted Open-source alternative to AI video platforms — Free AI image & video generation studio with…
★ 27.5k
awesome-gpt-image-2-API-and-Prompts
GPT-Image-2 API and Prompts
★ 17k
DALLE2-pytorch
Implementation of DALL-E 2, OpenAI's updated text-to-image synthesis neural network, in Pytorch
★ 11.3k
imagen-pytorch
Implementation of Imagen, Google's Text-to-Image Neural Network, in Pytorch
★ 8.4k
Dreambooth-Stable-Diffusion
Implementation of Dreambooth (https://arxiv.org/abs/2208.12242) with Stable Diffusion
★ 7.7k
Generative-Media-Skills
Multi-modal Generative Media Skills for AI Agents (Claude Code, Cursor, Gemini CLI). High-quality image,…
★ 4.2k
awesome-generative-ai
A curated list of Generative AI tools, works, models, and references
★ 3.5k
min-dalle
min(DALL·E) is a fast, minimal port of DALL·E Mini to PyTorch
★ 3.5k
awesome-generative-ai-apps
50+ open-source generative AI apps you can clone, deploy, and monetize — image generators, video tools,…
★ 3.1k
Stable-Diffusion
FLUX, Stable Diffusion, SDXL, SD3, LoRA, Fine Tuning, DreamBooth, Training, Automatic1111, Forge WebUI,…
★ 2.8k
VQGAN-CLIP
Just playing with getting VQGAN+CLIP running locally, rather than having to use colab.
★ 2.7k
paperbanana
Open source implementation and extension of Google Research’s PaperBanana for automated academic figures,…
★ 2.3k
RPG-DiffusionMaster
[ICML 2024] Mastering Text-to-Image Diffusion: Recaptioning, Planning, and Generating with Multimodal LLMs…
★ 1.8k
CogView
Text-to-Image generation. The repo for NeurIPS 2021 paper "CogView: Mastering Text-to-Image Generation via…
★ 1.8k
TokenFlow
Official Pytorch Implementation for "TokenFlow: Consistent Diffusion Features for Consistent Video Editing"…
★ 1.7k
cog-face-to-many
Turn any face into a video game character, pixel art, claymation, 3D or toy
★ 1.4k
WebAI2API
★ 1.3k
CogView4
CogView4, CogView3-Plus and CogView3(ECCV 2024)
★ 1.1k
text2room
Text2Room generates textured 3D meshes from a given text prompt using 2D text-to-image models (ICCV2023).
★ 1.1k
CogView2
official code repo for paper "CogView2: Faster and Better Text-to-Image Generation via Hierarchical…
★ 955
muse-maskgit-pytorch
Implementation of Muse: Text-to-Image Generation via Masked Generative Transformers, in Pytorch
★ 918
unfake.js
Fix AI pixel art and vector images right in your browser
★ 866
TF-ICON
[ICCV 2023] "TF-ICON: Diffusion-Based Training-Free Cross-Domain Image Composition" (Official Implementation)
★ 814
aphantasia
CLIP + FFT/DWT/RGB = text to image/video
★ 790
Attend-and-Excite
Official Implementation for "Attend-and-Excite: Attention-Based Semantic Guidance for Text-to-Image Diffusion…
★ 772
awesome-text-to-image-studies
A collection of awesome text-to-image generation studies.
★ 760
stable-diffusion-aesthetic-gradients
Personalization for Stable Diffusion via Aesthetic Gradients 🎨
★ 741
awesome-text-to-video
A Survey on Text-to-Video Generation/Synthesis.
★ 736
cog-face-to-sticker
face-to-sticker
★ 641
Liquid
(Accepted by IJCV) Liquid: Language Models are Scalable and Unified Multi-modal Generators
★ 640
text-to-image
Generative Adversarial Text to Image Synthesis / Please Star -->
★ 599
AI-Visual-Prompt-Cookbook
118+ plug-and-play JSON style packs for Nano Banana Pro, GPT Image & Midjourney. Copy one JSON, get a style.…
★ 580
Awesome-LLMs-meet-Multimodal-Generation
🔥🔥🔥 A curated list of papers on LLMs-based multimodal generation (image, video, 3D and audio).
★ 552
parti-pytorch
Implementation of Parti, Google's pure attention-based text-to-image neural network, in Pytorch
★ 538
UniTok
[NeurIPS 2025 Spotlight] A Unified Tokenizer for Visual Generation and Understanding
★ 530
awesome-aiart-pics-prompts
🎨 精选 3000+ Gemini Nano Banana Pro 高质量提示词与生成案例 |…
★ 473
StyleShot
★ 471
LakonLab
Official implementation of AsymFlow, pi-Flow, GMFlow
★ 464
Ovis-U1
An unified model that seamlessly integrates multimodal understanding, text-to-image generation, and image…
★ 450
DALLE-mtf
Open-AI's DALL-E for large scale training in mesh-tensorflow.
★ 431
Text-to-Image-Synthesis
Pytorch implementation of Generative Adversarial Text-to-Image Synthesis paper
★ 411
stable-flow
Official implementation for "Stable Flow: Vital Layers for Training-Free Image Editing" [CVPR 2025]
★ 409
diffusers-torchao
End-to-end recipes for optimizing diffusion models with torchao and diffusers (inference and FP8 training).
★ 399
MACE
[CVPR 2024] "MACE: Mass Concept Erasure in Diffusion Models" (Official Implementation)
★ 393
CLIP-Guided-Diffusion
Just playing with getting CLIP Guided Diffusion running locally, rather than having to use colab.
★ 386
flux-krea
Official GitHub repository for FLUX.1 Krea [dev].
★ 364
T2I-CompBench
[Neurips 2023 & TPAMI] T2I-CompBench (++) for Compositional Text-to-image Generation Evaluation
★ 346
Meissonic
[ICLR 2025] Official Implementation of Meissonic: Revitalizing Masked Generative Transformers for Efficient…
★ 345
DF-GAN
[CVPR2022 oral] A Simple and Effective Baseline for Text-to-Image Synthesis
★ 326
e4t-diffusion
Implementation of Encoder-based Domain Tuning for Fast Personalization of Text-to-Image Models
★ 324
Ovis-Image
Ovis-Image is a 7B text-to-image model specifically optimized for high-quality text rendering, designed to…
★ 318
AI-Influencer-Generator
Create and customize your AI influencer open-source
★ 301
DiffusionOPSD
🔥 On-Policy Self-Distillation in Diffusion Models
★ 239
Loomic
Open-source AI canvas design agent — alternative to Lovart / CapCut Video Studio / Canva AI. Chat-driven…
★ 200
doubao2api
Reverse-engineered Doubao (豆包) API → OpenAI-compatible REST service. Free multimodal chat,…
★ 197
MooshieUI
A front-end UI for ComfyUI made for beginner level users.
★ 182
ComfyUI-MiniMax-H3-Image-Studio
ComfyUI nodes and tested workflows for MiniMax H3 still-image generation, reference editing, and Qwen detail…
★ 127
ark-cli
The fastest way to put Volcengine Ark in your terminal and your AI agent — go from prompt to generated …
★ 114
pixelgpt-24x24
Tiny local text-to-24x24 pixel art model, trained on roughly 200K samples in 30 minutes on an RTX 5090.
★ 85
Omni-Rewriter
Open agentic prompt-expansion harness for image and video generation, bridging polished demos, public APIs,…
★ 84
gpic
Official Code for GPIC: A Giant Permissive Image Corpus for Visual Generation
★ 54
mediagen
AI image and video generation skill for Claude Code and other coding agents — Gemini, OpenAI and Kie AI…
★ 54
Image-AI-Generator-2026
AI Image Generator is a powerful and user-friendly desktop application designed to help creators produce…
★ 53
awesome-gpt-image-2-prompts
Awesome GPT Image 2 Prompts
★ 44
Moda
MODA: open fashion retrieval benchmark and models by Hopit AI. MODA (203M, open source), MODA Pro Lite (213M,…
★ 44
🔗 Related families

Measured from GitHub topics shared by both projects, weighted by how rare each topic is.