video-generation

88 projects share this GitHub topic

video-generation — OpenMontage ★55kvideo-generationToonflow-app — ★15kwaoowaoo — ★13.9kCogVideo — ★13kHunyuanVideo — ★12.3kViMax — ★11.2kvideo-shotcraft — ★6.9kdramaclaw — ★4.8kArcReel — ★4.3kGenerative-Media-Skills — ★4.2kMAGI-1 — ★3.7kSageAttention — ★3.7kInternGPT — ★3.2kPyramid-Flow — ★3.2kawesome-generative-ai-apps — ★3.1kyoutube-automation-agent — ★3kprintfilm — ★2.8kMimicMotion — ★2.6kvox-director — ★1.7kai-fusion-video — ★1.4klocally-uncensored — ★1.3kBernini — ★1.1k4DAnyone — ★1kmuapi-cli — ★1kmlx-serve — ★1kSEINE — ★967videocomposer — ★958Causal-Forcing — ★941Vista — ★889ControlVideo — ★863lmms-engine — ★824Text-To-Video-AI — ★817open-webui-tools — ★801MagicDance — ★776ima2-gen — ★729MagicDrive-V2 — ★716diffusion-forcing-transformer — ★709comfyui-mcp — ★707SparkVSR — ★700flatkey-cli — ★698Awesome-Video-World-Models-with-AR-Diffusion — ★698Toonflow-app★ 15kwaoowaoo★ 13.9kCogVideo★ 13kHunyuanVideo★ 12.3kViMax★ 11.2kvideo-shotcraft★ 6.9kdramaclaw★ 4.8kArcReel★ 4.3kGenerative-Media-Skills★ 4.2kMAGI-1★ 3.7kSageAttention★ 3.7kInternGPT★ 3.2kPyramid-Flow★ 3.2kawesome-generative-ai-ap…★ 3.1kyoutube-automation-agent★ 3kprintfilm★ 2.8kMimicMotion★ 2.6kvox-director★ 1.7kai-fusion-video★ 1.4klocally-uncensored★ 1.3kBernini★ 1.1k4DAnyone★ 1kmuapi-cli★ 1kmlx-serve★ 1kSEINE★ 967videocomposer★ 958Causal-Forcing★ 941Vista★ 889ControlVideo★ 863lmms-engine★ 824Text-To-Video-AI★ 817open-webui-tools★ 801MagicDance★ 776ima2-gen★ 729MagicDrive-V2★ 716diffusion-forcing-transf…★ 709comfyui-mcp★ 707SparkVSR★ 700flatkey-cli★ 698Awesome-Video-World-Mode…★ 698

Lines connect members that are measurably related to each other. Dot size reflects stars.

🧬 Members
OpenMontage
World's first open-source, agentic video production system. 12 production pipelines, 100+ tools, 700+ agent…
★ 55k
Toonflow-app
Toonflow 是开源一站式 AI 短剧创作工具,将小说、剧本快速转化为动画短剧。集成 AI…
★ 15k
waoowaoo
首家工业级全流程 AI 影视生产平台。Industry-first professional AI Agent platform for…
★ 13.9k
CogVideo
text and image to video generation: CogVideoX (2024) and CogVideo (ICLR 2023)
★ 13k
HunyuanVideo
HunyuanVideo: A Systematic Framework For Large Video Generation Model
★ 12.3k
ViMax
"ViMax: Agentic Video Generation (Director, Screenwriter, Producer, and Video Generator All-in-One)"
★ 11.2k
video-shotcraft
AI video skill for Claude Code & Codex — cinematic product videos with Remotion: 152 shot recipe cards, 209…
★ 6.9k
dramaclaw
A general-purpose AIGC video engine: script to finished film in one pipeline — dramas, ads, product videos,…
★ 4.8k
ArcReel
★ 4.3k
Generative-Media-Skills
Multi-modal Generative Media Skills for AI Agents (Claude Code, Cursor, Gemini CLI). High-quality image,…
★ 4.2k
MAGI-1
MAGI-1: Autoregressive Video Generation at Scale
★ 3.7k
SageAttention
[ICLR2025, ICML2025, NeurIPS2025 Spotlight] Quantized Attention achieves speedup of 2-5x compared to…
★ 3.7k
InternGPT
InternGPT (iGPT) is an open source demo platform where you can easily showcase your AI models. Now it…
★ 3.2k
Pyramid-Flow
[ICLR 2025] Pyramidal Flow Matching for Efficient Video Generative Modeling
★ 3.2k
awesome-generative-ai-apps
50+ open-source generative AI apps you can clone, deploy, and monetize — image generators, video tools,…
★ 3.1k
youtube-automation-agent
🎬 Fully automated YouTube channel management with AI agents. Creates, optimizes & publishes videos 24/7.…
★ 3k
printfilm
短剧平台 AI Short Film Motion Comic Generation Platform Industrial AI Motion Comic & Video Workbench
★ 2.8k
MimicMotion
High-Quality Human Motion Video Generation with Confidence-aware Pose Guidance
★ 2.6k
vox-director
Turn one topic into a finished Vox-style paper-collage explainer/ad video — automated end to end on Atlas…
★ 1.7k
ai-fusion-video
【融光】 - 基于 Agent 的全流程AI短剧/漫剧/视频创作平台 - Java & agentscope2.0 |…
★ 1.4k
locally-uncensored
Plug-and-play local AI studio: uncensored chat, image & video generation, coding agent. Runs abliterated LLMs…
★ 1.3k
Bernini
Bernini is a unified framework for video generation and editing that combines an MLLM-based semantic planner…
★ 1.1k
4DAnyone
[SIGGRAPH Asia 2026] 4DAnyone: Create Anyone in 4D from a Casual Monocular Video
★ 1k
muapi-cli
Official CLI for muapi.ai — generate images, videos & audio from the terminal. MCP server, 14 AI models,…
★ 1k
mlx-serve
Native LLM inference server for Apple Silicon. OpenAI + Anthropic API compatible. No Python. Includes MLX…
★ 1k
SEINE
[ICLR 2024] SEINE: Short-to-Long Video Diffusion Model for Generative Transition and Prediction
★ 967
videocomposer
Official repo for VideoComposer: Compositional Video Synthesis with Motion Controllability
★ 958
Causal-Forcing
[ICML 2026] Official codebase for "Causal Forcing: Autoregressive Diffusion Distillation Done Right for…
★ 941
Vista
[NeurIPS 2024] A Generalizable World Model for Autonomous Driving
★ 889
ControlVideo
[ICLR 2024] Official pytorch implementation of "ControlVideo: Training-free Controllable Text-to-Video…
★ 863
lmms-engine
A simple, unified multimodal models training engine. Lean, flexible, and built for hacking at scale.
★ 824
Text-To-Video-AI
Generate video from text using AI
★ 817
open-webui-tools
Open‑WebUI Tools is a modular toolkit designed to extend and enrich your Open WebUI instance, turning it…
★ 801
MagicDance
[ICML 2024] MagicPose(also known as MagicDance): Realistic Human Poses and Facial Expressions Retargeting…
★ 776
ima2-gen
Local-first visual generation runtime and studio for people and coding agents, with reproducible image and…
★ 729
MagicDrive-V2
[ICCV 2025] Official implementation of the paper “MagicDrive-V2: High-Resolution Long Video Generation for…
★ 716
diffusion-forcing-transformer
[ICML 2025] Official PyTorch Implementation of "History-Guided Video Diffusion"
★ 709
comfyui-mcp
Local-first, agent-native control plane for ComfyUI — MCP server + sidebar agent that generates images,…
★ 707
SparkVSR
[ECCV 2026] SparkVSR: Interactive Video Super-Resolution via Sparse Keyframe Propagation
★ 700
flatkey-cli
Flatkey media generation CLI for images, videos, audio, text, credits, and model discovery.
★ 698
Awesome-Video-World-Models-with-AR-Diffusion
A Curated List of Awesome Video World Models with AR Diffusion: Covering Algorithms, Applications, and…
★ 698
ChronoEdit
[ICLR 2026] ChronoEdit: Towards Temporal Reasoning for Image Editing and World Simulation
★ 697
Awesome-CVPR2026-CVPR2025-ICCV2025-CVPR2024-ECCV2026-ECCV2024-AIGC
A Collection of Papers and Codes for CVPR2026/CVPR2025/ICCV2025/CVPR2024/ECCV2026/ECCV2024 AIGC
★ 676
NOVA
[ICLR 2025] Autoregressive Video Generation without Vector Quantization
★ 655
Vibe-Workflow
Free, open-source alternative to Weavy AI, Krea Nodes, Freepik Spaces & FloraFauna AI — node-based AI…
★ 565
genblaze
Genblaze is an open source Python SDK for orchestrating generative AI media pipelines across video, audio,…
★ 558
wind-comic
Multi-agent AI pipeline that turns one line of text into a finished short-form drama: script, cinematic…
★ 548
forge-film
Multi-model DAG-driven parallel AI film generation — parallel speedup scales with scene independence;…
★ 535
Orkas-VideoStudio
Turn your coding agent into a video studio: describe a video in plain language, and your agent writes the…
★ 525
fantasy-portrait
FantasyPortrait: Enhancing Multi-Character Portrait Animation with Expression-Augmented Diffusion Transformers
★ 512
Cosmos-Drive-Dreams
Cosmos-Drive-Dreams: Scalable Synthetic Driving Data Generation with World Foundation Models
★ 509
DiffMorpher
Official Code for DiffMorpher: Unleashing the Capability of Diffusion Models for Image Morphing (CVPR 2024)
★ 504
Seedance-2.5-API
Python wrapper for ByteDance's Seedance 2.5 API — Text-to-Video, Image-to-Video, realistic human faces,…
★ 462
World-R1
[ICML 2026] World-R1: Reinforcing 3D Constraints for Text-to-Video Generation
★ 409
OpenDWM
An open source code repository of driving world models, with training, inferencing, evaluation tools, and…
★ 405
unified_video_action
Official PyTorch Implementation of Unified Video Action Model (RSS 2025)
★ 400
mega-data-factory
🏭 Mega Scale Multimodal DataPipeline for SOTA Foundation Models
★ 372
Epona
Official Code for Epona: Autoregressive Diffusion World Model for Autonomous Driving (ICCV 2025)
★ 370
HunyuanPortrait
[CVPR-2025] The official code of HunyuanPortrait: Implicit Condition Control for Enhanced Portrait Animation
★ 345
Seedance-2-API
Python wrapper for ByteDance's Seedance 2.0 , Seedance 2.5 and Seedance 2 Mini API — Text-to-Video,…
★ 344
sdk
AI video generation SDK — JSX for videos. One API for Kling, Flux, ElevenLabs, Veed. Built on Vercel AI SDK.
★ 335
PVDM
[CVPR'23] Video Probabilistic Diffusion Models in Projected Latent Space
★ 322
awesome-seedance-2.5-api-prompts
Seedance 2.5 API guide, prompts, parameters, and examples for video generation
★ 313
ComfyUI-H3-FaceRefine
Refine and improve the quality of small faces in MiniMax H3 video. Per-frame face tracking, crop, refine with…
★ 309
Echo-Memory
Official code for Echo-Memory: a controlled study of memory in action-conditioned video world models…
★ 249
ShipGenAI
🚀 50 production-ready Generative AI SaaS apps — brand them, ship them, keep 100% of the revenue. Stripe…
★ 225
Loomic
Open-source AI canvas design agent — alternative to Lovart / CapCut Video Studio / Canva AI. Chat-driven…
★ 200
guaardvark
The self-hosted AI workstation. Autonomous screen agents, 3-tier neural routing, parallel agent swarms, video…
★ 197
VEFX-Bench
VEFX-Bench: A Holistic Benchmark for Generic Video Editing and Visual Effects
★ 175
awesome-seedance-2.5-prompts-skills
100+ curated Seedance 2.5 prompts with real video previews, plus an installable Agent Skill that optimizes…
★ 172
awesome-ad-video-prompts
Curated, original high-craft prompts for AI video ads (Seedance 2.0 / Veo 3 / Kling / Runway). Companion to…
★ 156
NEWTON
NEWTON: Agentic Planning for Physically Grounded Video Generation
★ 143
LiveWorld
[ECCV 2026] Official implementation of paper LiveWorld: Simulating Out-of-Sight Dynamics in Generative Video…
★ 139
gflow-cli
Drive Google Flow from the command line: Veo video and Imagen images, scripted, batched and pipeline-ready.…
★ 139
vibe-creating-skill
Open-source, bilingual AI video-prompt skill — rewrite ideas into model-ready text-to-video prompts. A…
★ 136
MaineCoon
MaineCoon: Pursuing a Real-Time Audio-Visual Social World Model — technical report & project links. 🌐…
★ 121
DataClaw0
DataClaw: Agentic Tailoring Multimodal Data from Raw Streams — coming soon (code, weights, dataset &…
★ 117
VidForge
Open-source AI pipeline for short-form commerce videos: multimodal analysis, RAG scripts, multi-agent…
★ 108
LongLive-RAG
Official Implementation of LongLive-RAG: A general retrieval-augmented framework for long video generation.
★ 108
LongLive-RAG
Official Implementation of LongLive-RAG: A general retrieval-augmented framework for long video generation.
★ 107 · GitHub ↗
minimax-h3-opencode-skills
OpenCode skill suite for MiniMax H3 directing, routing, multishot planning, prompt generation, and review.
★ 103
MiniMax-H3-Director-Cut-Studio
Premiere-inspired PySide6 director studio for MiniMax H3 Ref2VA with AI shot planning, semantic media…
★ 89
Awesome-AVI
Awesome Audio-Visual Intelligence, Survey of Audio-Visual Intelligence
★ 86
Omni-Rewriter
Open agentic prompt-expansion harness for image and video generation, bridging polished demos, public APIs,…
★ 84
Pulse-of-Motion
The Pulse of Motion: Measuring Physical Frame Rate from Visual Dynamics
★ 74
mere-run
Run local image, text, speech, vision, music, and video workflows, plus model management and a loopback…
★ 69
mediagen
AI image and video generation skill for Claude Code and other coding agents — Gemini, OpenAI and Kie AI…
★ 54
awesome-agentic-world-model
🔥 Quo Vadis, World Modeling? Towards Interactive World Proxies for Continually Improving Agents
★ 50
🔗 Related families

Measured from GitHub topics shared by both projects, weighted by how rare each topic is.