video

26 Projekte teilen dieses GitHub-Topic

video — hyperframes ★43.4kvideosrs — ★29.2kagents — ★13.9kten-framework — ★11.1kautoclip — ★7.1kLLPlayer — ★4.1kSimpleMem — ★3.7kagentheroes — ★3.4kFireRed-OpenStoryline — ★3.3kAsk-Anything — ★3.3kmPLUG-Owl — ★2.5kvideo-starter-kit — ★2.4kvox-director — ★1.7kSALMONN — ★1.5ktongflow — ★1kVideo-MME — ★791python_video_stab — ★773videogan — ★706SparkVSR — ★700Ego4d — ★620vatic — ★613LLaVA-Mini — ★577Video-MME-v2 — ★367VEFX-Bench — ★175Awesome-AVI — ★86Pulse-of-Motion — ★74srs★ 29.2kagents★ 13.9kten-framework★ 11.1kautoclip★ 7.1kLLPlayer★ 4.1kSimpleMem★ 3.7kagentheroes★ 3.4kFireRed-OpenStoryline★ 3.3kAsk-Anything★ 3.3kmPLUG-Owl★ 2.5kvideo-starter-kit★ 2.4kvox-director★ 1.7kSALMONN★ 1.5ktongflow★ 1kVideo-MME★ 791python_video_stab★ 773videogan★ 706SparkVSR★ 700Ego4d★ 620vatic★ 613LLaVA-Mini★ 577Video-MME-v2★ 367VEFX-Bench★ 175Awesome-AVI★ 86Pulse-of-Motion★ 74

Linien verbinden Mitglieder, die messbar miteinander verwandt sind. Die Punktgröße spiegelt die Sterne wider.

🧬 Mitglieder
hyperframes
Write HTML. Render video. Built for agents.
★ 43.4k
srs
SRS is a simple, high-performance, AI-driven real-time media server supporting RTMP, WebRTC, HLS, HTTP-FLV,…
★ 29.2k
agents
A framework for building realtime voice AI agents 🤖🎙️📹
★ 13.9k
ten-framework
Open-source framework for conversational voice AI agents
★ 11.1k
autoclip
AutoClip : AI-powered video clipping and highlight generation · 一款智能高光提取与剪辑的二创工…
★ 7.1k
LLPlayer
The media player for language learning, with dual subtitles, AI-generated subtitles, real-time translation,…
★ 4.1k
SimpleMem
SimpleMem: Efficient Lifelong Memory for LLM Agents — Text & Multimodal
★ 3.7k
agentheroes
Generate, animate and schedule your AI characters 🤖
★ 3.4k
FireRed-OpenStoryline
FireRed-OpenStoryline is an AI video editing agent that transforms manual editing into intention-driven…
★ 3.3k
Ask-Anything
[CVPR2024 Highlight][VideoChatGPT] ChatGPT with video understanding! And many more supported LMs such as…
★ 3.3k
mPLUG-Owl
mPLUG-Owl: The Powerful Multi-modal Large Language Model Family
★ 2.5k
video-starter-kit
Enable AI models for video production in the browser
★ 2.4k
vox-director
Turn one topic into a finished Vox-style paper-collage explainer/ad video — automated end to end on Atlas…
★ 1.7k
SALMONN
SALMONN family: A suite of advanced multi-modal LLMs
★ 1.5k
tongflow
TongFlow — Multimodal GenAI Studio
★ 1k
Video-MME
✨✨[CVPR 2025] Video-MME: The First-Ever Comprehensive Evaluation Benchmark of Multi-modal LLMs in Video…
★ 791
python_video_stab
A Python package to stabilize videos using OpenCV
★ 773
videogan
Generating Videos with Scene Dynamics. NIPS 2016.
★ 706
SparkVSR
[ECCV 2026] SparkVSR: Interactive Video Super-Resolution via Sparse Keyframe Propagation
★ 700
Ego4d
Ego4d dataset repository. Download the dataset, visualize, extract features & example usage of the dataset
★ 620
vatic
Efficiently Scaling Up Video Annotation with Crowdsourced Marketplaces. IJCV 2012
★ 613
LLaVA-Mini
LLaVA-Mini is a unified large multimodal model (LMM) that can support the understanding of images,…
★ 577
Video-MME-v2
Video-MME-v2: Towards the Next Stage in Benchmarks for Comprehensive Video Understanding
★ 367
VEFX-Bench
VEFX-Bench: A Holistic Benchmark for Generic Video Editing and Visual Effects
★ 175
Awesome-AVI
Awesome Audio-Visual Intelligence, Survey of Audio-Visual Intelligence
★ 86
Pulse-of-Motion
The Pulse of Motion: Measuring Physical Frame Rate from Visual Dynamics
★ 74
🔗 Verwandte Familien

Gemessen anhand der von beiden Projekten geteilten GitHub-Themen, gewichtet nach der Seltenheit jedes Themas.