foundation-models

58 projets partagent ce topic GitHub

foundation-models — ColossalAI ★41.4kfoundation-modelsLLaVA — ★24.9kunilm — ★22.2kJanus — ★17.8kpyod — ★10kTabPFN — ★7.6kdata-juicer — ★7kopensquilla — ★6.8kYuE — ★6.3kchronos-forecasting — ★5.8knixtla — ★4kLimiX — ★3.8kNExT-GPT — ★3.6kAsk-Anything — ★3.3kautodistill — ★2.7kInternVideo — ★2.3kawesome-llm-powered-agent — ★2.3kMambaVision — ★2.2kalpaca_eval — ★2kgraph-fraud-detection-papers — ★1.9klag-llama — ★1.6kawesome-japanese-llm — ★1.4kawesome-vlm-architectures — ★1.3kKnowledgeEditingPapers — ★1.2kAwesome-TimeSeries-SpatioTemporal-LM-LLM — ★1.2kAwesome-Foundation-Models — ★1.2kFoundation-Models-Framework-Lab — ★1.2kONE-PEACE — ★1.1kPointLLM — ★1.1kVoxPoser — ★833ModelsGenesis — ★787Awesome-Model-Merging-Methods-Theories-Applications — ★778GigaAM — ★688Awesome-Reasoning-Foundation-Models — ★656EmerNeRF — ★639Aether — ★603tokenize-anything — ★601MMMU — ★593Groma — ★586HoloMotion — ★583Swarm — ★574LLaVA★ 24.9kunilm★ 22.2kJanus★ 17.8kpyod★ 10kTabPFN★ 7.6kdata-juicer★ 7kopensquilla★ 6.8kYuE★ 6.3kchronos-forecasting★ 5.8knixtla★ 4kLimiX★ 3.8kNExT-GPT★ 3.6kAsk-Anything★ 3.3kautodistill★ 2.7kInternVideo★ 2.3kawesome-llm-powered-agen…★ 2.3kMambaVision★ 2.2kalpaca_eval★ 2kgraph-fraud-detection-pa…★ 1.9klag-llama★ 1.6kawesome-japanese-llm★ 1.4kawesome-vlm-architecture…★ 1.3kKnowledgeEditingPapers★ 1.2kAwesome-TimeSeries-Spati…★ 1.2kAwesome-Foundation-Model…★ 1.2kFoundation-Models-Framew…★ 1.2kONE-PEACE★ 1.1kPointLLM★ 1.1kVoxPoser★ 833ModelsGenesis★ 787Awesome-Model-Merging-Me…★ 778GigaAM★ 688Awesome-Reasoning-Founda…★ 656EmerNeRF★ 639Aether★ 603tokenize-anything★ 601MMMU★ 593Groma★ 586HoloMotion★ 583Swarm★ 574

Les traits relient les membres réellement apparentés entre eux. La taille des points suit les étoiles.

🧬 Membres
ColossalAI
Making large AI models cheaper, faster and more accessible
★ 41.4k
LLaVA
[NeurIPS'23 Oral] Visual Instruction Tuning (LLaVA) built towards GPT-4V level capabilities and beyond.
★ 24.9k
unilm
Large-scale Self-supervised Pre-training Across Tasks, Languages, and Modalities
★ 22.2k
Janus
Janus-Series: Unified Multimodal Understanding and Generation Models
★ 17.8k
pyod
A Python library for anomaly detection across tabular, time series, graph, text, image, and audio data. 60+…
★ 10k
TabPFN
⚡ TabPFN: Foundation Model for Tabular Data ⚡
★ 7.6k
data-juicer
Data processing for and with foundation models! 🍎 🍋 🌽 ➡️ ➡️🍸 🍹 🍷
★ 7k
opensquilla
OpenSquilla — Token-Efficient AI Agent with same budget, higher intelligence density
★ 6.8k
YuE
YuE: Open Full-song Music Generation Foundation Model, something similar to Suno.ai but open
★ 6.3k
chronos-forecasting
Chronos: Pretrained Models for Time Series Forecasting
★ 5.8k
nixtla
TimeGPT-1: production ready pre-trained Time Series Foundation Model for forecasting and anomaly detection.…
★ 4k
LimiX
LimiX: Unleashing Structured-Data Modeling Capability for Generalist Intelligence…
★ 3.8k
NExT-GPT
Code and models for ICML 2024 paper, NExT-GPT: Any-to-Any Multimodal Large Language Model
★ 3.6k
Ask-Anything
[CVPR2024 Highlight][VideoChatGPT] ChatGPT with video understanding! And many more supported LMs such as…
★ 3.3k
autodistill
Images to inference with no labeling (use foundation models to train supervised models).
★ 2.7k
InternVideo
[ECCV2024] Video Foundation Models & Data for Multimodal Understanding
★ 2.3k
awesome-llm-powered-agent
Awesome things about LLM-powered agents. Papers / Repos / Blogs / ...
★ 2.3k
MambaVision
[CVPR 2025] Official PyTorch Implementation of MambaVision: A Hybrid Mamba-Transformer Vision Backbone
★ 2.2k
alpaca_eval
An automatic evaluator for instruction-following language models. Human-validated, high-quality, cheap, and…
★ 2k
graph-fraud-detection-papers
A curated list of Graph/Transformer-based fraud, anomaly, and outlier detection papers & resources
★ 1.9k
lag-llama
Lag-Llama: Towards Foundation Models for Probabilistic Time Series Forecasting
★ 1.6k
awesome-japanese-llm
日本語LLMまとめ - Overview of Japanese LLMs
★ 1.4k
awesome-vlm-architectures
Curated visual catalog of 155+ vision-language model (VLM/MLLM) architectures: papers, diagrams, training…
★ 1.3k
KnowledgeEditingPapers
Must-read Papers on Knowledge Editing for Large Language Models.
★ 1.2k
Awesome-TimeSeries-SpatioTemporal-LM-LLM
A professional list on Large (Language) Models and Foundation Models (LLM, LM, FM) for Time Series,…
★ 1.2k
Awesome-Foundation-Models
A curated list of foundation models for vision and language tasks
★ 1.2k
Foundation-Models-Framework-Lab
A practical lab for building, testing, and evaluating apps with Apple's Foundation Models framework.
★ 1.2k
ONE-PEACE
A general representation model across vision, audio, language modalities. Paper: ONE-PEACE: Exploring One…
★ 1.1k
PointLLM
[ECCV 2024 Best Paper Candidate & TPAMI 2025] PointLLM: Empowering Large Language Models to Understand Point…
★ 1.1k
VoxPoser
VoxPoser: Composable 3D Value Maps for Robotic Manipulation with Language Models
★ 833
ModelsGenesis
[MICCAI 2019 Young Scientist Award] [MEDIA 2020 Best Paper Award] Models Genesis, one of the first…
★ 787
Awesome-Model-Merging-Methods-Theories-Applications
Model Merging in LLMs, MLLMs, and Beyond: Methods, Theories, Applications and Opportunities. ACM Computing…
★ 778
GigaAM
Foundational Model for Speech Recognition Tasks
★ 688
Awesome-Reasoning-Foundation-Models
✨✨Latest Papers and Benchmarks in Reasoning with Foundation Models
★ 656
EmerNeRF
PyTorch Implementation of EmerNeRF: Emergent Spatial-Temporal Scene Decomposition via Self-Supervision
★ 639
Aether
[ICCV 2025 & ICCV 2025 RIWM Outstanding Paper] Aether: Geometric-Aware Unified World Modeling
★ 603
tokenize-anything
[ECCV 2024] Tokenize Anything via Prompting
★ 601
MMMU
This repo contains evaluation code for the paper "MMMU: A Massive Multi-discipline Multimodal Understanding…
★ 593
Groma
[ECCV2024] Grounded Multimodal Large Language Model with Localized Visual Tokenization
★ 586
HoloMotion
HoloMotion: A Foundation Model for Whole-Body Humanoid Control
★ 583
Swarm
Type-safe tools, on-device inference, multi-agent workflows, memory, and guardrails, in one Swift-native…
★ 574
LLM4TS
Large Language & Foundation Models for Time Series.
★ 566
HPT
Heterogeneous Pre-trained Transformer (HPT) as Scalable Policy Learner.
★ 541
TPA
[NeurIPS 2025 Spotlight] TPA: Tensor ProducT ATTenTion Transformer (https://arxiv.org/abs/2501.06425)
★ 460
VLABench
Official repo of VLABench, a large scale benchmark designed for fairly evaluating VLA, Embodied Agent, and…
★ 450
Olympus
[CVPR 2025 Highlight] Official code for "Olympus: A Universal Task Router for Computer Vision Tasks"
★ 428
ai-powered-search
The codebase for the book "AI-Powered Search" (Manning Publications, 2025) and associated "AI-Powered Search:…
★ 404
BioReason
BioReason: Incentivizing Multimodal Biological Reasoning within a DNA-LLM Model | NeurIPS '25
★ 402
MindVideo
Official code base for MinD-Video
★ 391
HyperSIGMA
The official repo for [TPAMI'25] "HyperSIGMA: Hyperspectral Intelligence Comprehension Foundation Model"
★ 379
lmms-finetune
A minimal codebase for finetuning large multimodal models, supporting llava-1.5/1.6, llava-interleave,…
★ 373
mega-data-factory
🏭 Mega Scale Multimodal DataPipeline for SOTA Foundation Models
★ 372
CarDreamer
World Model based Autonomous Driving Platform in CARLA :car:
★ 359
fondant
Production-ready data processing made easy and shareable
★ 358
GRID-playground
Platform for General Robot Intelligence Development
★ 345
ViP-LLaVA
[CVPR2024] ViP-LLaVA: Making Large Multimodal Models Understand Arbitrary Visual Prompts
★ 338
meta-prompting
Official implementation of Meta Prompting for AI Systems (https://arxiv.org/abs/2311.11482)
★ 320
Awesome-Multimodal-LLM-Autonomous-Driving
[WACV 2024 Survey Paper] Multimodal Large Language Models for Autonomous Driving
★ 311
🔗 Familles voisines

Mesuré à partir des topics GitHub communs aux deux projets, pondérés par leur rareté.