transformer

59 proyectos comparten este topic de GitHub

transformer — transformers ★164.7ktransformervllm — ★90.6ksglang — ★33kheretic — ★29.6kMNN — ★16knano-vllm — ★15.2kRWKV-LM — ★14.7klm-evaluation-harness — ★13.8kpetals — ★10.5kMindSearch — ★6.9kkimi-k3-in-c — ★6.9kFasterTransformer — ★6.4kcactus — ★6ksemantic-router — ★5.5kbeat-ai — ★4.7kbeat-ai — ★4.7kbeat-ai — ★4.7kbeatai — ★4.7kAIGC-Interview-Book — ★4.5ktowhee — ★3.5kvllm-ascend — ★2.7kai-engineering-resources — ★2.6kmPLUG-Owl — ★2.5kEasyLM — ★2.5kMoBA — ★2.2kAwesome-LLM-Long-Context-Modeling — ★2.2kaici — ★2.1kFireRedASR — ★2kgraph-fraud-detection-papers — ★1.9kLLMCompiler — ★1.9knndeploy — ★1.9ksafe-rlhf — ★1.6kVibeThinker — ★1.6kparallax — ★1.4kGPA — ★1.3kTinyEngram — ★1.1kCogView2 — ★955papers-for-molecular-design-using-DL — ★951GenAI_LLM_timeline — ★950YOLOS — ★903kat — ★847vllm★ 90.6ksglang★ 33kheretic★ 29.6kMNN★ 16knano-vllm★ 15.2kRWKV-LM★ 14.7klm-evaluation-harness★ 13.8kpetals★ 10.5kMindSearch★ 6.9kkimi-k3-in-c★ 6.9kFasterTransformer★ 6.4kcactus★ 6ksemantic-router★ 5.5kbeat-ai★ 4.7kbeat-ai★ 4.7kbeat-ai★ 4.7kbeatai★ 4.7kAIGC-Interview-Book★ 4.5ktowhee★ 3.5kvllm-ascend★ 2.7kai-engineering-resources★ 2.6kmPLUG-Owl★ 2.5kEasyLM★ 2.5kMoBA★ 2.2kAwesome-LLM-Long-Context…★ 2.2kaici★ 2.1kFireRedASR★ 2kgraph-fraud-detection-pa…★ 1.9kLLMCompiler★ 1.9knndeploy★ 1.9ksafe-rlhf★ 1.6kVibeThinker★ 1.6kparallax★ 1.4kGPA★ 1.3kTinyEngram★ 1.1kCogView2★ 955papers-for-molecular-des…★ 951GenAI_LLM_timeline★ 950YOLOS★ 903kat★ 847

Las líneas conectan a los miembros que están mediblemente relacionados entre sí. El tamaño de los puntos refleja las estrellas.

🧬 Miembros
transformers
🤗 Transformers: the model-definition framework for state-of-the-art machine learning models in text,…
★ 164.7k
vllm
A high-throughput and memory-efficient inference and serving engine for LLMs
★ 90.6k
sglang
SGLang is a high-performance serving framework for large language models and multimodal models.
★ 33k
heretic
Fully automatic censorship removal for language models
★ 29.6k
MNN
MNN: A blazing-fast, lightweight inference engine battle-tested by Alibaba, powering high-performance…
★ 16k
nano-vllm
Nano vLLM
★ 15.2k
RWKV-LM
RWKV (pronounced RwaKuv) is an RNN with great LLM performance, which can also be directly trained like a GPT…
★ 14.7k
lm-evaluation-harness
A framework for few-shot evaluation of language models.
★ 13.8k
petals
🌸 Run LLMs at home, BitTorrent-style. Fine-tuning and inference up to 10x faster than offloading
★ 10.5k
MindSearch
🔍 An LLM-based Multi-agent Framework of Web Search Engine (like Perplexity.ai Pro and SearchGPT)
★ 6.9k
kimi-k3-in-c
A 2.78-trillion-parameter Kimi K3 running inference on a single CPU in 8.24 GB of RAM. Portable C99: no BLAS,…
★ 6.9k
FasterTransformer
Transformer related optimization, including BERT, GPT
★ 6.4k
cactus
Quantization, kernels, runtime and inference engine for mobiles, wearables, smart home and robots.
★ 6k
semantic-router
A programmable Mixture-of-Models router for heterogeneous LLM inference
★ 5.5k
beat-ai
★ 4.7k
beat-ai
★ 4.7k
beat-ai
★ 4.7k
beatai
★ 4.7k
AIGC-Interview-Book
★ 4.5k
towhee
Towhee is a framework that is dedicated to making neural data processing pipelines simple and fast.
★ 3.5k
vllm-ascend
Community maintained hardware plugin for vLLM on Ascend
★ 2.7k
ai-engineering-resources
Research papers and blogs to transition to AI Engineering
★ 2.6k
mPLUG-Owl
mPLUG-Owl: The Powerful Multi-modal Large Language Model Family
★ 2.5k
EasyLM
Large language models (LLMs) made easy, EasyLM is a one stop solution for pre-training, finetuning,…
★ 2.5k
MoBA
MoBA: Mixture of Block Attention for Long-Context LLMs
★ 2.2k
Awesome-LLM-Long-Context-Modeling
📰 Must-read papers and blogs on LLM based Long Context Modeling 🔥
★ 2.2k
aici
AICI: Prompts as (Wasm) Programs
★ 2.1k
FireRedASR
Open-source industrial-grade ASR models supporting Mandarin, Chinese dialects and English, achieving a new…
★ 2k
graph-fraud-detection-papers
A curated list of Graph/Transformer-based fraud, anomaly, and outlier detection papers & resources
★ 1.9k
LLMCompiler
[ICML 2024] LLMCompiler: An LLM Compiler for Parallel Function Calling
★ 1.9k
nndeploy
一款简单易用和高性能的AI部署框架 | An Easy-to-Use and High-Performance AI Deployment Framework
★ 1.9k
safe-rlhf
Safe RLHF: Constrained Value Alignment via Safe Reinforcement Learning from Human Feedback
★ 1.6k
VibeThinker
Tiny Model, Big Logic: Diversity-Driven Optimization Elicits Large-Model Reasoning Ability in VibeThinker-1.5B
★ 1.6k
parallax
Parallax is a distributed model serving framework that lets you build your own AI cluster anywhere
★ 1.4k
GPA
[AutoArk] GPA (General Purpose Audio) can do ASR, TTS and voice conversion with one tiny model!
★ 1.3k
TinyEngram
Research of DeepSeek Engram Architecture based on Qwen-3 and Stable Diffusion series.
★ 1.1k
CogView2
official code repo for paper "CogView2: Faster and Better Text-to-Image Generation via Hierarchical…
★ 955
papers-for-molecular-design-using-DL
List of Molecular and Material design using Generative AI and Deep Learning
★ 951
GenAI_LLM_timeline
ChatGPT, GenerativeAI and LLMs Timeline
★ 950
YOLOS
[NeurIPS 2021] You Only Look at One Sequence
★ 903
kat
[ICLR2025] Kolmogorov-Arnold Transformer
★ 847
HiVT
[CVPR 2022] HiVT: Hierarchical Vector Transformer for Multi-Agent Motion Prediction
★ 745
SqueezeLLM
[ICML 2024] SqueezeLLM: Dense-and-Sparse Quantization
★ 723
diffusion-forcing-transformer
[ICML 2025] Official PyTorch Implementation of "History-Guided Video Diffusion"
★ 709
InterFuser
[CoRL 2022] InterFuser: Safety-Enhanced Autonomous Driving Using Interpretable Sensor Fusion Transformer
★ 653
vidur
Accurate, large-scale, and extensible simulator for LLM inference Systems
★ 642
multilspy
multilspy is a lsp client library in Python intended to be used to build applications around language servers.
★ 604
memit
Mass-editing thousands of facts into a transformer memory (ICLR 2023)
★ 555
krasis
Krasis is a Hybrid LLM runtime which focuses on efficient running of larger models on consumer grade VRAM…
★ 516
SparseBEV
[ICCV 2023 & TPAMI 2026] SparseBEV: High-Performance Sparse 3D Object Detection from Multi-Camera Videos
★ 468
JetStream
JetStream is a throughput and memory optimized engine for LLM inference on XLA devices, starting with TPUs…
★ 456
KVQuant
[NeurIPS 2024] KVQuant: Towards 10 Million Context Length LLM Inference with KV Cache Quantization
★ 433
KIVI
[ICML 2024] KIVI: A Tuning-Free Asymmetric 2bit Quantization for KV Cache
★ 431
Awesome-Attention-Heads
An awesome repository & A comprehensive survey on interpretability of LLM attention heads.
★ 413
OpenGraph
[EMNLP'2024] "OpenGraph: Towards Open Graph Foundation Models"
★ 338
chat.petals.dev
💬 Chatbot web app + HTTP and Websocket endpoints for LLM inference with the Petals client
★ 321
cevahir-ai
Full-stack open-source AI engine for building language models — tokenizer training, transformer…
★ 219
neo-unify
Toy-scale unified multimodal model experiments — encoder-free understanding & generation with…
★ 51
TeaLeaves
End-to-end pipeline for seeing how LLMs actually process your prompts. Capture attention across every layer,…
★ 42
🔗 Familias relacionadas

Medido a partir de los temas de GitHub compartidos por ambos proyectos, ponderado por cuán raros son cada uno de los temas.