compression

22 projets partagent ce topic GitHub

compression — headroom ★63.2kcompressionDeepSpeed — ★42.8kPaddleNLP — ★13kSimpleMem — ★3.7kvortex — ★3.1kaimet — ★2.7kbolt — ★2.5kAwesome-Efficient-LLM — ★2kCompressAI — ★1.6kmodel-optimization — ★1.6kAwesome-Knowledge-Distillation-of-LLMs — ★1.3knncf — ★1.2kMindPipe — ★1kswin2sr — ★690DFloat11 — ★652awesome_computer_science — ★573wavemap — ★567KVQuant — ★430Context-Engine — ★402BK-SDM — ★320picollm — ★315distill — ★174DeepSpeed★ 42.8kPaddleNLP★ 13kSimpleMem★ 3.7kvortex★ 3.1kaimet★ 2.7kbolt★ 2.5kAwesome-Efficient-LLM★ 2kCompressAI★ 1.6kmodel-optimization★ 1.6kAwesome-Knowledge-Distil…★ 1.3knncf★ 1.2kMindPipe★ 1kswin2sr★ 690DFloat11★ 652awesome_computer_science★ 573wavemap★ 567KVQuant★ 430Context-Engine★ 402BK-SDM★ 320picollm★ 315distill★ 174

Les traits relient les membres réellement apparentés entre eux. La taille des points suit les étoiles.

🧬 Membres
headroom
Compress tool outputs, logs, files, and RAG chunks before they reach the LLM. 20% fewer tokens for coding…
★ 63.2k
DeepSpeed
DeepSpeed is a deep learning optimization library that makes distributed training and inference easy,…
★ 42.8k
PaddleNLP
Easy-to-use and powerful LLM and SLM library with awesome model zoo.
★ 13k
SimpleMem
SimpleMem: Efficient Lifelong Memory for LLM Agents — Text & Multimodal
★ 3.7k
vortex
An extensible, state-of-the-art framework for columnar compression, and the fastest FOSS columnar file…
★ 3.1k
aimet
AIMET is a library that provides advanced quantization and compression techniques for trained neural network…
★ 2.7k
bolt
10x faster matrix and vector operations
★ 2.5k
Awesome-Efficient-LLM
A curated list for Efficient Large Language Models
★ 2k
CompressAI
A PyTorch library and evaluation platform for end-to-end compression research
★ 1.6k
model-optimization
A toolkit to optimize ML models for deployment for Keras and TensorFlow, including quantization and pruning.
★ 1.6k
Awesome-Knowledge-Distillation-of-LLMs
This repository collects papers for "A Survey on Knowledge Distillation of Large Language Models". We break…
★ 1.3k
nncf
Neural Network Compression Framework for enhanced OpenVINO™ inference
★ 1.2k
MindPipe
A powerful model compression framework for LLMs and LVLMs, adapted for NVIDIA GPUs and Huawei Ascend NPUs.
★ 1k
swin2sr
[ECCV] Swin2SR: SwinV2 Transformer for Compressed Image Super-Resolution and Restoration. Advances in Image…
★ 690
DFloat11
DFloat11 [NeurIPS '25]: Lossless Compression of LLMs and DiTs for Efficient GPU Inference
★ 652
awesome_computer_science
The complete syllabus of Computer Science and Engineering. Roadmap, Checklist for Beginners.
★ 573
wavemap
Fast, efficient and accurate multi-resolution, multi-sensor 3D occupancy mapping
★ 567
KVQuant
[NeurIPS 2024] KVQuant: Towards 10 Million Context Length LLM Inference with KV Cache Quantization
★ 430
Context-Engine
Context-Engine MCP - Agentic Context Compression Suite
★ 402
BK-SDM
A Compressed Stable Diffusion for Efficient Text-to-Image Generation [ECCV'24]
★ 320
picollm
On-device LLM Inference Powered by X-Bit Quantization
★ 315
distill
Context intelligence layer for LLM agents: persistent memory with write-time dedup, sensitivity tagging,…
★ 174
🔗 Familles voisines

Mesuré à partir des topics GitHub communs aux deux projets, pondérés par leur rareté.