gpu

59 Projekte teilen dieses GitHub-Topic

gpu — pytorch ★101.7kgpufastai — ★28.1kWebGL-Fluid-Simulation — ★16.5ktvm — ★13.6kflashinfer — ★6.3klemonade — ★5.6kcuml — ★5.2kkoharu — ★4.9kMegEngine — ★4.8kexecutorch — ★4.8ktiny-cuda-nn — ★4.5kllm-d — ★4.4kdeepflow — ★4.3kjittor — ★3.2kheavydb — ★3.1kchitu — ★3kleptonai — ★2.8kdifftaichi — ★2.7ktorchrec — ★2.6kdeepdetect — ★2.6kopenlake — ★2.4ktrainer — ★2.2knode-llama-cpp — ★2.2kbeta9 — ★1.8kglim — ★1.7ktt-metal — ★1.7kgpu-hot — ★1.6kuccl — ★1.5kgpu-io — ★1.5kisaac_ros_visual_slam — ★1.4kgpu_poor — ★1.4kmodal-examples — ★1.2kwebgl-wind — ★1.1kTalos — ★988femtoGPT — ★935GPUMD — ★811tty7 — ★804can-i-finetune-this — ★790isaac_ros_nvblox — ★724DFloat11 — ★642VerletIntegration — ★581fastai★ 28.1kWebGL-Fluid-Simulation★ 16.5ktvm★ 13.6kflashinfer★ 6.3klemonade★ 5.6kcuml★ 5.2kkoharu★ 4.9kMegEngine★ 4.8kexecutorch★ 4.8ktiny-cuda-nn★ 4.5kllm-d★ 4.4kdeepflow★ 4.3kjittor★ 3.2kheavydb★ 3.1kchitu★ 3kleptonai★ 2.8kdifftaichi★ 2.7ktorchrec★ 2.6kdeepdetect★ 2.6kopenlake★ 2.4ktrainer★ 2.2knode-llama-cpp★ 2.2kbeta9★ 1.8kglim★ 1.7ktt-metal★ 1.7kgpu-hot★ 1.6kuccl★ 1.5kgpu-io★ 1.5kisaac_ros_visual_slam★ 1.4kgpu_poor★ 1.4kmodal-examples★ 1.2kwebgl-wind★ 1.1kTalos★ 988femtoGPT★ 935GPUMD★ 811tty7★ 804can-i-finetune-this★ 790isaac_ros_nvblox★ 724DFloat11★ 642VerletIntegration★ 581

Linien verbinden Mitglieder, die messbar miteinander verwandt sind. Die Punktgröße spiegelt die Sterne wider.

🧬 Mitglieder
pytorch
Tensors and Dynamic neural networks in Python with strong GPU acceleration
★ 101.7k
fastai
The fastai deep learning library
★ 28.1k
WebGL-Fluid-Simulation
Play with fluids in your browser (works even on mobile)
★ 16.5k
tvm
Open Machine Learning Compiler Framework
★ 13.6k
flashinfer
FlashInfer: Kernel Library for LLM Serving
★ 6.3k
lemonade
Lemonade helps users discover and run local AI apps by serving optimized LLMs right from their own GPUs and…
★ 5.6k
cuml
cuML - RAPIDS Machine Learning Library
★ 5.2k
koharu
ML-powered manga translator, written in Rust.
★ 4.9k
MegEngine
MegEngine 是一个快速、可拓展、易于使用且支持自动求导的深度学习框架
★ 4.8k
executorch
On-device AI across mobile, embedded and edge for PyTorch
★ 4.8k
tiny-cuda-nn
Lightning fast C++/CUDA neural network framework
★ 4.5k
llm-d
Achieve state of the art inference performance with modern accelerators on Kubernetes
★ 4.4k
deepflow
eBPF Observability - Distributed Tracing and Profiling
★ 4.3k
jittor
Jittor is a high-performance deep learning framework based on JIT compiling and meta-operators.
★ 3.2k
heavydb
HeavyDB (formerly MapD/OmniSciDB)
★ 3.1k
chitu
High-performance inference framework for large language models, focusing on efficiency, flexibility, and…
★ 3k
leptonai
A Pythonic framework to simplify AI service building
★ 2.8k
difftaichi
10 differentiable physical simulators built with Taichi differentiable programming (DiffTaichi, ICLR 2020)
★ 2.7k
torchrec
Pytorch domain library for recommendation systems
★ 2.6k
deepdetect
Deep Learning Server and CLI for Torch and TensorRT
★ 2.6k
openlake
OpenLake is a high performance storage engine for efficient LLM inference and GPU Training
★ 2.4k
trainer
Distributed AI Model Training and LLM Fine-Tuning on Kubernetes
★ 2.2k
node-llama-cpp
Run AI models locally on your machine with node.js bindings for llama.cpp. Enforce a JSON schema on the model…
★ 2.2k
beta9
Ultrafast serverless GPU inference, sandboxes, and background jobs
★ 1.8k
glim
GLIM: versatile and extensible point cloud-based 3D localization and mapping framework
★ 1.7k
tt-metal
:metal: TT-NN operator library, and TT-Metalium low level kernel programming model.
★ 1.7k
gpu-hot
🔥 Real-time NVIDIA GPU dashboard
★ 1.6k
uccl
UCCL is an efficient communication library for GPUs, covering collectives, P2P (e.g., KV cache transfer, RL…
★ 1.5k
gpu-io
A GPU-accelerated computing library for running physics simulations and other GPGPU computations in a web…
★ 1.5k
isaac_ros_visual_slam
Visual SLAM/odometry package based on NVIDIA-accelerated cuVSLAM
★ 1.4k
gpu_poor
Calculate token/s & GPU memory requirement for any LLM. Supports llama.cpp/ggml/bnb/QLoRA quantization
★ 1.4k
modal-examples
Examples of programs built using Modal
★ 1.2k
webgl-wind
Wind power visualization with WebGL particles
★ 1.1k
Talos
GPU worker client for the Talos network. Pairs with your Talos account, serves open-model inference jobs over…
★ 988
femtoGPT
Pure Rust implementation of a minimal Generative Pretrained Transformer
★ 935
GPUMD
Graphics Processing Units Molecular Dynamics
★ 811
tty7
A terminal workbench in pure Rust: shells, persistent sessions, SSH, coding agents. GPU-rendered on Zed's…
★ 804
can-i-finetune-this
Estimate whether a Hugging Face model fits and fine-tunes on your local GPU.
★ 790
isaac_ros_nvblox
NVIDIA-accelerated 3D scene reconstruction and Nav2 local costmap provider using nvblox
★ 724
DFloat11
DFloat11 [NeurIPS '25]: Lossless Compression of LLMs and DiTs for Efficient GPU Inference
★ 642
VerletIntegration
A real-time particle simulation that uses Verlet Integration
★ 581
neurokernel
Neurokernel Project
★ 565
popsift
PopSift is an implementation of the SIFT algorithm in CUDA.
★ 498
blub
3D fluid simulation experiments in Rust, using WebGPU-rs (WIP)
★ 483
isaac_ros_pose_estimation
Deep learned, NVIDIA-accelerated 3D object pose estimation
★ 475
awsome-distributed-ai
Collection of best practices, reference architectures, model training examples and utilities to train large…
★ 473
warpx
WarpX is an advanced Particle-In-Cell code.
★ 470
zinc
Zig INferenCe Engine — Local LLM inference on AMD GPUs and Apple Silicon
★ 467
cucim
cuCIM - RAPIDS GPU-accelerated image processing library
★ 463
JetStream
JetStream is a throughput and memory optimized engine for LLM inference on XLA devices, starting with TPUs…
★ 456
hoomd-blue
Molecular dynamics and Monte Carlo soft matter simulation on GPUs.
★ 444
rag-chatbot
RAG (Retrieval-augmented generation) ChatBot that provides answers based on contextual information extracted…
★ 433
WaterBall
Fluid simulation on a sphere🌏
★ 390
knowhere
Vector search engine inside Milvus, integrating FAISS, HNSW, DiskANN.
★ 371
SPH_Taichi
A high-performance implementation of SPH in Taichi.
★ 319
mppi_numba
A GPU implementation of Model Predictive Path Integral (MPPI) control that uses a probabilistic…
★ 310
isaac_ros_common
Common utilities, packages, scripts, and testing infrastructure for Isaac ROS packages.
★ 307
attyx
GPU accelerated terminal for agentic workflows
★ 239
agentfm-core
AgentFM is a peer-to-peer network that turns everyday computers into a decentralized AI supercomputer.…
★ 135
🔗 Verwandte Familien

Gemessen anhand der von beiden Projekten geteilten GitHub-Themen, gewichtet nach der Seltenheit jedes Themas.