gguf

26 projets partagent ce topic GitHub

gguf — llmfit ★30.9kggufllamafile — ★25.5kwhichllm — ★6kshimmy — ★5.7kOGAM — ★2.8kmaid — ★2.6khandy-ollama — ★2.5kcomfyui_LLM_party — ★2.3knode-llama-cpp — ★2.1kgollama — ★1.8kauto-round — ★1.5kFun-ASR — ★1.4kOuteTTS — ★1.4kAtomic-Chat — ★1.2katomic-agent — ★1.1klocally-uncensored — ★980llama3.java — ★816Box — ★709Audar-ASR-V1 — ★560catai — ★498gerbil — ★472overlay-translator — ★424mlx-serve — ★380pmetal — ★306Soup — ★80FerryAI — ★56llamafile★ 25.5kwhichllm★ 6kshimmy★ 5.7kOGAM★ 2.8kmaid★ 2.6khandy-ollama★ 2.5kcomfyui_LLM_party★ 2.3knode-llama-cpp★ 2.1kgollama★ 1.8kauto-round★ 1.5kFun-ASR★ 1.4kOuteTTS★ 1.4kAtomic-Chat★ 1.2katomic-agent★ 1.1klocally-uncensored★ 980llama3.java★ 816Box★ 709Audar-ASR-V1★ 560catai★ 498gerbil★ 472overlay-translator★ 424mlx-serve★ 380pmetal★ 306Soup★ 80 · GitHub ↗FerryAI★ 56 · GitHub ↗

Les traits relient les membres réellement apparentés entre eux. La taille des points suit les étoiles.

🧬 Membres
llmfit
Hundreds of models & providers. One command to find what runs on your hardware.
★ 30.9k
llamafile
Distribute and run LLMs with a single file.
★ 25.5k
whichllm
Find the local LLM that actually runs and performs best on your hardware. Ranked by real, recency-aware…
★ 6k
shimmy
⚡ Pure-Rust WebGPU inference engine — OpenAI-API compatible, GGUF native, runs on any GPU. No Python. No…
★ 5.7k
OGAM
The Swiss Army Knife of Offline AI. Chat, see, speak, and generate images on your phone or Mac — GGUF LLMs,…
★ 2.8k
maid
Maid is a free and open source application for interfacing with llama.cpp models locally, and with Anthropic,…
★ 2.6k
handy-ollama
动手学Ollama,CPU玩转大模型部署,在线阅读地址:https://datawhalechina.github.io/handy-ollam…
★ 2.5k
comfyui_LLM_party
LLM Agent Framework in ComfyUI includes MCP sever, Omost,GPT-sovits, ChatTTS,GOT-OCR2.0, and FLUX prompt…
★ 2.3k
node-llama-cpp
Run AI models locally on your machine with node.js bindings for llama.cpp. Enforce a JSON schema on the model…
★ 2.1k
gollama
Go manage your Ollama models
★ 1.8k
auto-round
A SOTA quantization algorithm for high-accuracy low-bit LLM inference, seamlessly optimized for CPU/XPU/CUDA,…
★ 1.5k
Fun-ASR
Open-source LLM-based ASR model family for Chinese, dialect, accent, and multilingual speech, with FunASR,…
★ 1.4k
OuteTTS
Interface for OuteTTS models.
★ 1.4k
Atomic-Chat
Local AI app and inference engine for agents. Run open-weight LLMs locally — private, 100% offline on your…
★ 1.2k
atomic-agent
Local First Ai Agent. Optimized for Local Ai models. Long context window. Proper tools callings. Runs…
★ 1.1k
locally-uncensored
Plug-and-play local AI studio: uncensored chat, image & video generation, coding agent. Runs abliterated LLMs…
★ 980
llama3.java
Llama 3+ inference in pure Java
★ 816
Box
The most advanced, fully offline client-side AI suite on Android today.
★ 709
Audar-ASR-V1
Arabic-first generative speech recognition — Audar-ASR-V1 (Flash + Turbo). #1 on the Open Universal Arabic…
★ 560
catai
Run AI ✨ assistant locally! with simple API for Node.js 🚀
★ 498
gerbil
A desktop app for running Large Language Models locally.
★ 472
overlay-translator
★ 424
mlx-serve
Native LLM inference server for Apple Silicon. OpenAI + Anthropic API compatible. No Python. Includes MLX…
★ 380
pmetal
PMetal: high-performance Apple Silicon framework for local LLM inference, LoRA/QLoRA fine-tuning, serving,…
★ 306
Soup
Soup turns the pain of LLM fine-tuning into a simple workflow. One config, one command, done.
★ 80 · GitHub ↗
FerryAI
Native AI inference for PHP 8.3+ - run ONNX, GGUF (llama.cpp) and RubixML models directly in your PHP process…
★ 56 · GitHub ↗
🔗 Familles voisines

Mesuré à partir des topics GitHub communs aux deux projets, pondérés par leur rareté.