gguf

21 progetti condividono questo topic GitHub

gguf — llmfit ★34.6kggufllamafile — ★25.8kkoboldcpp — ★11.6kshimmy — ★5.6kSoup — ★4.2kOGAM — ★3khandy-ollama — ★2.5katomic-agent — ★2.4knode-llama-cpp — ★2.2kgollama — ★1.8kauto-round — ★1.5kOuteTTS — ★1.4kAtomic-Chat — ★1.4klocally-uncensored — ★1.3kmlx-serve — ★1kllama3.java — ★815Box — ★790gerbil — ★471TensorSharp — ★388halofpx — ★41FerryAI — ★41llamafile★ 25.8kkoboldcpp★ 11.6kshimmy★ 5.6kSoup★ 4.2kOGAM★ 3khandy-ollama★ 2.5katomic-agent★ 2.4knode-llama-cpp★ 2.2kgollama★ 1.8kauto-round★ 1.5kOuteTTS★ 1.4kAtomic-Chat★ 1.4klocally-uncensored★ 1.3kmlx-serve★ 1kllama3.java★ 815Box★ 790gerbil★ 471TensorSharp★ 388halofpx★ 41 · GitHub ↗FerryAI★ 41

Le linee collegano membri che sono misurabilmente correlati tra loro. La dimensione dei punti riflette le stelle.

🧬 Membri
llmfit
Hundreds of models & providers. One command to find what runs on your hardware.
★ 34.6k
llamafile
Distribute and run LLMs with a single file.
★ 25.8k
koboldcpp
Run GGUF models easily with a KoboldAI UI. One File. Zero Install.
★ 11.6k
shimmy
⚡ Pure-Rust WebGPU inference engine — OpenAI-API compatible, GGUF native, runs on any GPU. No Python. No…
★ 5.6k
Soup
Fine-tune LLMs from one YAML. Layer streaming trains an 8B model on a 4 GB laptop GPU.
★ 4.2k
OGAM
The Swiss Army Knife of Offline AI. Chat, see, speak, and generate images on your phone or Mac — GGUF LLMs,…
★ 3k
handy-ollama
动手学Ollama,CPU玩转大模型部署,在线阅读地址:https://datawhalechina.github.io/handy-ollam…
★ 2.5k
atomic-agent
Local First Ai Agent. Optimized for Local Ai models. Long context window. Proper tools callings. Runs…
★ 2.4k
node-llama-cpp
Run AI models locally on your machine with node.js bindings for llama.cpp. Enforce a JSON schema on the model…
★ 2.2k
gollama
Go manage your Ollama models
★ 1.8k
auto-round
A SOTA quantization algorithm for high-accuracy low-bit LLM inference, seamlessly optimized for CPU/XPU/CUDA,…
★ 1.5k
OuteTTS
Interface for OuteTTS models.
★ 1.4k
Atomic-Chat
Local AI app and inference engine for agents. Run open-weight LLMs locally — private, 100% offline on your…
★ 1.4k
locally-uncensored
Plug-and-play local AI studio: uncensored chat, image & video generation, coding agent. Runs abliterated LLMs…
★ 1.3k
mlx-serve
Native LLM inference server for Apple Silicon. OpenAI + Anthropic API compatible. No Python. Includes MLX…
★ 1k
llama3.java
Llama 3+ inference in pure Java
★ 815
Box
The most advanced, fully offline client-side AI suite on Android today.
★ 790
gerbil
A desktop app for running Large Language Models locally.
★ 471
TensorSharp
A native .NET LLM inference engine for GGUF models. TensorSharp provides a console application, a web-based…
★ 388
halofpx
Run Ornith, Qwen, Nemotron & DeepSeek optimized on AMD Strix Halo — unified OpenAI-compatible server with…
★ 41 · GitHub ↗
FerryAI
Native AI inference for PHP 8.3+ - run ONNX, GGUF (llama.cpp) and RubixML models directly in your PHP process…
★ 41
🔗 Famiglie affini

Misurato dai temi di GitHub condivisi da entrambi i progetti, ponderato in base a quanto è raro ciascun tema.