embeddings

113 proyectos comparten este topic de GitHub

embeddings — supabase ★107.2kembeddingsclaude-mem — ★89kRAG_Techniques — ★28.9kWeKnora — ★19.1kturbovec — ★14.5ktxtai — ★12.8klangchain4j — ★12.7kInsForge — ★12.5kFlagEmbedding — ★12kseatunnel — ★9.5klance — ★6.9kpostgresml — ★6.8kpytorch-metric-learning — ★6.3khoncho — ★6.3ksemble — ★5.7kDaft — ★5.7ktext-embeddings-inference — ★5kAutoRAG — ★5kOpenMemory — ★4.4kruby_llm — ★4.3klightly — ★3.8khub — ★3.5kawesome-generative-ai — ★3.5ktowhee — ★3.5kmteb — ★3.4kSocratiCode — ★3.2kfastembed — ★3.1kprompttools — ★3kml-surveys — ★2.9kbootcamp — ★2.4kmemsearch — ★2.4ksie — ★2.4kailia-models — ★2.4kvearch — ★2.3kgenerative-ai-docs — ★2.3kmodel2vec — ★2.2kReadAny — ★2.1kagentset — ★2kfeatureform — ★2kDot — ★1.9kGPTDiscord — ★1.9kclaude-mem★ 89kRAG_Techniques★ 28.9kWeKnora★ 19.1kturbovec★ 14.5ktxtai★ 12.8klangchain4j★ 12.7kInsForge★ 12.5kFlagEmbedding★ 12kseatunnel★ 9.5klance★ 6.9kpostgresml★ 6.8kpytorch-metric-learning★ 6.3khoncho★ 6.3ksemble★ 5.7kDaft★ 5.7ktext-embeddings-inferenc…★ 5kAutoRAG★ 5kOpenMemory★ 4.4kruby_llm★ 4.3klightly★ 3.8khub★ 3.5kawesome-generative-ai★ 3.5ktowhee★ 3.5kmteb★ 3.4kSocratiCode★ 3.2kfastembed★ 3.1kprompttools★ 3kml-surveys★ 2.9kbootcamp★ 2.4kmemsearch★ 2.4ksie★ 2.4kailia-models★ 2.4kvearch★ 2.3kgenerative-ai-docs★ 2.3kmodel2vec★ 2.2kReadAny★ 2.1kagentset★ 2kfeatureform★ 2kDot★ 1.9kGPTDiscord★ 1.9k

Las líneas conectan a los miembros que están mediblemente relacionados entre sí. El tamaño de los puntos refleja las estrellas.

🧬 Miembros
supabase
The Postgres development platform. Supabase gives you a dedicated Postgres database to build your web,…
★ 107.2k
claude-mem
Persistent Context Across Sessions for Every Agent – Captures everything your agent does during sessions,…
★ 89k
RAG_Techniques
This repository showcases various advanced techniques for Retrieval-Augmented Generation (RAG) systems. Each…
★ 28.9k
WeKnora
Open-source LLM knowledge platform: turn raw documents into a queryable RAG, an autonomous reasoning agent,…
★ 19.1k
turbovec
A vector index built on TurboQuant, written in Rust with Python bindings
★ 14.5k
txtai
💡 All-in-one AI framework for semantic search, LLM orchestration and language model workflows
★ 12.8k
langchain4j
LangChain4j is an idiomatic, open-source Java library for building LLM-powered applications on the JVM. It…
★ 12.7k
InsForge
The all-in-one, open-source backend platform for agentic coding. InsForge gives your coding agent database,…
★ 12.5k
FlagEmbedding
Retrieval and Retrieval-augmented LLMs
★ 12k
seatunnel
SeaTunnel is a multimodal, high-performance, distributed, massive data integration tool.
★ 9.5k
lance
Open Lakehouse Format for Multimodal AI. Convert from Parquet in 2 lines of code for 100x faster random…
★ 6.9k
postgresml
Postgres with GPUs for ML/AI apps.
★ 6.8k
pytorch-metric-learning
The easiest way to use deep metric learning in your application. Modular, flexible, and extensible. Written…
★ 6.3k
honcho
Memory library for building stateful agents
★ 6.3k
semble
Fast and Accurate Code Search for Agents. Uses ~98% fewer tokens than grep+read
★ 5.7k
Daft
High-performance data engine for AI and multimodal workloads. Process images, audio, video, and structured…
★ 5.7k
text-embeddings-inference
A blazing fast inference solution for text embeddings models
★ 5k
AutoRAG
AutoRAG: Now your agent can find anything in your computer. It gets smarter if you are using it frequently.
★ 5k
OpenMemory
Local persistent memory store for LLM applications including claude desktop, github copilot, codex,…
★ 4.4k
ruby_llm
One delightful Ruby framework for every major AI provider. Build AI agents, chatbots, RAG apps, and…
★ 4.3k
lightly
A python library for self-supervised learning on images.
★ 3.8k
hub
A library for transfer learning by reusing parts of TensorFlow models.
★ 3.5k
awesome-generative-ai
A curated list of Generative AI tools, works, models, and references
★ 3.5k
towhee
Towhee is a framework that is dedicated to making neural data processing pipelines simple and fast.
★ 3.5k
mteb
MTEB: State-of-the-art evaluation of embeddings across languages and modalities
★ 3.4k
SocratiCode
Enterprise-grade (40m+ LOC) codebase intelligence, zero-setup, local & private Plugin/Skill/Extension or MCP:…
★ 3.2k
fastembed
Fast, Accurate, Lightweight Python library to make State of the Art Embedding
★ 3.1k
prompttools
Open-source tools for prompt testing and experimentation, with support for both LLMs (e.g. OpenAI, LLaMA) and…
★ 3k
ml-surveys
📋 Survey papers summarizing advances in deep learning, NLP, CV, graphs, reinforcement learning,…
★ 2.9k
bootcamp
Dealing with all unstructured data, such as reverse image search, audio search, molecular search, video…
★ 2.4k
memsearch
A persistent, unified memory layer for all your AI agents (e.g. Claude Code, Codex), backed by Markdown and…
★ 2.4k
sie
Open-source inference server and production cluster for all the models your agent needs.
★ 2.4k
ailia-models
The collection of pre-trained, state-of-the-art AI models for ailia SDK
★ 2.4k
vearch
Distributed vector search for AI-native applications
★ 2.3k
generative-ai-docs
This repository is deprecated and will be archived
★ 2.3k
model2vec
Fast State-of-the-Art Static Embeddings
★ 2.2k
ReadAny
AI-powered cross-platform e-book reader with semantic search, RAG chat, local vector store, notes, TTS, and…
★ 2.1k
agentset
The open-source RAG platform: built-in citations, deep research, 22+ file formats, partitions, MCP server,…
★ 2k
featureform
The Virtual Feature Store. Turn your existing data infrastructure into a feature store.
★ 2k
Dot
Text-To-Speech, RAG, and LLMs. All local!
★ 1.9k
GPTDiscord
A robust, all-in-one GPT interface for Discord. ChatGPT-style conversations, image generation, AI-moderation,…
★ 1.9k
awesome-embedding-models
A curated list of awesome embedding models tutorials, projects and communities.
★ 1.8k
grepai
Semantic Search & Call Graphs for AI Agents (100% Local)
★ 1.8k
ModernBERT
Bringing BERT into modernity via both architecture changes and scaling
★ 1.7k
LLPhant
LLPhant - A comprehensive PHP Generative AI Framework using OpenAI GPT 4. Inspired by Langchain
★ 1.7k
magnitude
A fast, efficient universal vector embedding utility package.
★ 1.7k
lightly-train
All-in-one training for vision models (YOLO, ViTs, RT-DETR, DINOv3): pretraining, fine-tuning, distillation.
★ 1.6k
hivemind
Hivemind turns your traces into reusable skills across agents
★ 1.5k
amazon-bedrock-samples
This repository contains examples for customers to get started using the Amazon Bedrock Service. This…
★ 1.5k
korvus
Korvus is a search SDK that unifies the entire RAG pipeline in a single database query. Built on top of…
★ 1.5k
Awesome-LLM-RAG
Awesome-LLM-RAG: a curated list of advanced retrieval augmented generation (RAG) in Large Language Models
★ 1.3k
SeaGOAT
local-first semantic code search engine
★ 1.3k
graph
GPU-accelerated force graph layout and rendering
★ 1.2k
obsidian-Smart2Brain
An Obsidian plugin to interact with your privacy focused AI-Assistant making your second brain even smarter!
★ 1.2k
vlms-zero-to-hero
This series will take you on a journey from the fundamentals of NLP and Computer Vision to the cutting edge…
★ 1.2k
chromem-go
Embeddable vector database for Go with Chroma-like interface and zero third-party dependencies. In-memory…
★ 1k
fastembed-rs
Rust library for generating vector embeddings, reranking locally!
★ 978
vectordb-recipes
Resource, examples & tutorials for multimodal AI, RAG and agents using vector search and LLMs
★ 970
vectordb
Epsilla is a high performance Vector Database Management System
★ 875
VectorCode
A code repository indexing tool to supercharge your LLM experience.
★ 874
rag_api
ID-based RAG FastAPI: Integration with Langchain and PostgreSQL/pgvector
★ 871
NeumAI
Neum AI is a best-in-class framework to manage the creation and synchronization of vector embeddings at large…
★ 865
Magick
Magick is a cutting-edge toolkit for a new kind of AI builder. Make Magick with us!
★ 845
lingoose
🪿 LinGoose is a Go framework for building awesome AI/LLM applications.
★ 835
contrastors
Train Models Contrastively in Pytorch
★ 799
llm-apps-java-spring-ai
Samples showing how to build Java applications powered by Generative AI and LLMs using Spring AI and Spring…
★ 766
Azure-AIGEN-demos
Microsoft Foundry (demos, documentation, accelerators).
★ 755
embedJs
A NodeJS RAG framework to easily work with LLMs and embeddings
★ 604
multimodal-agents-course
An MCP Multimodal AI Agent with eyes and ears!
★ 576
AnglE
Train and Infer Powerful Sentence Embeddings with AnglE | 🔥 SOTA on STS and MTEB Leaderboard
★ 573
adaptive-classifier
A flexible, adaptive classification system for dynamic text classification
★ 569
LLM-Kit
🚀WebUI integrated platform for latest LLMs | 各大语言模型的全流程工具 WebUI…
★ 552
embedbase
A dead-simple API to build LLM-powered apps
★ 524
graphrag-rs
GraphRAG-rs is a high-performance, state-of-the-art Rust implementation of GraphRAG (Graph-based Retrieval…
★ 522
chatgpt-your-files
Production-ready MVP for securely chatting with your documents using pgvector
★ 514
ReMind
Your Local Artificial Memory on your Device.
★ 513
restai
RESTai is an AIaaS (AI as a Service) open-source platform. Supports many public and local LLM suported by…
★ 512
ai-template
Mercury - Train your own custom GPT. Chat with any file, or website.
★ 453
octocode
Semantic code searcher and codebase utility
★ 436
tinyvector
A tiny embedding database in pure Rust.
★ 436
gen-cv
Vision AI Solution Accelerator
★ 433
awesome-rag
A curated list of retrieval-augmented generation (RAG) in large language models
★ 429
mlx-embeddings
MLX-Embeddings is the best package for running Vision and Language Embedding models locally on your Mac using…
★ 423
super-rag
Super performant RAG pipelines for AI apps. Summarization, Retrieve/Rerank and Code Interpreters in one…
★ 395
vectorhub
Quickly and easily build AI website or application by using embeddings!
★ 386
embedding_studio
Embedding Studio is a framework which allows you transform your Vector Database into a feature-rich Search…
★ 382
TrueMemory
The memory your AI should have had from the start. Automatic capture, automatic recall, 100% local. One…
★ 370
VectorDB-Plugin
Program that lets you ask questions about your documents, audio, and video files.
★ 369
vicinity
Lightweight Nearest Neighbors with Flexible Backends
★ 349
qvac
QVAC - Local AI SDK and libraries for building private, cross-platform, peer-to-peer AI applications. Run…
★ 337
vectorai
Vector AI — A platform for building vector based applications. Encode, query and analyse data using vectors.
★ 322
DataChad
Ask questions about any data source by leveraging langchains
★ 321
marm-memory
Local-first 3-in-1 AI memory layer & MCP server for Claude Code, Codex, Grok, Gemini, VS Code and Cursor.…
★ 320
rushdb
RushDB is a graph + vector database and memory layer for AI agents. Push any JSON, get typed, searchable,…
★ 319
MegaMemory
Persistent project knowledge graph for coding agents. MCP server with semantic search, in-process embeddings,…
★ 309
text2text
Text2Text Language Modeling Toolkit
★ 304
PharosRAG
Pharos — local-first agentic RAG for your team's document library: multi-format ingest, hybrid retrieval,…
★ 292
Kwipu
Ask questions across your Markdown notes using a fully local Graph RAG engine. Built for Obsidian vaults,…
★ 265
ClawMem
On-device memory layer for AI agents. Claude Code, Hermes and OpenClaw. Hooks + MCP server + hybrid RAG…
★ 195
semble_rs
Fast, AI-agent-native code search in Rust — hybrid BM25 + semantic, Tree-sitter AST chunking, dependency &…
★ 187
VectorRAG.Net
VectorRAG.Net is a .NET-native high-performance vector database library for semantic search and RAG…
★ 169
flowflow
Voice notes for iPhone and macOS - 100% Rust, Dioxus, local-first (SQLite + LanceDB + RIG)
★ 160
self-hosted-ai-stack
Deploy a complete self-hosted AI stack with Docker Compose: Ollama, LiteLLM, AnythingLLM, Whisper,…
★ 129
vault-curate
Hybrid semantic search and AI curation for your vault. Combines BM25 keyword, on-device WebGPU embeddings,…
★ 127
llm-wiki-memory
Local, git-versioned memory for AI coding agents. No RAG, no Docker, no external service. Capture, compile,…
★ 125
agentkeeper
Own your AI memory — import ChatGPT, Claude and Gemini exports, see what each AI knows about you.…
★ 119
mfs
A context harness for AI agents: all your scattered context — code, memory, docs, databases, SaaS — in…
★ 116
engram
🧠 Hybrid long-term memory plugin for OpenClaw agents — SQLite+FTS5 for structured facts, LanceDB for…
★ 103
Vera
Local code search combining BM25, vector similarity, and cross-encoder reranking. Parses 60+ languages with…
★ 101
rag.computer
Self-hostable RAG platform - document ingestion, embedding, and vector search behind a simple REST API
★ 98
OpenDocuments
Self-hosted RAG platform for AI document search across GitHub, Notion, Google Drive, local files, and web…
★ 97
LodeDB
World's fastest and most compact embedded vector database: exact by default, multimodal, local-first, and…
★ 89
FerryAI
Native AI inference for PHP 8.3+ - run ONNX, GGUF (llama.cpp) and RubixML models directly in your PHP process…
★ 56 · GitHub ↗
🔗 Familias relacionadas

Medido a partir de los temas de GitHub compartidos por ambos proyectos, ponderado por cuán raros son cada uno de los temas.