data-engineering

28 Projekte teilen dieses GitHub-Topic

data-engineering — Made-With-ML ★48.8kdata-engineeringkestra — ★28kairbyte — ★22kdagster — ★16.1ksemantica — ★11.5kcocoindex — ★11.4kmage-ai — ★8.8kunstract — ★7.2kfeast — ★7.1kDeepAnalyze — ★4.6ksql-translator — ★4.3kqsv — ★3.7kmlops-course — ★3.4ksecond-brain-ai-assistant-course — ★3.1khamilton — ★2.6kSAG — ★2.4kmetarank — ★2.4kbytewax — ★2kfeathr — ★1.9kmlrun — ★1.7kktx — ★1.6kduckle — ★1.3kactive_workflow — ★866NeumAI — ★864agents — ★430Comprehensive-Data-Science-AI-Project-Portfolio — ★333extractor — ★320hflow — ★202kestra★ 28kairbyte★ 22kdagster★ 16.1ksemantica★ 11.5kcocoindex★ 11.4kmage-ai★ 8.8kunstract★ 7.2kfeast★ 7.1kDeepAnalyze★ 4.6ksql-translator★ 4.3kqsv★ 3.7kmlops-course★ 3.4ksecond-brain-ai-assistan…★ 3.1khamilton★ 2.6kSAG★ 2.4kmetarank★ 2.4kbytewax★ 2kfeathr★ 1.9kmlrun★ 1.7kktx★ 1.6kduckle★ 1.3kactive_workflow★ 866NeumAI★ 864agents★ 430Comprehensive-Data-Scien…★ 333extractor★ 320hflow★ 202

Linien verbinden Mitglieder, die messbar miteinander verwandt sind. Die Punktgröße spiegelt die Sterne wider.

🧬 Mitglieder
Made-With-ML
Learn how to develop, deploy and iterate on production-grade ML applications.
★ 48.8k
kestra
Event Driven Orchestration & Scheduling Platform for Mission Critical Applications
★ 28k
airbyte
Open-source data movement for ELT pipelines and AI agents — from APIs, databases & files to warehouses,…
★ 22k
dagster
An orchestration platform for the development, production, and observation of data assets.
★ 16.1k
semantica
Graph-Native Infrastructure for Context and Accountable AI Systems
★ 11.5k
cocoindex
Incremental engine for long horizon agents 🌟 Star if you like it!
★ 11.4k
mage-ai
🧙 Build, run, and manage data pipelines for integrating and transforming data.
★ 8.8k
unstract
LLM-Driven Extraction of Unstructured Data — Built for API Deployments & ETL Pipeline Workflows
★ 7.2k
feast
The Open Source Feature Store for AI/ML
★ 7.1k
DeepAnalyze
★ 4.6k
sql-translator
SQL Translator is a tool for converting natural language queries into SQL code using artificial intelligence.…
★ 4.3k
qsv
Blazing-fast Data-Wrangling toolkit
★ 3.7k
mlops-course
Learn how to design, develop, deploy and iterate on production-grade ML applications.
★ 3.4k
second-brain-ai-assistant-course
Learn to build your Second Brain AI assistant with LLMs, agents, RAG, fine-tuning, LLMOps and AI systems…
★ 3.1k
hamilton
Apache Hamilton helps data scientists and engineers define testable, modular, self-documenting dataflows,…
★ 2.6k
SAG
A new SOTA for RAG — an original retrieval architecture and an open-source knowledge base for humans and…
★ 2.4k
metarank
A low code Machine Learning personalized ranking service for articles, listings, search results,…
★ 2.4k
bytewax
Python Stream Processing
★ 2k
feathr
Feathr – A scalable, unified data and AI engineering platform for enterprise
★ 1.9k
mlrun
MLRun is an open source MLOps platform for quickly building and managing continuous ML applications across…
★ 1.7k
ktx
ktx is an executable context layer for data and analytics agents 🐙 Allow Claude Code, Codex, or other AI…
★ 1.6k
duckle
Open-source ETL/ELT you deploy on your own servers or cloud. Built on DuckDB: no-code/low-code visual…
★ 1.3k
active_workflow
Polyglot workflows without leaving the comfort of your technology stack.
★ 866
NeumAI
Neum AI is a best-in-class framework to manage the creation and synchronization of vector embeddings at large…
★ 864
agents
AI agent tooling for data engineering workflows.
★ 430
Comprehensive-Data-Science-AI-Project-Portfolio
A curated collection of AI, data engineering, and DevOps projects featuring real-world applications, advanced…
★ 333
extractor
Use LLMs to robustly extract web data
★ 320
hflow
SDK for robotics teams to verify the quality of their data used for AI model training.
★ 202
🔗 Verwandte Familien

Gemessen anhand der von beiden Projekten geteilten GitHub-Themen, gewichtet nach der Seltenheit jedes Themas.