data-quality

18 projets partagent ce topic GitHub

data-quality — Made-With-ML ★48.9kdata-qualityapplied-ml — ★30kOpenMetadata — ★14.6kfg-data-profiling — ★13.7kcleanlab — ★11.6kfiftyone — ★10.9kevidently — ★7.8kfeast — ★7.2kmlops-course — ★3.4kwhylogs — ★2.8kfeatureform — ★2kfeathr — ★1.9kCurator — ★1.7kcleanvision — ★1.2kduckle — ★885Failed-ML — ★753dingo — ★730mega-data-factory — ★370applied-ml★ 30kOpenMetadata★ 14.6kfg-data-profiling★ 13.7kcleanlab★ 11.6kfiftyone★ 10.9kevidently★ 7.8kfeast★ 7.2kmlops-course★ 3.4kwhylogs★ 2.8kfeatureform★ 2kfeathr★ 1.9kCurator★ 1.7kcleanvision★ 1.2kduckle★ 885Failed-ML★ 753dingo★ 730mega-data-factory★ 370

Les traits relient les membres réellement apparentés entre eux. La taille des points suit les étoiles.

🧬 Membres
Made-With-ML
Learn how to develop, deploy and iterate on production-grade ML applications.
★ 48.9k
applied-ml
📚 Papers & tech blogs by companies sharing their work on data science & machine learning in production.
★ 30k
OpenMetadata
The Open Context Layer for Data and AI , OpenMetadata is the open platform for building trusted data context…
★ 14.6k
fg-data-profiling
1 Line of code data quality profiling & exploratory data analysis for Pandas and Spark DataFrames.
★ 13.7k
cleanlab
Cleanlab's open-source library is the standard data-centric AI package for data quality and machine learning…
★ 11.6k
fiftyone
Refine high-quality datasets and visual AI models
★ 10.9k
evidently
Evidently is ​​an open-source ML and LLM observability framework. Evaluate, test, and monitor any…
★ 7.8k
feast
The Open Source Feature Store for AI/ML
★ 7.2k
mlops-course
Learn how to design, develop, deploy and iterate on production-grade ML applications.
★ 3.4k
whylogs
An open-source data logging library for machine learning models and data pipelines. 📚 Provides visibility…
★ 2.8k
featureform
The Virtual Feature Store. Turn your existing data infrastructure into a feature store.
★ 2k
feathr
Feathr – A scalable, unified data and AI engineering platform for enterprise
★ 1.9k
Curator
Scalable data pre processing and curation toolkit for LLMs
★ 1.7k
cleanvision
Automatically find issues in image datasets and practice data-centric computer vision.
★ 1.2k
duckle
Open-source ETL/ELT on DuckDB. Write, wire, or draw one pipeline: 350+ components, 160+ connectors, dbt, CDC,…
★ 885
Failed-ML
Compilation of high-profile real-world examples of failed machine learning projects
★ 753
dingo
Dingo: A Comprehensive AI Data, Model and Application Quality Evaluation Tool
★ 730
mega-data-factory
🏭 Mega Scale Multimodal DataPipeline for SOTA Foundation Models
★ 370
🔗 Familles voisines

Mesuré à partir des topics GitHub communs aux deux projets, pondérés par leur rareté.