Data radar: unclassified entries
Momentum based on available 7-day snapshots for Data: measured with two snapshots, estimated otherwise. Each value shows its own source window. This page reads stored snapshots only.
Radar filters
Tool ranking
Mixed ranking: measured growth takes priority; entries without two snapshots are estimated. Raw values keep their source-specific window, shown under each tool; a change requires at least two snapshots from the last 7 days.
- 1Active
spark-vllm-docker
OtherDocker configuration for running VLLM on dual DGX Sparks
githubmeasured growthOpen source ↗
2 211GitHub stars+19 (+0.87 %) - 2Active
steampipe
OtherZero-ETL, infinite possibilities. Live query APIs, code & more with SQL. No DB required.
githubmeasured growthOpen source ↗
7 948GitHub stars+13 (+0.16 %) - 32 074GitHub stars+10 (+0.48 %)
- 4Active
redash
OtherMake Your Company Data Driven. Connect to any data source, easily visualize, dashboard and share your data.
githubmeasured growthOpen source ↗
28 778GitHub stars+10 (+0.03 %) - 5Active
devlake
OtherApache DevLake is an open-source dev data platform to ingest, analyze, and visualize the fragmented data from DevOps tools, extracting insights for engineering excellence, developer experience, and community growth.
githubmeasured growthOpen source ↗
3 119GitHub stars+3 (+0.10 %) - 6Active
jaffle-shop
Other🥪🦘 An open source sandbox project exploring dbt workflows via a fictional sandwich shop's data.
githubmeasured growthOpen source ↗
346GitHub stars+1 (+0.29 %) - 7376GitHub stars+1 (+0.27 %)
- 81 797GitHub stars+1 (+0.06 %)
- 9Active
gatk
OtherOfficial code repository for GATK versions 4 and up
githubmeasured growthOpen source ↗
1 993GitHub stars+1 (+0.05 %) - 10Active
kyuubi
OtherApache Kyuubi is a distributed and multi-tenant gateway to provide serverless SQL on data warehouses and lakehouses.
githubmeasured growthOpen source ↗
2 364GitHub stars+1 (+0.04 %) - 114 160GitHub starsstable
- 12Active
deequ
OtherDeequ is a library built on top of Apache Spark for defining "unit tests for data", which measure data quality in large datasets.
githubmeasured growthOpen source ↗
3 642GitHub starsstable - 13Active
ytsaurus
OtherYTsaurus is a scalable and fault-tolerant open-source big data platform.
githubmeasured growthOpen source ↗
2 203GitHub starsstable - 14Active
This extension makes vscode seamlessly work with dbt™: Auto-complete, preview, column lineage, AI docs generation, health checks, cost estimation etc
githubmeasured growthOpen source ↗
584GitHub starsstable - 15Active
This package contains macros and models to find DAG issues automatically
githubmeasured growthOpen source ↗
574GitHub starsstable - 16Active
tuva-core
OtherMain repo including core data model, data marts, data quality tests, and terminology sets.
githubmeasured growthOpen source ↗
325GitHub starsstable - 17255GitHub starsstable
- 18246GitHub starsstable
- 19231GitHub starsstable
- 20Active
RoaringBitmap
OtherA better compressed bitset in Java: used by Apache Spark, Netflix Atlas, Apache Pinot, Tablesaw, and many others
githubestimated momentumOpen source ↗
3 922GitHub stars— - 21Active
zio-quill
OtherCompile-time Language Integrated Queries for Scala
githubestimated momentumOpen source ↗
2 166GitHub stars— - 22Active
auron
OtherThe Auron accelerator for distributed computing framework (e.g., Spark) leverages native vectorized execution to accelerate query processing
githubestimated momentumOpen source ↗
1 795GitHub stars— - 23612GitHub stars—
Learning resources
Ranked by measured growth, normalized across sources. These resources remain available separately and do not take part in the main tool ranking.
- 1ActiveOpen source ↗
alluxio
OtherAlluxio, data orchestration for analytics and machine learning in the cloud
githubResourcemeasured growth7 234GitHub starsstable - 2ActiveOpen source ↗
awesome-dbt
OtherA curated list of awesome dbt resources
githubResourcemeasured growth1 721GitHub stars-1 (-0.06 %)