Skip to content
CAI
Produce a survey ↗Verify a survey

apache/druid

Apache Druid is a distributed, columnar, real-time analytics database designed for high-performance data ingestion and interactive querying. The system supports a wide variety of data sources, including Kafka, Kinesis, HDFS, S3, and various cloud storage backends, while offering flexible storage options such as S3, GCS, Cassandra, and SQL databases. It provides extensive aggregation capabilities through DataSketches, Bloom filters, and histogram aggregations, enabling fast approximate and exact statistical analysis. Additionally, the platform features a comprehensive security model with LDAP and database-backed authentication, role-based authorization, and support for Kerberos and JWT-based access control.

46.8

Weak · 29 July 2026

967k

lines of production code

Java

with TypeScript

11

bus factor · 765 authors in all

3

measurements over time

CAI band scale
CAI trend line

How it got here

2012–2018 · Security, storage, and statistical extensions

This period focused on hardening the platform with comprehensive security features, including basic authentication, LDAP support, and Kerberos integration. It also expanded data handling capabilities by adding support for diverse storage backends like PostgreSQL, Cassandra, and S3, while introducing advanced statistical aggregations via DataSketches and histogram extensions.

93 changes

2019–2020 · cloud storage and security extensions

This period focused on expanding cloud storage support by adding extensions for Google Cloud Storage, Aliyun OSS, and HDFS, alongside new aggregators like Bloom Filters and t-digest sketches. The release also introduced significant security enhancements, including LDAP and OIDC authentication, while providing comprehensive configuration templates and Docker setups for easier deployment.

50 changes

2021–2026 · Extensibility and Observability

This period focused on expanding Druid's ecosystem by introducing new connectors for Iceberg, Delta Lake, and RabbitMQ, alongside advanced aggregation types like compressed BigDecimals and DDsketches. Concurrently, the project significantly enhanced observability through OpenTelemetry and Prometheus integrations, while establishing a robust testing infrastructure with Docker-based containers and SQL-level validation tools.

46 changes

CAI lens gauges

Survey your own repository

apache/druid was measured the same way every project in this corpus was: the same rubric, at a pinned commit, with the result published in full. Point the surveyor at a repository you know and see whether you agree with it.

Survey a repository

About this page

  • The description of this project is derived from its own commit history, not from its README.
  • The score is its highest published measurement, taken on 29 July 2026 at a pinned commit. It is not a live figure and does not change until the project is measured again.
  • Measured at commit a7bcc37897 — the exact code this score is about.
  • Scored under rubric rubric-2026.08.18. Score the same commit under that rubric and you get the same number.
CAI link cards