druid

Data processing engine

A high-performance real-time analytics database for fast queries and ingest

Apache Druid: a high performance real-time analytics database.

GitHub

14k stars
584 watching
4k forks
Language: Java
last commit: almost 2 years ago
Linked from 6 awesome lists

druid

Backlinks from these awesome lists:

Related projects:

RepositoryDescriptionStars
apache/pigEnables data processing and transformation in large files using a high-level language with compile-time optimizations for efficient execution on distributed computing frameworks.682
apache/sparkAn analytics engine designed to handle large-scale data processing and analysis40,170
apache/tezA system that enables flexible data processing pipelines using a low-level engine for higher-level frameworks482
apache/datafusion-ballistaDistributed query engine for Apache DataFusion applications1,580
apache/systemdsAn end-to-end data science platform that integrates data integration, machine learning model training, and deployment1,038
asavinov/bistroA general-purpose data analysis engine that processes batch and stream data in a column-oriented manner7
apache/impalaA high-performance query engine designed to handle large-scale data processing and analytics1,164
ivelum/djangoqlAn advanced search library for Django models with auto-completion and support for logical operators and table joins.1,025
allegro/turniloA web application providing a user-friendly interface to explore and visualize data in Apache Druid733
apache/sedonaA software framework that enables developers to process spatial data at any scale within modern cluster computing systems.1,974
apache/samzaA distributed stream processing framework for handling high-volume data streams with fault tolerance and durability guarantees817
h2oai/h2o-2An analytics engine that provides fast and scalable predictive modeling capabilities for big data2,224
webdb-app/appA comprehensive database IDE with features like versioning and data inference, designed to simplify database development and management.192
cloudera/impalaA distributed SQL query engine for analyzing large datasets in Hadoop clusters34
dalmatinerdb/dqeA distributed, in-memory query engine built on top of Erlang, designed to handle high-performance data processing and analytics tasks10