hive

Data warehouse tool

A software project that enables data warehousing and management of large datasets using SQL

Apache Hive

GitHub

6k stars
326 watching
5k forks
Language: Java
last commit: almost 2 years ago
Linked from 2 awesome lists

apachebig-datadatabasehadoophivejavasql

Backlinks from these awesome lists:

Related projects:

RepositoryDescriptionStars
apache/hudiA platform for storing and managing big data in cloud storage, enabling incremental processing and optimized querying of large datasets5,498
apache/shardingsphereA distributed SQL query and transaction engine for sharding, scaling, encryption, and more on any database20,034
apache/kyuubiAn Apache project providing a distributed and multi-tenant gateway to enable serverless SQL on data warehouses and lakehouses2,116
apache/kylinAn OLAP engine designed to handle Big Data with sub-second query latency and seamless integration with BI tools.3,661
apache/cassandraA highly scalable, partitioned row store that allows flexible data distribution and organization.8,906
apache/hbaseA distributed, versioned, column-oriented store designed to scale and manage large amounts of structured data5,246
apache/datafusionA query engine that supports various data formats and allows customization of its functionality.6,462
dbeaver/dbeaverA multi-platform tool for connecting to and managing various databases40,942
crate/crateA distributed and scalable SQL database for storing and analyzing massive amounts of data in near real-time.4,139
apache/igniteA distributed, in-memory database system for high-performance computing and data processing4,834
apache/drillA distributed query layer for Hadoop and NoSQL data storage systems, supporting various query languages.1,949
apache/arrowA toolkit for efficient data interchange and in-memory analytics in various languages14,728
apache/datafusion-ballistaDistributed query engine for Apache DataFusion applications1,580
apache/iotdbA time-series data management system for industrial IoT applications5,651
apache/tezA system that enables flexible data processing pipelines using a low-level engine for higher-level frameworks482