DBMS > Apache Druid vs. Hive vs. Spark SQL
System Properties Comparison Apache Druid vs. Hive vs. Spark SQL
Please select another system to include it in the comparison.
|Editorial information provided by DB-Engines|
|Name||Apache Druid Xexclude from comparison||Hive Xexclude from comparison||Spark SQL Xexclude from comparison|
|Description||Open-source analytics data store designed for sub-second OLAP queries on high dimensionality and high cardinality data||data warehouse software for querying and managing large distributed datasets, built on Hadoop||Spark SQL is a component on top of 'Spark Core' for structured data processing|
|Primary database model||Relational DBMS|
Time Series DBMS
|Relational DBMS||Relational DBMS|
|Developer||Apache Software Foundation and contributors||Apache Software Foundation initially developed by Facebook||Apache Software Foundation|
|Current release||25.0.0, January 2023||3.1.3, April 2022||3.3.0 ( 2.13), June 2022|
|License Commercial or Open Source||Open Source Apache license v2||Open Source Apache Version 2||Open Source Apache 2.0|
|Cloud-based only Only available as a cloud service||no||no||no|
|DBaaS offerings (sponsored links) Database as a Service|
Providers of DBaaS offerings, please contact us to be listed.
|Server operating systems||Linux|
|All OS with a Java VM||Linux|
|Data scheme||yes schema-less columns are supported||yes||yes|
|Typing predefined data types such as float or date||yes||yes||yes|
|XML support Some form of processing data in XML format, e.g. support for XML data structures, and/or support for XPath, XQuery or XSLT.||no||no|
|SQL Support of SQL||SQL for querying||SQL-like DML and DDL statements||SQL-like DML and DDL statements|
|APIs and other access methods||JDBC|
RESTful HTTP/JSON API
|Supported programming languages||Clojure|
|Server-side scripts Stored procedures||no||yes user defined functions and integration of map-reduce||no|
|Partitioning methods Methods for storing different data on different nodes||Sharding manual/auto, time-based||Sharding||yes, utilizing Spark Core|
|Replication methods Methods for redundantly storing data on multiple nodes||yes, via HDFS, S3 or other storage engines||selectable replication factor||none|
|MapReduce Offers an API for user-defined Map/Reduce methods||no||yes query execution via MapReduce|
|Consistency concepts Methods to ensure consistency in a distributed system||Immediate Consistency||Eventual Consistency|
|Foreign keys Referential integrity||no||no||no|
|Transaction concepts Support to ensure data integrity after non-atomic manipulations of data||no||no||no|
|Concurrency Support for concurrent manipulation of data||yes||yes||yes|
|Durability Support for making data persistent||yes||yes||yes|
|In-memory capabilities Is there an option to define some or all structures to be held in-memory only.||no||no|
|User concepts Access control||RBAC using LDAP or Druid internals for users and groups for read/write by datasource and system||Access rights for users, groups and roles||no|
More information provided by the system vendor
We invite representatives of system vendors to contact us for updating and extending the system information,
Related products and services
We invite representatives of vendors of related products to contact us for presenting information about their offerings here.
|Apache Druid||Hive||Spark SQL|
|DB-Engines blog posts|
Why is Hadoop not listed in the DB-Engines Ranking?
|Recent citations in the news|
Apache Druid 25.0 Delivers Multi-Stage Query Engine and ...
Stream Big, Think Bigger: Analyze Streaming Data at Scale
As CIO budgets tighten, Apache Superset looks like a Big Data winner
Imply Announces Major Open Source Contribution for Apache Druid ...
Apache Druid’s Role in Modern Data Analytics
provided by Google News
Apache Iceberg promises to change the economics of cloud-based data analytics
Are Databases Becoming Just Query Engines for Big Object Stores?
The Apache Software Foundation Announces Apache® Linkis™ as ...
What Is Apache Hive? (Definition, Benefits, Challenges)
Ground Labs Introduces Enterprise Recon 2.8 with New Way of ...
provided by Google News
Apache® Kyuubi Becomes Top-Level Project
A Deep Dive into Custom Spark Transformers for ML Pipelines
Data Engineer (Apache Airflow, Hive, Spark, SQL, AWS)
Accelerating SQL Queries on a Modern Real-Time Database
Data chess game: Databricks vs. Snowflake, part 1
provided by Google News
Kubernetes Admin-Apache Druid cluster
DevOps Engineer Fully Remote
Jr. Big Data Engineer - 114065
Machine Learning Engineer
Sr. Big Data Engineer
Senior Data Engineer
Entry Level Data Engineer with BigData - 1
Scala/Spark Data Engineer
Big Data Support Engineer (Spark, SQL)
Senior Member of Technical Staff
Spark (Databricks) Engineer – All Levels
Share this page