PrestoDB (or Presto)
19 Reviews and Ratings
What is PrestoDB (or Presto)?
PrestoDB is an open-source, distributed SQL query engine for interactive analytics and open data-lakehouse workloads. It lets data teams query data where it resides—across data lakes, object storage, relational databases, NoSQL systems, warehouses, and streaming platforms—through a single ANSI SQL interface.
PrestoDB uses a connector architecture to access external systems such as Hive, Iceberg, Delta Lake, Hudi, Kafka, PostgreSQL, MySQL, Oracle Database, MongoDB, Elasticsearch, BigQuery, Redshift, and cloud storage-backed lakehouse tables. A single query can join data from multiple connected sources without requiring the data to be copied into a separate database first.
The engine separates query compute from storage and distributes query work across a cluster. Its coordinator plans and schedules queries, while worker nodes process data in parallel. Query optimization features include cost-based optimization, pushdown to supported source systems, caching, materialized views, resource groups, and exchange materialization. PrestoDB can be deployed on premises or in any cloud environment, including through Docker and Kubernetes.
PrestoDB supports ad hoc analysis, reporting, dashboards, and large-scale interactive SQL workloads. JDBC, REST, and client integrations allow it to be used from SQL clients, BI tools, notebooks, and custom applications. Presto on Spark is available for teams that want to use Presto’s SQL engine and optimizer with Spark for large-scale batch processing.
Categories & Use Cases
Videos
Product Demos
Technical Details
| Mobile Application | No |
|---|
FAQs
What is PrestoDB (or Presto)?
Presto is an open source SQL query engine designed to run queries on data stored in Hadoop or in traditional databases.
Teradata supported development of Presto followed the acquisition of Hadapt and Revelytix.