About this project
Presto is an open-source distributed SQL query engine designed for big data analytics. It enables fast, interactive queries over various data sources including Hive, relational databases, and object stores, using a familiar SQL interface. The project supports both Java-based execution and a C++ native implementation leveraging Velox for enhanced performance. This repository provides the official source code, build instructions, and development guidelines. Key features include:
- **Distributed Architecture**: Presto separates coordinator and worker nodes to scale query execution across clusters.
- **Multi-Source Connectivity**: Connectors allow querying data from Hive, MySQL, PostgreSQL, Kafka, and more.
- **High Performance**: Optimized for low-latency, interactive queries on large datasets.
- **Native Execution**: The presto-native-execution module offers a C++ rewrite using Velox for improved efficiency.
- **Extensible**: Users can write custom connectors and functions.
For deployment, refer to the official installation documentation. The project requires Java 17, Maven 3.6.3+, and supports Linux and macOS. Building from source involves running `./mvnw clean install`. The repository includes comprehensive unit tests and a web-based console built with React. Developers can contribute by following the contribution guidelines and joining the community Slack channel. Presto is licensed under Apache License 2.0.
Comments
0 Rating appears after 10 ratings
Sign in to join the discussion.