6,598 open-source and SaaS tools, with GitHub stats refreshed every day.

Trino

Open source

A distributed SQL query engine for big data analytics that runs fast queries across many data sources, formerly known as PrestoSQL.

trino.io
Trino homepage screenshot
GitHub stars
13k
Last commit
today
Repository age
7 years
Version
483
Licence
Apache-2.0
Self-hosted
Yes

About Trino

Trino is a distributed SQL query engine built for analytics on large data sets. It was formerly called PrestoSQL, and it lets analysts and engineers run interactive SQL against data wherever it lives rather than first moving it into a single warehouse.

The project's topics point to its role in the data lake ecosystem, with connectors and integrations for systems such as Hive, Hadoop, Iceberg and Delta Lake, and a JDBC driver for applications. Trino is a Maven project written in Java, and its repository covers development guidelines, plugin implementors, a security policy and reproducible builds. It runs as a cluster of servers and has a web UI.

Trino is released under the Apache-2.0 licence and is deployed and operated by the user, with deployment instructions and end-user documentation in the project's user manual. It is a fit for data platform teams that need fast, standards-based SQL over many sources, rather than for individual analysts looking for a point-and-click tool.

Key features

  • Distributed SQL query engine
  • Fast analytics on big data
  • Connectors for data lake formats like Iceberg
  • JDBC access for applications
  • Plugin architecture for custom connectors
  • Reproducible builds since version 449

Good fit for

  • →Interactive SQL over a data lake
  • →Querying multiple data sources from one engine
  • →Big data analytics for data platform teams
Built with
Java
Tags
sql
query-engine
big-data
analytics
data-lake
distributed
java
presto
iceberg

Open-source alternatives to Trino

See all

SaaS alternatives to Trino

See all