Trino is a highly parallel and distributed query engine, that is built from the ground up for efficient, low latency analytics. It is an ANSI SQL compliant query engine, that works with BI tools such as R, Tableau, Power BI, Superset and many others. It helps to natively query data in Hadoop, S3, Cassandra, MySQL, and many others, without the need for complex, slow, and error-prone processes for copying the data. 
 
Access data from multiple systems within a single query. For example, join historic log data stored in an S3 object storage with customer data stored in a MySQL relational database.

Trino is a highly parallel and distributed query engine, that is built from the ground up for efficient, low latency analytics. It is an ANSI SQL compliant query engine, that works with BI tools such as R, Tableau, Power BI, Superset and many others. It helps to natively query data in Hadoop, S3, Cassandra, MySQL, and many others, without the need for complex, slow, and error-prone processes for copying the data. 

Trino - A query engine that runs at ludicrous speed

Apache Hudi (pronounced Hoodie) stands for Hadoop Upserts Deletes and Incrementals. Hudi manages the storage of large analytical datasets on DFS (Cloud stores, HDFS or any Hadoop FileSystem compatible storage). As an organization, Hudi can help you build an efficient data lake, solving some of the most complex, low-level storage management problems, while putting data into hands of your data analysts, engineers and scientists much quicker.
 
Features: 
<ul>
<li>Upsert support with fast, pluggable indexing</li>
<li>Atomically publish data with rollback support</li>
<li>Snapshot isolation between writer &amp; queries</li>
<li>Savepoints for data recovery</li>
<li>Manages file sizes, layout using statistics</li>
<li>Async compaction of row &amp; columnar data</li>
<li>Timeline metadata to track lineage</li>
<li>Optimize data lake layout with clustering</li>
</ul>

Apache Hudi (pronounced Hoodie) stands for Hadoop Upserts Deletes and Incrementals. Hudi manages the storage of large analytical datasets on DFS (Cloud stores, HDFS or any Hadoop FileSystem compatible storage). As an organization, Hudi can help you build an efficient data lake, solving some of the most complex, low-level storage management problems, while putting data into hands of your data analysts, engineers and scientists much quicker.

Apache Hudi - Streaming Data Lake Platform

Discover open source projects across all platforms

Projects

Trino - A query engine that runs at ludicrous speed

Apache Hudi - Streaming Data Lake Platform

TechStack

Tagcloud

License

Suggested keywords:

Projects

Trino - A query engine that runs at ludicrous speed

Apache Hudi - Streaming Data Lake Platform

TechStack

Tagcloud

License