Reader small image

You're reading from  Modern Data Architecture on AWS

Product typeBook
Published inAug 2023
PublisherPackt
ISBN-139781801813396
Edition1st Edition
Concepts
Right arrow
Author (1)
Behram Irani
Behram Irani
author image
Behram Irani

Behram Irani is currently a technology leader with Amazon Web Services (AWS) specializing in data, analytics and AI/ML. He has spent over 18 years in the tech industry helping organizations, from start-ups to large-scale enterprises, modernize their data platforms. In the last 6 years working at AWS, Behram has been a thought leader in the data, analytics and AI/ML space; publishing multiple papers and leading the digital transformation efforts for many organizations across the globe. Behram has completed his Bachelor of Engineering in Computer Science from the University of Pune and has an MBA degree from the University of Florida.
Read more about Behram Irani

Right arrow

Analytics using Presto, Trino, and Hive on Amazon EMR

If you recall from Chapter 5, we introduced Amazon EMR as one of the services for processing big data. EMR has over 25 open source projects, and we went through a use case where Apache Spark in EMR was leveraged to solve a data processing problem. EMR also has a few projects that assist in ad hoc query execution and allow users to interactively executive SQL queries to get the data stored in the S3 data lake. Let’s shed some light on these projects.

Presto/Trino

Presto is an open source project that provides a fast analytics query execution engine for data stored in many types of storage, most commonly used with data stored in data lakes. Presto, also known as PrestoDB, was first created on Facebook. In 2019, Presto development eventually forked into two, with PrestoDB and PrestoSQL. To keep the name confusion at a minimum, PrestoSQL was renamed Trino in 2020.

Amazon EMR supports both PrestoDB and Trino. Organizations...

lock icon
The rest of the page is locked
Previous PageNext Page
You have been reading a chapter from
Modern Data Architecture on AWS
Published in: Aug 2023Publisher: PacktISBN-13: 9781801813396

Author (1)

author image
Behram Irani

Behram Irani is currently a technology leader with Amazon Web Services (AWS) specializing in data, analytics and AI/ML. He has spent over 18 years in the tech industry helping organizations, from start-ups to large-scale enterprises, modernize their data platforms. In the last 6 years working at AWS, Behram has been a thought leader in the data, analytics and AI/ML space; publishing multiple papers and leading the digital transformation efforts for many organizations across the globe. Behram has completed his Bachelor of Engineering in Computer Science from the University of Pune and has an MBA degree from the University of Florida.
Read more about Behram Irani