PrestoDB in HPE Ezmeral Unified Analytics – Milind Bhandarkar, HPE
HPE Ezmeral Unified Analytics is an end-to-end data & AI/ML platform that consists of several popular open-source frameworks for data engineering, data analytics, data science, & ML engineering in a well-integrated packaging. These open-source frameworks include Apache Spark, Apache Airflow, Apache Superset, PrestoDB, MLFlow, Kubeflow, and Feast. This platform is built atop Kubernetes and provides built in security. In this talk we will focus on the role of PrestoDB in Unified Analytics as a fast SQL query engine, and also as a secure data access layer. We will discuss some of our value-additions to PrestoDB, such as a distributed memory-centric columnar caching layer that provides both explicit and transparent caching for dataset fragments, often leading to 3x to 4x query performance. We will conclude by proposing to make caching pluggable in PrestoDB and discussing future directions.
How to Get the Most from Presto by Connecting to OpenMetadata – Suresh Srinivas & Sriharsha Chinatalapani, Collate
PrestoDB is a fast and efficient software that can change the way teams access their data. OpenMetadata is a metadata management software that is here to change the data pipeline space. Together they can take your data, metadata and the entire company data culture to the next level. In this talk, Suresh and Sriharsha will show you how to connect your Presto database, ingest its metadata and get your data right. Use your company’s data to its full potential with OpenMetadata’s unified solution with data cataloging, discoverability, quality, lineage, operations, and collaboration all in one central UI. OpenMetadata’s philosophies: 1. Building Metadata Standards with a Schema and API first approach 2. Automation with bots abstraction for automating mundane tasks – quality, data deletion, tag propagation. 3. Top-Notch UX to enable everybody in the company to unlock the value of data.
Running PrestoDB on Kubernetes with Ahana Cloud and AWS EKS
PrestoDB is built to be cloud agnostic and container-friendly, but getting it to run on Kubernetes in the cloud can be challenging. In this talk, Gary Stafford (AWS) and Dipti Borkar (Ahana) will discuss: Why use the in-VPC deployment model with AWS and demo, etc – Deploying PrestoDB on AWS EKS using the Ahana Cloud managed service within the user’s AWS account.
Building a Modern Data Platform with Presto – Denis Krivenko, Platform24
Hadoop era is gone. Cloud computing is today’s reality. But… What if you cannot use public clouds? What if your cloud does not provide data platform capabilities? What if you want your solution to be cloud agnostic? In this case you create your own cloud native data platform on Kubernetes. In the session Denis will talk about reasons for building analytics data platform solution in Platform24, cloud native data platform architecture principles, data stack they use and why Presto plays one of the key roles in it.