How to Boost a Cloud-Native Lakehouse with 2X Performance
Offered By: The ASF via YouTube
Course Description
Overview
Discover how to enhance a cloud-native lakehouse's performance by 200% in this 48-minute conference talk. Explore the challenges of using Apache Spark as a lake analysis engine on Kubernetes, including resource management, task scheduling, storage interconnection, elastic scalability, and high reliability. Learn from Shi Shaofeng, Chief Architect at Kyligence and Apache Kylin committer, as he shares insights on building a cloud-native lake analytics engine using open-source technologies such as Kubernetes, Spark, Gluten, Volcano, and Kyuubi. Gain valuable knowledge from Kyligence's experience serving various clients and their contributions to the open-source community in the realm of cloud computing and big data technology integration.
Syllabus
How To Boost A Cloud-Native Lakehouse With 2X Performance
Taught by
The ASF
Related Courses
CS115x: Advanced Apache Spark for Data Science and Data EngineeringUniversity of California, Berkeley via edX Big Data Analytics
University of Adelaide via edX Big Data Essentials: HDFS, MapReduce and Spark RDD
Yandex via Coursera Big Data Analysis: Hive, Spark SQL, DataFrames and GraphFrames
Yandex via Coursera Introduction to Apache Spark and AWS
University of London International Programmes via Coursera