Running Remote Shuffle Service to Solve Apache Spark's Dynamic Resource Allocation Challenge on Kubernetes
Offered By: CNCF [Cloud Native Computing Foundation] via YouTube
Course Description
Overview
Explore a novel solution to Apache Spark's dynamic resource allocation (DRA) challenge on Kubernetes using an open-source remote shuffle service (RSS). Gain insights into Spark's DRA in Kubernetes, learn how the RSS alleviates resource contention issues, and discover a more reliable and scalable solution for big data processing. Understand how offloading shuffle data to remote storage outside Spark's executor pods can decouple storage and compute, supporting dynamic scaling needs. Delve into the implementation details, benefits, and potential impact of this approach for efficient large-scale data processing in machine learning and ETL use cases.
Syllabus
Running Remote Shuffle Service to Solve a Well-Known Challenge for Apa... Melody Yang & Keyong Zhou
Taught by
CNCF [Cloud Native Computing Foundation]
Related Courses
Financial Sustainability: The Numbers side of Social Enterprise+Acumen via NovoEd Cloud Computing Concepts: Part 2
University of Illinois at Urbana-Champaign via Coursera Developing Repeatable ModelsĀ® to Scale Your Impact
+Acumen via Independent Managing Microsoft Windows Server Active Directory Domain Services
Microsoft via edX Introduction aux conteneurs
Microsoft Virtual Academy via OpenClassrooms