YoVDO

PandasUDFs - Scaling Ensembles for Improved Predictions

Offered By: Databricks via YouTube

Tags

Big Data Courses Apache Spark Courses Databricks Courses Data Processing Courses Predictive Modeling Courses Distributed Computing Courses Ensemble Models Courses

Course Description

Overview

Discover how to leverage PandasUDFs as a powerful technique for scaling ensemble models in this 38-minute Databricks talk. Learn to transform development code into scalable solutions for category-specific predictions, dramatically reducing runtime from hours to minutes. Explore the general usage, types of PandasUDFs, strategies for overcoming data limits, and equivalent approaches in R and Koalas. Gain insights into applying this method to scale from single models to entire ensembles, enhancing prediction accuracy across diverse categories.

Syllabus

Introduction
The Problem
PandasUDFs
Use Cases
Data Limits
Other Frameworks


Taught by

Databricks

Related Courses

Coding the Matrix: Linear Algebra through Computer Science Applications
Brown University via Coursera
كيف تفكر الآلات - مقدمة في تقنيات الحوسبة
King Fahd University of Petroleum and Minerals via Rwaq (رواق)
Datascience et Analyse situationnelle : dans les coulisses du Big Data
IONIS via IONIS
Data Lakes for Big Data
EdCast
統計学Ⅰ:データ分析の基礎 (ga014)
University of Tokyo via gacco