SQL for Efficient Data Organization in Machine Learning
Offered By: Snorkel AI via YouTube
Course Description
Overview
Save Big on Coursera Plus. 7,000+ courses at $160 off. Limited Time Only!
Explore how SQL can enhance data organization for machine learning in this 11-minute video presentation by Columbia PhD student Zachary Huang. Learn about JoinBoost, a lightweight Python library that transforms tree training algorithms over normalized databases into pure SQL queries. Discover how this innovative approach addresses the mismatch between ML data organization requirements and traditional database structures, offering a simplified, all-in-one data stack solution. Gain insights into JoinBoost's compatibility with various DBMS and data stacks, its exceptional performance and scalability, and how it outperforms specialized ML libraries like LightGBM in terms of speed and scalability for random forests and gradient boosting algorithms.
Syllabus
Introduction
Background
Example
Problem Statement
Taught by
Snorkel AI
Related Courses
Machine Learning with Python: Zero to GBMsJovian Machine Learning in Healthcare: Fundamentals & Applications
Northeastern University via Coursera Machine/Deep Learning for Mining Quality Prediction-Enhanced
Coursera Project Network via Coursera Mining Quality Prediction Using Machine & Deep Learning
Coursera Project Network via Coursera Modeling Time Series and Sequential Data
SAS via Coursera