Hugging Face Datasets - Dataset Builder Scripts for Beginners
Offered By: James Briggs via YouTube
Course Description
Overview
Learn how to work with dataset builder scripts in Hugging Face Datasets using Python. Explore the download manager, Apache Arrow datatypes, and techniques for creating compressed files. Discover how to generate examples, finish split generators, and add datasets to Hugging Face. Gain insights into using datasets for similarity search, semantic search, vector similarity search, classification, and question-answering tasks. Apply these skills to streamline the training and fine-tuning of models with PyTorch and TensorFlow.
Syllabus
Intro
Creating Compressed Files
Creating Dataset Build Script
Download Manager
Finishing Split Generator
Generate Examples Method
Add Dataset to Hugging Face
Apache Arrow Features
What's Next?
Taught by
James Briggs
Related Courses
U&P AI - Natural Language Processing (NLP) with PythonUdemy What's New in Cognitive Search and Cool Frameworks with PyTorch - Episode 5
Microsoft via YouTube Stress Testing Qdrant - Semantic Search with 90,000 Vectors - Lightning Fast Search Microservice
David Shapiro ~ AI via YouTube Semantic Search for AI - Testing Out Qdrant Neural Search
David Shapiro ~ AI via YouTube Spotify's Podcast Search Explained
James Briggs via YouTube