Deconstructing Text Embedding Models - Understanding Tokenizers and Model Selection
Offered By: EuroPython Conference via YouTube
Course Description
Overview
Explore the intricacies of text embedding models in this 44-minute EuroPython Conference talk. Delve into the critical role of tokenizers in model selection, moving beyond reliance on benchmarks like the Massive Text Embedding Benchmark (MTEB). Learn to assess model suitability for specific datasets based on tokenizer performance, and discover strategies for optimizing tokenizers during the fine-tuning process of embedding models. Gain insights into making informed decisions when choosing text embedding models for unique data characteristics.
Syllabus
Deconstructing the text embedding models — Kacper Łukawski
Taught by
EuroPython Conference
Related Courses
TensorFlow: Working with NLPLinkedIn Learning Introduction to Video Editing - Video Editing Tutorials
Great Learning via YouTube HuggingFace Crash Course - Sentiment Analysis, Model Hub, Fine Tuning
Python Engineer via YouTube GPT3 and Finetuning the Core Objective Functions - A Deep Dive
David Shapiro ~ AI via YouTube How to Build a Q&A AI in Python - Open-Domain Question-Answering
James Briggs via YouTube