YoVDO

Retrieval Augmentation and Semantic Search at Scale

Offered By: Linux Foundation via YouTube

Tags

Vector Search Courses Scalability Courses Semantic Search Courses Retrieval Augmented Generation Courses

Course Description

Overview

Save Big on Coursera Plus. 7,000+ courses at $160 off. Limited Time Only!
Explore the intricacies of Vector Search and Retrieval-Augmented Generation in this 32-minute talk by Ash Vardanian from Unum Cloud. Delve into technical benchmarks and uncover the challenges that hinder various solutions from scaling effectively. Gain insights from multiple CLIP-like AI pre-training experiments and learn about serving over 10 Billion vectors from a single machine. Examine the design decisions behind the USearch and UForm open-source libraries, and discover answers to crucial questions such as selecting the optimal GPU for inference workloads and serving search results from SSDs instead of RAM. Understand the tools and techniques that enable efficient retrieval augmentation and semantic search at scale, addressing the high costs associated with AI work.

Syllabus

Retrieval Augmentation and Semantic Search at Scale - Ash Vardanian, Unum Cloud


Taught by

Linux Foundation

Tags

Related Courses

AWS Flash - Operationalize Generative AI Applications (FMOps/LLMOps)
Amazon Web Services via AWS Skill Builder
AWS Flash - Operationalize Generative AI Applications (FMOps/LLMOps) (Simplified Chinese)
Amazon Web Services via AWS Skill Builder
Building Retrieval Augmented Generation (RAG) workflows with Amazon OpenSearch Service
Amazon Web Services via AWS Skill Builder
Advanced Prompt Engineering for Everyone
Vanderbilt University via Coursera
Advanced Retrieval for AI with Chroma
DeepLearning.AI via Coursera