Retrieval Augmentation and Semantic Search at Scale
Offered By: Linux Foundation via YouTube
Course Description
Overview
Explore the intricacies of Vector Search and Retrieval-Augmented Generation in this 32-minute talk by Ash Vardanian from Unum Cloud. Delve into technical benchmarks and uncover the challenges that hinder various solutions from scaling effectively. Gain insights from multiple CLIP-like AI pre-training experiments and learn about serving over 10 Billion vectors from a single machine. Examine the design decisions behind the USearch and UForm open-source libraries, and discover answers to crucial questions such as selecting the optimal GPU for inference workloads and serving search results from SSDs instead of RAM. Understand the tools and techniques that enable efficient retrieval augmentation and semantic search at scale, addressing the high costs associated with AI work.
Syllabus
Retrieval Augmentation and Semantic Search at Scale - Ash Vardanian, Unum Cloud
Taught by
Linux Foundation
Tags
Related Courses
Pinecone Vercel Starter Template and RAG - Live Code Review Part 2Pinecone via YouTube Will LLMs Kill Search? The Future of Information Retrieval
Aleksa Gordić - The AI Epiphany via YouTube RAG But Better: Rerankers with Cohere AI - Improving Retrieval Pipelines
James Briggs via YouTube Advanced RAG - Contextual Compressors and Filters - Lecture 4
Sam Witteveen via YouTube LangChain Multi-Query Retriever for RAG - Advanced Technique for Broader Vector Space Search
James Briggs via YouTube