Record Deduplication with Python
Offered By: PyCon US via YouTube
Course Description
Overview
Discover the power of Record Deduplication techniques in this informative PyCon US talk. Learn how to identify duplicate records in datasets lacking unique identifiers using Python, without requiring advanced Data Science expertise. Explore real-world applications in government and business, including the Australian Census case study that led to significant population estimate revisions. Gain insights into the main concepts of Record Deduplication, common workflows, algorithms, and essential Python tools and libraries. Suitable for intermediate-level Python developers, this 29-minute presentation equips you with practical knowledge to clean and compare attributes in a fuzzy manner, enabling effective data deduplication for various critical applications.
Syllabus
Talk: Flávio Juvenal da Silva Junior - 1 + 1 = 1 or Record Deduplication with Python
Taught by
PyCon US
Related Courses
Intro to Python for Brand New ProgrammersPyCon US via YouTube Comprehending Comprehensions
PyCon US via YouTube Data Analysis with SQLite and Python
PyCon US via YouTube Build a Production Ready GraphQL API Using Python
PyCon US via YouTube Web Development With A Python-backed Frontend - Featuring HTMX and Tailwind
PyCon US via YouTube