Class Central is learner-supported. When you buy through links on our site, we may earn an affiliate commission.

YouTube

Record Deduplication with Python

PyCon US via YouTube

Overview

Save Big on Coursera Plus. 7,000+ courses at $160 off. Limited Time Only!
Discover the power of Record Deduplication techniques in this informative PyCon US talk. Learn how to identify duplicate records in datasets lacking unique identifiers using Python, without requiring advanced Data Science expertise. Explore real-world applications in government and business, including the Australian Census case study that led to significant population estimate revisions. Gain insights into the main concepts of Record Deduplication, common workflows, algorithms, and essential Python tools and libraries. Suitable for intermediate-level Python developers, this 29-minute presentation equips you with practical knowledge to clean and compare attributes in a fuzzy manner, enabling effective data deduplication for various critical applications.

Syllabus

Talk: Flávio Juvenal da Silva Junior - 1 + 1 = 1 or Record Deduplication with Python

Taught by

PyCon US

Reviews

Start your review of Record Deduplication with Python

Never Stop Learning.

Get personalized course recommendations, track subjects and courses with reminders, and more.

Someone learning on their laptop while sitting on the floor.