Fuzzy Data Matching with SQL: Enhancing Data Quality and Query Performance

Author:   Jim Lehmer
Publisher:   O'Reilly Media
ISBN:  

9781098152277


Pages:   250
Publication Date:   13 October 2023
Format:   Paperback
Availability:   In Print   Availability explained
This item will be ordered in for you from one of our suppliers. Upon receipt, we will promptly dispatch it out to you. For in store availability, please contact us.

Our Price $158.37 Quantity:  
Add to Cart

Share |

Fuzzy Data Matching with SQL: Enhancing Data Quality and Query Performance


Add your own review!

Overview

If you were handed two different but related sets of data, what tools would you use to find the matches? What if all you had was SQL SELECT access to a database? In this practical book, author Jim Lehmer provides best practices, techniques, and tricks to help you import, clean, match, score, and think about heterogeneous data using SQL. DBAs, programmers, business analysts, and data scientists will learn how to identify and remove duplicates, parse strings, extract data from XML and JSON, generate SQL using SQL, regularize data and prepare datasets, and apply data quality and ETL approaches for finding the similarities and differences between various expressions of the same data. Full of real-world techniques, the examples in the book contain working code. You'll learn how to: Identity and remove duplicates in two different datasets using SQL Regularize data and achieve data quality using SQL Extract data from XML and JSON Generate SQL using SQL to increase your productivity Prepare datasets for import, merging, and better analysis using SQL Report results using SQL Apply data quality and ETL approaches to finding similarities and differences between various expressions of the same data ""

Full Product Details

Author:   Jim Lehmer
Publisher:   O'Reilly Media
Imprint:   O'Reilly Media
ISBN:  

9781098152277


ISBN 10:   1098152271
Pages:   250
Publication Date:   13 October 2023
Audience:   General/trade
Format:   Paperback
Publisher's Status:   Active
Availability:   In Print   Availability explained
This item will be ordered in for you from one of our suppliers. Upon receipt, we will promptly dispatch it out to you. For in store availability, please contact us.

Table of Contents

Reviews

Author Information

James Lehmer has been ""in computers"" for over three decades in various software development roles - programmer, systems programmer, software engineer, team lead, and software architect. He has worked on a variety of operating systems with a number of programming languages. James currently works in a Windows shop coding primarily in C#, but with his background in cross-platform development, he often gets tapped to deal with any *IX boxes that enter his environment.

Tab Content 6

Author Website:  

Customer Reviews

Recent Reviews

No review item found!

Add your own review!

Countries Available

All regions
Latest Reading Guide

MRG2025CC

 

Shopping Cart
Your cart is empty
Shopping cart
Mailing List