Isbn: 9783031007071 - an introduction to duplicate detection (24 risultati)

Perfeziona la tua ricerca

  • Libri (24)

a

Fascia di prezzo personalizzata (EUR)

a

  • Lingua: Inglese

    Editore: Springer, 2010

    3031007077 / 9783031007071

    • Brossura

    Da: GreatBookPrices, Columbia, MD, U.S.A.GreatBookPrices

    Venditore con 5 stelle
    Contatta il venditore

    Condizione: Usato - Come nuovo

    EUR 27,36

    EUR 2,30 spedizione 
    Spedito in U.S.A.

    Quantità: Più di 20 disponibili

    Condizione: As New. Unread book in perfect condition.

  • Lingua: Inglese

    Editore: Springer, 2010

    3031007077 / 9783031007071

    • Brossura

    Da: GreatBookPrices, Columbia, MD, U.S.A.GreatBookPrices

    Venditore con 5 stelle
    Contatta il venditore

    Condizione: Nuovo

    EUR 28,80

    EUR 2,30 spedizione 
    Spedito in U.S.A.

    Quantità: Più di 20 disponibili

    Condizione: New.

  • Lingua: Inglese

    Editore: Springer International Publishing AG, Cham, 2010

    3031007077 / 9783031007071

    • Brossura

    Da: Grand Eagle Retail, Bensenville, IL, U.S.A.Grand Eagle Retail

    Venditore con 5 stelle
    Contatta il venditore

    Condizione: Nuovo

    EUR 31,18

     Spedizione gratuita 
    Spedito in U.S.A.

    Quantità: 1 disponibili

    Paperback. Condizione: new. Paperback. With the ever increasing volume of data, data quality problems abound. Multiple, yet different representations of the same real-world objects in data, duplicates, are one of the most intriguing data quality problems. The effects of such duplicates are detrimental; for instance, bank customers can obtain duplicate identities, inventory levels are monitored incorrectly, catalogs are mailed multiple times to the same household, etc. Automatically detecting duplicates is difficult: First, duplicate representations are usually not identical but slightly differ in their values. Second, in principle all pairs of records should be compared, which is infeasible for large volumes of data. This lecture examines closely the two main components to overcome these difficulties: (i) Similarity measures are used to automatically identify duplicates when comparing two records. Well-chosen similarity measures improve the effectiveness of duplicate detection. (ii) Algorithms are developed to perform on very large volumes of data in search for duplicates. Well-designed algorithms improve the efficiency of duplicate detection. Finally, we discuss methods to evaluate the success of duplicate detection. Table of Contents: Data Cleansing: Introduction and Motivation / Problem Definition / Similarity Functions / Duplicate Detection Algorithms / Evaluating Detection Success / Conclusion and Outlook / Bibliography With the ever increasing volume of data, data quality problems abound. Table of Contents: Data Cleansing: Introduction and Motivation / Problem Definition / Similarity Functions / Duplicate Detection Algorithms / Evaluating Detection Success / Conclusion and Outlook / Bibliography Shipping may be from multiple locations in the US or from the UK, depending on stock availability.

  • Lingua: Inglese

    Editore: Springer International Publishing AG, CH, 2010

    3031007077 / 9783031007071

    • Brossura

    Da: Rarewaves.com USA, London, LONDO, Regno UnitoRarewaves.com USA

    Venditore con 5 stelle
    Contatta il venditore

    Condizione: Nuovo

    EUR 31,63

     Spedizione gratuita 
    Spedito da Regno Unito a U.S.A.

    Quantità: 1 disponibili

    Paperback. Condizione: New. With the ever increasing volume of data, data quality problems abound. Multiple, yet different representations of the same real-world objects in data, duplicates, are one of the most intriguing data quality problems. The effects of such duplicates are detrimental; for instance, bank customers can obtain duplicate identities, inventory levels are monitored incorrectly, catalogs are mailed multiple times to the same household, etc. Automatically detecting duplicates is difficult: First, duplicate representations are usually not identical but slightly differ in their values. Second, in principle all pairs of records should be compared, which is infeasible for large volumes of data. This lecture examines closely the two main components to overcome these difficulties: (i) Similarity measures are used to automatically identify duplicates when comparing two records. Well-chosen similarity measures improve the effectiveness of duplicate detection. (ii) Algorithms are developed to perform on very large volumes of data in search for duplicates. Well-designed algorithms improve the efficiency of duplicate detection. Finally, we discuss methods to evaluate the success of duplicate detection. Table of Contents: Data Cleansing: Introduction and Motivation / Problem Definition / Similarity Functions / Duplicate Detection Algorithms / Evaluating Detection Success / Conclusion and Outlook / Bibliography.

  • Lingua: Inglese

    Editore: Springer, 2010

    3031007077 / 9783031007071

    • Brossura

    Da: PBShop.store UK, Fairford, GLOS, Regno UnitoPBShop.store UK

    Venditore con 5 stelle
    Contatta il venditore

    Condizione: Nuovo

    EUR 27,73

    EUR 3,83 spedizione 
    Spedito da Regno Unito a U.S.A.

    Quantità: 2 disponibili

    PAP. Condizione: New. New Book. Shipped from UK. Established seller since 2000.

  • Lingua: Inglese

    Editore: Springer International Publishing AG, CH, 2010

    3031007077 / 9783031007071

    • Brossura

    Da: Rarewaves USA, HEBRON, KY, U.S.A.Rarewaves USA

    Venditore con 5 stelle
    Contatta il venditore

    Condizione: Nuovo

    EUR 34,65

     Spedizione gratuita 
    Spedito in U.S.A.

    Quantità: 1 disponibili

    Paperback. Condizione: New. With the ever increasing volume of data, data quality problems abound. Multiple, yet different representations of the same real-world objects in data, duplicates, are one of the most intriguing data quality problems. The effects of such duplicates are detrimental; for instance, bank customers can obtain duplicate identities, inventory levels are monitored incorrectly, catalogs are mailed multiple times to the same household, etc. Automatically detecting duplicates is difficult: First, duplicate representations are usually not identical but slightly differ in their values. Second, in principle all pairs of records should be compared, which is infeasible for large volumes of data. This lecture examines closely the two main components to overcome these difficulties: (i) Similarity measures are used to automatically identify duplicates when comparing two records. Well-chosen similarity measures improve the effectiveness of duplicate detection. (ii) Algorithms are developed to perform on very large volumes of data in search for duplicates. Well-designed algorithms improve the efficiency of duplicate detection. Finally, we discuss methods to evaluate the success of duplicate detection. Table of Contents: Data Cleansing: Introduction and Motivation / Problem Definition / Similarity Functions / Duplicate Detection Algorithms / Evaluating Detection Success / Conclusion and Outlook / Bibliography.

  • Lingua: Inglese

    Editore: Springer, 2010

    3031007077 / 9783031007071

    • Brossura

    Da: Ria Christie Collections, Uxbridge, Regno UnitoRia Christie Collections

    Venditore con 5 stelle
    Contatta il venditore

    Condizione: Nuovo

    EUR 28,62

    EUR 10,91 spedizione 
    Spedito da Regno Unito a U.S.A.

    Quantità: Più di 20 disponibili

    Condizione: New. In English.

  • Lingua: Inglese

    Editore: Springer, 2010

    3031007077 / 9783031007071

    • Brossura

    Da: Books Puddle, New York, NY, U.S.A.Books Puddle

    Venditore con 4 stelle
    Contatta il venditore

    Condizione: Nuovo

    EUR 40,06

    EUR 3,48 spedizione 
    Spedito in U.S.A.

    Quantità: 4 disponibili

    Condizione: New. 1st edition NO-PA16APR2015-KAP.

  • Lingua: Inglese

    Editore: Springer, 2010

    3031007077 / 9783031007071

    • Brossura

    Da: GreatBookPricesUK, Woodford Green, Regno UnitoGreatBookPricesUK

    Venditore con 5 stelle
    Contatta il venditore

    Condizione: Nuovo

    EUR 27,72

    EUR 17,46 spedizione 
    Spedito da Regno Unito a U.S.A.

    Quantità: Più di 20 disponibili

    Condizione: New.

  • Lingua: Inglese

    Editore: Springer 2010-03, 2010

    3031007077 / 9783031007071

    • Brossura

    Da: Chiron Media, Wallingford, Regno UnitoChiron Media

    Venditore con 5 stelle
    Contatta il venditore

    Condizione: Nuovo

    EUR 29,10

    EUR 18,03 spedizione 
    Spedito da Regno Unito a U.S.A.

    Quantità: 10 disponibili

    PF. Condizione: New.

  • Lingua: Inglese

    Editore: Springer, 2010

    3031007077 / 9783031007071

    • Brossura

    Da: GreatBookPricesUK, Woodford Green, Regno UnitoGreatBookPricesUK

    Venditore con 5 stelle
    Contatta il venditore

    Condizione: Usato - Come nuovo

    EUR 30,48

    EUR 17,46 spedizione 
    Spedito da Regno Unito a U.S.A.

    Quantità: Più di 20 disponibili

    Condizione: As New. Unread book in perfect condition.

  • Lingua: Inglese

    Editore: Springer, 2010

    3031007077 / 9783031007071

    • Brossura

    Da: Speedyhen, Hertfordshire, Regno UnitoSpeedyhen

    Venditore con 5 stelle
    Contatta il venditore

    Condizione: Nuovo

    EUR 24,64

    EUR 47,72 spedizione 
    Spedito da Regno Unito a U.S.A.

    Quantità: 2 disponibili

    Condizione: NEW.

  • Lingua: Inglese

    Editore: Springer, 2010

    3031007077 / 9783031007071

    • Brossura

    Da: AHA-BUCH GmbH, Einbeck, GermaniaAHA-BUCH GmbH

    Venditore con 5 stelle
    Contatta il venditore

    Condizione: Nuovo

    EUR 46,10

    EUR 30,50 spedizione 
    Spedito da Germania a U.S.A.

    Quantità: 1 disponibili

    Taschenbuch. Condizione: Neu. Druck auf Anfrage Neuware - Printed after ordering - With the ever increasing volume of data, data quality problems abound. Multiple, yet different representations of the same real-world objects in data, duplicates, are one of the most intriguing data quality problems. The effects of such duplicates are detrimental; for instance, bank customers can obtain duplicate identities, inventory levels are monitored incorrectly, catalogs are mailed multiple times to the same household, etc. Automatically detecting duplicates is difficult: First, duplicate representations are usually not identical but slightly differ in their values. Second, in principle all pairs of records should be compared, which is infeasible for large volumes of data. This lecture examines closely the two main components to overcome these difficulties: (i) Similarity measures are used to automatically identify duplicates when comparing two records. Well-chosen similarity measures improve the effectiveness of duplicate detection. (ii) Algorithms are developed to perform on very large volumes of data in search for duplicates. Well-designed algorithms improve the efficiency of duplicate detection. Finally, we discuss methods to evaluate the success of duplicate detection. Table of Contents: Data Cleansing: Introduction and Motivation / Problem Definition / Similarity Functions / Duplicate Detection Algorithms / Evaluating Detection Success / Conclusion and Outlook / Bibliography.

  • Lingua: Inglese

    Editore: Springer International Publishing AG, CH, 2010

    3031007077 / 9783031007071

    • Brossura

    Da: Rarewaves USA United, HEBRON, KY, U.S.A.Rarewaves USA United

    Venditore con 5 stelle
    Contatta il venditore

    Condizione: Nuovo

    EUR 36,55

    EUR 43,57 spedizione 
    Spedito in U.S.A.

    Quantità: 1 disponibili

    Paperback. Condizione: New. With the ever increasing volume of data, data quality problems abound. Multiple, yet different representations of the same real-world objects in data, duplicates, are one of the most intriguing data quality problems. The effects of such duplicates are detrimental; for instance, bank customers can obtain duplicate identities, inventory levels are monitored incorrectly, catalogs are mailed multiple times to the same household, etc. Automatically detecting duplicates is difficult: First, duplicate representations are usually not identical but slightly differ in their values. Second, in principle all pairs of records should be compared, which is infeasible for large volumes of data. This lecture examines closely the two main components to overcome these difficulties: (i) Similarity measures are used to automatically identify duplicates when comparing two records. Well-chosen similarity measures improve the effectiveness of duplicate detection. (ii) Algorithms are developed to perform on very large volumes of data in search for duplicates. Well-designed algorithms improve the efficiency of duplicate detection. Finally, we discuss methods to evaluate the success of duplicate detection. Table of Contents: Data Cleansing: Introduction and Motivation / Problem Definition / Similarity Functions / Duplicate Detection Algorithms / Evaluating Detection Success / Conclusion and Outlook / Bibliography.

  • Lingua: Inglese

    Editore: Springer, Berlin|Springer International Publishing|Morgan & Claypool|Springer, 2010

    3031007077 / 9783031007071

    • Brossura

    Da: moluna, Greven, Germaniamoluna

    Venditore con 5 stelle
    Contatta il venditore

    Condizione: Nuovo

    EUR 32,08

    EUR 48,99 spedizione 
    Spedito da Germania a U.S.A.

    Quantità: 2 disponibili

    Condizione: New. With the ever increasing volume of data, data quality problems abound. Multiple, yet different representations of the same real-world objects in data, duplicates, are one of the most intriguing data quality problems. The effects of such duplicates are detri.

  • Lingua: Inglese

    Editore: Springer International Publishing AG, Cham, 2010

    3031007077 / 9783031007071

    • Brossura

    Da: AussieBookSeller, Truganina, VIC, AustraliaAussieBookSeller

    Venditore con 5 stelle
    Contatta il venditore

    Condizione: Nuovo

    EUR 52,78

    EUR 32,24 spedizione 
    Spedito da Australia a U.S.A.

    Quantità: 1 disponibili

    Paperback. Condizione: new. Paperback. With the ever increasing volume of data, data quality problems abound. Multiple, yet different representations of the same real-world objects in data, duplicates, are one of the most intriguing data quality problems. The effects of such duplicates are detrimental; for instance, bank customers can obtain duplicate identities, inventory levels are monitored incorrectly, catalogs are mailed multiple times to the same household, etc. Automatically detecting duplicates is difficult: First, duplicate representations are usually not identical but slightly differ in their values. Second, in principle all pairs of records should be compared, which is infeasible for large volumes of data. This lecture examines closely the two main components to overcome these difficulties: (i) Similarity measures are used to automatically identify duplicates when comparing two records. Well-chosen similarity measures improve the effectiveness of duplicate detection. (ii) Algorithms are developed to perform on very large volumes of data in search for duplicates. Well-designed algorithms improve the efficiency of duplicate detection. Finally, we discuss methods to evaluate the success of duplicate detection. Table of Contents: Data Cleansing: Introduction and Motivation / Problem Definition / Similarity Functions / Duplicate Detection Algorithms / Evaluating Detection Success / Conclusion and Outlook / Bibliography With the ever increasing volume of data, data quality problems abound. Table of Contents: Data Cleansing: Introduction and Motivation / Problem Definition / Similarity Functions / Duplicate Detection Algorithms / Evaluating Detection Success / Conclusion and Outlook / Bibliography Shipping may be from our Sydney, NSW warehouse or from our UK or US warehouse, depending on stock availability.

  • Altre immagini

    Lingua: Inglese

    Editore: Springer, 2010

    3031007077 / 9783031007071

    • Brossura

    Da: preigu, Osnabrück, Germaniapreigu

    Venditore con 5 stelle
    Contatta il venditore

    Condizione: Nuovo

    EUR 26,50

    EUR 70,00 spedizione 
    Spedito da Germania a U.S.A.

    Quantità: 5 disponibili

    Taschenbuch. Condizione: Neu. An Introduction to Duplicate Detection | Felix Nauman (u. a.) | Taschenbuch | Synthesis Lectures on Data Management | ix | Englisch | 2010 | Springer | EAN 9783031007071 | Verantwortliche Person für die EU: Springer Verlag GmbH, Tiergartenstr. 17, 69121 Heidelberg, juergen[dot]hartmann[at]springer[dot]com | Anbieter: preigu.

  • Lingua: Inglese

    Editore: Springer International Publishing AG, CH, 2010

    3031007077 / 9783031007071

    • Brossura

    Da: Rarewaves.com UK, London, Regno UnitoRarewaves.com UK

    Venditore con 5 stelle
    Contatta il venditore

    Condizione: Nuovo

    EUR 29,25

    EUR 75,65 spedizione 
    Spedito da Regno Unito a U.S.A.

    Quantità: 1 disponibili

    Paperback. Condizione: New. With the ever increasing volume of data, data quality problems abound. Multiple, yet different representations of the same real-world objects in data, duplicates, are one of the most intriguing data quality problems. The effects of such duplicates are detrimental; for instance, bank customers can obtain duplicate identities, inventory levels are monitored incorrectly, catalogs are mailed multiple times to the same household, etc. Automatically detecting duplicates is difficult: First, duplicate representations are usually not identical but slightly differ in their values. Second, in principle all pairs of records should be compared, which is infeasible for large volumes of data. This lecture examines closely the two main components to overcome these difficulties: (i) Similarity measures are used to automatically identify duplicates when comparing two records. Well-chosen similarity measures improve the effectiveness of duplicate detection. (ii) Algorithms are developed to perform on very large volumes of data in search for duplicates. Well-designed algorithms improve the efficiency of duplicate detection. Finally, we discuss methods to evaluate the success of duplicate detection. Table of Contents: Data Cleansing: Introduction and Motivation / Problem Definition / Similarity Functions / Duplicate Detection Algorithms / Evaluating Detection Success / Conclusion and Outlook / Bibliography.

  • Lingua: Inglese

    Editore: Springer, 2010

    3031007077 / 9783031007071

    • Brossura
    • Print on Demand

    Da: Brook Bookstore On Demand, Napoli, NA, ItaliaBrook Bookstore On Demand

    Venditore con 5 stelle
    Contatta il venditore

    Condizione: Nuovo

    EUR 26,21

    EUR 4,00 spedizione 
    Spedito da Italia a U.S.A.

    Quantità: Più di 20 disponibili

    Condizione: new. Questo è un articolo print on demand.

  • Lingua: Inglese

    Editore: Springer-Nature New York Inc, 2010

    3031007077 / 9783031007071

    • Brossura
    • Print on Demand

    Da: Revaluation Books, Exeter, Regno UnitoRevaluation Books

    Venditore con 5 stelle
    Contatta il venditore

    Condizione: Nuovo

    EUR 30,02

    EUR 11,64 spedizione 
    Spedito da Regno Unito a U.S.A.

    Quantità: 2 disponibili

    Paperback. Condizione: Brand New. 86 pages. 9.25x7.51x9.25 inches. In Stock. This item is printed on demand.

  • Lingua: Inglese

    Editore: Springer, 2010

    3031007077 / 9783031007071

    • Brossura
    • Print on Demand

    Da: Majestic Books, Hounslow, Regno UnitoMajestic Books

    Venditore con 4 stelle
    Contatta il venditore

    Condizione: Nuovo

    EUR 36,79

    EUR 7,57 spedizione 
    Spedito da Regno Unito a U.S.A.

    Quantità: 4 disponibili

    Condizione: New. Print on Demand.

  • Lingua: Inglese

    Editore: Springer, 2010

    3031007077 / 9783031007071

    • Brossura
    • Print on Demand

    Da: Biblios, frankfurt am main, HESSE, GermaniaBiblios

    Venditore con 4 stelle
    Contatta il venditore

    Condizione: Nuovo

    EUR 37,65

    EUR 9,95 spedizione 
    Spedito da Germania a U.S.A.

    Quantità: 4 disponibili

    Condizione: New. PRINT ON DEMAND.

  • Lingua: Inglese

    Editore: Springer International Publishing Mrz 2010, 2010

    3031007077 / 9783031007071

    • Brossura
    • Print on Demand

    Da: BuchWeltWeit Ludwig Meier e.K., Bergisch Gladbach, GermaniaBuchWeltWeit Ludwig Meier e.K.

    Venditore con 5 stelle
    Contatta il venditore

    Condizione: Nuovo

    EUR 26,74

    EUR 23,00 spedizione 
    Spedito da Germania a U.S.A.

    Quantità: 2 disponibili

    Taschenbuch. Condizione: Neu. This item is printed on demand - it takes 3-4 days longer - Neuware -With the ever increasing volume of data, data quality problems abound. Multiple, yet different representations of the same real-world objects in data, duplicates, are one of the most intriguing data quality problems. The effects of such duplicates are detrimental; for instance, bank customers can obtain duplicate identities, inventory levels are monitored incorrectly, catalogs are mailed multiple times to the same household, etc. Automatically detecting duplicates is difficult: First, duplicate representations are usually not identical but slightly differ in their values. Second, in principle all pairs of records should be compared, which is infeasible for large volumes of data. This lecture examines closely the two main components to overcome these difficulties: (i) Similarity measures are used to automatically identify duplicates when comparing two records. Well-chosen similarity measures improve the effectiveness of duplicate detection. (ii) Algorithms are developed to perform on very large volumes of data in search for duplicates. Well-designed algorithms improve the efficiency of duplicate detection. Finally, we discuss methods to evaluate the success of duplicate detection. Table of Contents: Data Cleansing: Introduction and Motivation / Problem Definition / Similarity Functions / Duplicate Detection Algorithms / Evaluating Detection Success / Conclusion and Outlook / Bibliography 88 pp. Englisch.

  • Lingua: Inglese

    Editore: Springer, Springer Mär 2010, 2010

    3031007077 / 9783031007071

    • Brossura
    • Print on Demand

    Da: buchversandmimpf2000, Emtmannsberg, BAYE, Germaniabuchversandmimpf2000

    Venditore con 5 stelle
    Contatta il venditore

    Condizione: Nuovo

    EUR 26,74

    EUR 60,00 spedizione 
    Spedito da Germania a U.S.A.

    Quantità: 1 disponibili

    Taschenbuch. Condizione: Neu. This item is printed on demand - Print on Demand Titel. Neuware -With the ever increasing volume of data, data quality problems abound. Multiple, yet different representations of the same real-world objects in data, duplicates, are one of the most intriguing data quality problems. The effects of such duplicates are detrimental; for instance, bank customers can obtain duplicate identities, inventory levels are monitored incorrectly, catalogs are mailed multiple times to the same household, etc. Automatically detecting duplicates is difficult: First, duplicate representations are usually not identical but slightly differ in their values. Second, in principle all pairs of records should be compared, which is infeasible for large volumes of data. This lecture examines closely the two main components to overcome these difficulties: (i) Similarity measures are used to automatically identify duplicates when comparing two records. Well-chosen similarity measures improve the effectiveness of duplicate detection. (ii) Algorithms are developed to perform on very large volumes of data in search for duplicates. Well-designed algorithms improve the efficiency of duplicate detection. Finally, we discuss methods to evaluate the success of duplicate detection. Table of Contents: Data Cleansing: Introduction and Motivation / Problem Definition / Similarity Functions / Duplicate Detection Algorithms / Evaluating Detection Success / Conclusion and Outlook / BibliographySpringer-Verlag KG, Sachsenplatz 4-6, 1201 Wien 88 pp. Englisch.