Isbn: 9783031007071 - an introduction to duplicate detection (synthesis lectures on data management) (23 resultados)

ISBN: 
Refinar con la Búsqueda avanzada

Filtrar la búsqueda

  • Libros (23)

a

Intervalo de precios personalizado (EUR)

a

  • Idioma: Inglés

    Editorial: Springer, 2010

    3031007077 / 9783031007071

    • Tapa blanda

    Librería: GreatBookPrices, Columbia, MD, Estados Unidos de AmericaGreatBookPrices

    Vendedor de 5 estrellas
    Contactar con el vendedor

    Condición: Usado - Como Nuevo

    EUR 27,40

    Envío por EUR 2,35 
    Se envía dentro de Estados Unidos de America

    Cantidad disponible: Más de 20 disponibles

    Condición: As New. Unread book in perfect condition.

  • Idioma: Inglés

    Editorial: Springer International Publishing AG, Cham, 2010

    3031007077 / 9783031007071

    • Tapa blanda

    Librería: Grand Eagle Retail, Bensenville, IL, Estados Unidos de AmericaGrand Eagle Retail

    Vendedor de 5 estrellas
    Contactar con el vendedor

    Condición: Nuevo

    EUR 31,66

     Gastos de envío gratis 
    Se envía dentro de Estados Unidos de America

    Cantidad disponible: 1 disponible

    Paperback. Condición: new. Paperback. With the ever increasing volume of data, data quality problems abound. Multiple, yet different representations of the same real-world objects in data, duplicates, are one of the most intriguing data quality problems. The effects of such duplicates are detrimental; for instance, bank customers can obtain duplicate identities, inventory levels are monitored incorrectly, catalogs are mailed multiple times to the same household, etc. Automatically detecting duplicates is difficult: First, duplicate representations are usually not identical but slightly differ in their values. Second, in principle all pairs of records should be compared, which is infeasible for large volumes of data. This lecture examines closely the two main components to overcome these difficulties: (i) Similarity measures are used to automatically identify duplicates when comparing two records. Well-chosen similarity measures improve the effectiveness of duplicate detection. (ii) Algorithms are developed to perform on very large volumes of data in search for duplicates. Well-designed algorithms improve the efficiency of duplicate detection. Finally, we discuss methods to evaluate the success of duplicate detection. Table of Contents: Data Cleansing: Introduction and Motivation / Problem Definition / Similarity Functions / Duplicate Detection Algorithms / Evaluating Detection Success / Conclusion and Outlook / Bibliography With the ever increasing volume of data, data quality problems abound. Table of Contents: Data Cleansing: Introduction and Motivation / Problem Definition / Similarity Functions / Duplicate Detection Algorithms / Evaluating Detection Success / Conclusion and Outlook / Bibliography Shipping may be from multiple locations in the US or from the UK, depending on stock availability.…

  • Idioma: Inglés

    Editorial: Springer, 2010

    3031007077 / 9783031007071

    • Tapa blanda

    Librería: GreatBookPrices, Columbia, MD, Estados Unidos de AmericaGreatBookPrices

    Vendedor de 5 estrellas
    Contactar con el vendedor

    Condición: Nuevo

    EUR 29,25

    Envío por EUR 2,35 
    Se envía dentro de Estados Unidos de America

    Cantidad disponible: Más de 20 disponibles

    Condición: New.

  • Idioma: Inglés

    Editorial: Springer International Publishing AG, CH, 2010

    3031007077 / 9783031007071

    • Tapa blanda

    Librería: Rarewaves.com USA, London, LONDO, Reino UnidoRarewaves.com USA

    Vendedor de 5 estrellas
    Contactar con el vendedor

    Condición: Nuevo

    EUR 31,67

     Gastos de envío gratis 
    Se envía de Reino Unido a Estados Unidos de America

    Cantidad disponible: 1 disponible

    Paperback. Condición: New. With the ever increasing volume of data, data quality problems abound. Multiple, yet different representations of the same real-world objects in data, duplicates, are one of the most intriguing data quality problems. The effects of such duplicates are detrimental; for instance, bank customers can obtain duplicate identities, inventory levels are monitored incorrectly, catalogs are mailed multiple times to the same household, etc. Automatically detecting duplicates is difficult: First, duplicate representations are usually not identical but slightly differ in their values. Second, in principle all pairs of records should be compared, which is infeasible for large volumes of data. This lecture examines closely the two main components to overcome these difficulties: (i) Similarity measures are used to automatically identify duplicates when comparing two records. Well-chosen similarity measures improve the effectiveness of duplicate detection. (ii) Algorithms are developed to perform on very large volumes of data in search for duplicates. Well-designed algorithms improve the efficiency of duplicate detection. Finally, we discuss methods to evaluate the success of duplicate detection. Table of Contents: Data Cleansing: Introduction and Motivation / Problem Definition / Similarity Functions / Duplicate Detection Algorithms / Evaluating Detection Success / Conclusion and Outlook / Bibliography.…

  • Idioma: Inglés

    Editorial: Springer, 2010

    3031007077 / 9783031007071

    • Tapa blanda

    Librería: PBShop.store UK, Fairford, GLOS, Reino UnidoPBShop.store UK

    Vendedor de 5 estrellas
    Contactar con el vendedor

    Condición: Nuevo

    EUR 28,04

    Envío por EUR 3,87 
    Se envía de Reino Unido a Estados Unidos de America

    Cantidad disponible: 2 disponibles

    PAP. Condición: New. New Book. Shipped from UK. Established seller since 2000.

  • Idioma: Inglés

    Editorial: Springer International Publishing AG, CH, 2010

    3031007077 / 9783031007071

    • Tapa blanda

    Librería: Rarewaves USA, HEBRON, KY, Estados Unidos de AmericaRarewaves USA

    Vendedor de 5 estrellas
    Contactar con el vendedor

    Condición: Nuevo

    EUR 34,47

     Gastos de envío gratis 
    Se envía dentro de Estados Unidos de America

    Cantidad disponible: 1 disponible

    Paperback. Condición: New. With the ever increasing volume of data, data quality problems abound. Multiple, yet different representations of the same real-world objects in data, duplicates, are one of the most intriguing data quality problems. The effects of such duplicates are detrimental; for instance, bank customers can obtain duplicate identities, inventory levels are monitored incorrectly, catalogs are mailed multiple times to the same household, etc. Automatically detecting duplicates is difficult: First, duplicate representations are usually not identical but slightly differ in their values. Second, in principle all pairs of records should be compared, which is infeasible for large volumes of data. This lecture examines closely the two main components to overcome these difficulties: (i) Similarity measures are used to automatically identify duplicates when comparing two records. Well-chosen similarity measures improve the effectiveness of duplicate detection. (ii) Algorithms are developed to perform on very large volumes of data in search for duplicates. Well-designed algorithms improve the efficiency of duplicate detection. Finally, we discuss methods to evaluate the success of duplicate detection. Table of Contents: Data Cleansing: Introduction and Motivation / Problem Definition / Similarity Functions / Duplicate Detection Algorithms / Evaluating Detection Success / Conclusion and Outlook / Bibliography.…

  • Idioma: Inglés

    Editorial: Springer, 2010

    3031007077 / 9783031007071

    • Tapa blanda

    Librería: Ria Christie Collections, Uxbridge, Reino UnidoRia Christie Collections

    Vendedor de 5 estrellas
    Contactar con el vendedor

    Condición: Nuevo

    EUR 31,86

    Envío por EUR 11,03 
    Se envía de Reino Unido a Estados Unidos de America

    Cantidad disponible: Más de 20 disponibles

    Condición: New. In English.

  • Idioma: Inglés

    Editorial: Springer, 2010

    3031007077 / 9783031007071

    • Tapa blanda

    Librería: GreatBookPricesUK, Woodford Green, Reino UnidoGreatBookPricesUK

    Vendedor de 5 estrellas
    Contactar con el vendedor

    Condición: Nuevo

    EUR 28,03

    Envío por EUR 17,65 
    Se envía de Reino Unido a Estados Unidos de America

    Cantidad disponible: Más de 20 disponibles

    Condición: New.

  • Idioma: Inglés

    Editorial: Springer 2010-03, 2010

    3031007077 / 9783031007071

    • Tapa blanda

    Librería: Chiron Media, Wallingford, Reino UnidoChiron Media

    Vendedor de 5 estrellas
    Contactar con el vendedor

    Condición: Nuevo

    EUR 29,42

    Envío por EUR 18,23 
    Se envía de Reino Unido a Estados Unidos de America

    Cantidad disponible: 10 disponibles

    PF. Condición: New.

  • Idioma: Inglés

    Editorial: Springer, 2010

    3031007077 / 9783031007071

    • Tapa blanda

    Librería: GreatBookPricesUK, Woodford Green, Reino UnidoGreatBookPricesUK

    Vendedor de 5 estrellas
    Contactar con el vendedor

    Condición: Usado - Como Nuevo

    EUR 30,83

    Envío por EUR 17,65 
    Se envía de Reino Unido a Estados Unidos de America

    Cantidad disponible: Más de 20 disponibles

    Condición: As New. Unread book in perfect condition.

  • Idioma: Inglés

    Editorial: Springer, 2010

    3031007077 / 9783031007071

    • Tapa blanda

    Librería: Books Puddle, Woodside, NY, Estados Unidos de AmericaBooks Puddle

    Vendedor de 4 estrellas
    Contactar con el vendedor

    Condición: Nuevo

    EUR 45,47

    Envío por EUR 3,55 
    Se envía dentro de Estados Unidos de America

    Cantidad disponible: 4 disponibles

    Condición: New. 1st edition NO-PA16APR2015-KAP.

  • Idioma: Inglés

    Editorial: Springer, 2010

    3031007077 / 9783031007071

    • Tapa blanda

    Librería: AHA-BUCH GmbH, Einbeck, AlemaniaAHA-BUCH GmbH

    Vendedor de 5 estrellas
    Contactar con el vendedor

    Condición: Nuevo

    EUR 28,99

    Envío por EUR 35,00 
    Se envía de Alemania a Estados Unidos de America

    Cantidad disponible: 1 disponible

    Taschenbuch. Condición: Neu. Druck auf Anfrage Neuware - Printed after ordering - With the ever increasing volume of data, data quality problems abound. Multiple, yet different representations of the same real-world objects in data, duplicates, are one of the most intriguing data quality problems. The effects of such duplicates are detrimental; for instance, bank customers can obtain duplicate identities, inventory levels are monitored incorrectly, catalogs are mailed multiple times to the same household, etc. Automatically detecting duplicates is difficult: First, duplicate representations are usually not identical but slightly differ in their values. Second, in principle all pairs of records should be compared, which is infeasible for large volumes of data. This lecture examines closely the two main components to overcome these difficulties: (i) Similarity measures are used to automatically identify duplicates when comparing two records. Well-chosen similarity measures improve the effectiveness of duplicate detection. (ii) Algorithms are developed to perform on very large volumes of data in search for duplicates. Well-designed algorithms improve the efficiency of duplicate detection. Finally, we discuss methods to evaluate the success of duplicate detection. Table of Contents: Data Cleansing: Introduction and Motivation / Problem Definition / Similarity Functions / Duplicate Detection Algorithms / Evaluating Detection Success / Conclusion and Outlook / Bibliography. …

  • Idioma: Inglés

    Editorial: Springer, 2010

    3031007077 / 9783031007071

    • Tapa blanda

    Librería: Speedyhen, Hertfordshire, Reino UnidoSpeedyhen

    Vendedor de 5 estrellas
    Contactar con el vendedor

    Condición: Nuevo

    EUR 24,91

    Envío por EUR 48,24 
    Se envía de Reino Unido a Estados Unidos de America

    Cantidad disponible: 2 disponibles

    Condición: NEW.

  • Idioma: Inglés

    Editorial: Springer International Publishing AG, CH, 2010

    3031007077 / 9783031007071

    • Tapa blanda

    Librería: Rarewaves USA United, HEBRON, KY, Estados Unidos de AmericaRarewaves USA United

    Vendedor de 5 estrellas
    Contactar con el vendedor

    Condición: Nuevo

    EUR 36,39

    Envío por EUR 44,44 
    Se envía dentro de Estados Unidos de America

    Cantidad disponible: 1 disponible

    Paperback. Condición: New. With the ever increasing volume of data, data quality problems abound. Multiple, yet different representations of the same real-world objects in data, duplicates, are one of the most intriguing data quality problems. The effects of such duplicates are detrimental; for instance, bank customers can obtain duplicate identities, inventory levels are monitored incorrectly, catalogs are mailed multiple times to the same household, etc. Automatically detecting duplicates is difficult: First, duplicate representations are usually not identical but slightly differ in their values. Second, in principle all pairs of records should be compared, which is infeasible for large volumes of data. This lecture examines closely the two main components to overcome these difficulties: (i) Similarity measures are used to automatically identify duplicates when comparing two records. Well-chosen similarity measures improve the effectiveness of duplicate detection. (ii) Algorithms are developed to perform on very large volumes of data in search for duplicates. Well-designed algorithms improve the efficiency of duplicate detection. Finally, we discuss methods to evaluate the success of duplicate detection. Table of Contents: Data Cleansing: Introduction and Motivation / Problem Definition / Similarity Functions / Duplicate Detection Algorithms / Evaluating Detection Success / Conclusion and Outlook / Bibliography.…

  • Idioma: Inglés

    Editorial: Springer, Berlin|Springer International Publishing|Morgan & Claypool|Springer, 2010

    3031007077 / 9783031007071

    • Tapa blanda

    Librería: moluna, Greven, Alemaniamoluna

    Vendedor de 5 estrellas
    Contactar con el vendedor

    Condición: Nuevo

    EUR 32,08

    Envío por EUR 48,99 
    Se envía de Alemania a Estados Unidos de America

    Cantidad disponible: 2 disponibles

    Condición: New. With the ever increasing volume of data, data quality problems abound. Multiple, yet different representations of the same real-world objects in data, duplicates, are one of the most intriguing data quality problems. The effects of such duplicates are detri.

  • Idioma: Inglés

    Editorial: Springer International Publishing AG, Cham, 2010

    3031007077 / 9783031007071

    • Tapa blanda

    Librería: AussieBookSeller, Truganina, VIC, AustraliaAussieBookSeller

    Vendedor de 5 estrellas
    Contactar con el vendedor

    Condición: Nuevo

    EUR 52,63

    Envío por EUR 32,88 
    Se envía de Australia a Estados Unidos de America

    Cantidad disponible: 1 disponible

    Paperback. Condición: new. Paperback. With the ever increasing volume of data, data quality problems abound. Multiple, yet different representations of the same real-world objects in data, duplicates, are one of the most intriguing data quality problems. The effects of such duplicates are detrimental; for instance, bank customers can obtain duplicate identities, inventory levels are monitored incorrectly, catalogs are mailed multiple times to the same household, etc. Automatically detecting duplicates is difficult: First, duplicate representations are usually not identical but slightly differ in their values. Second, in principle all pairs of records should be compared, which is infeasible for large volumes of data. This lecture examines closely the two main components to overcome these difficulties: (i) Similarity measures are used to automatically identify duplicates when comparing two records. Well-chosen similarity measures improve the effectiveness of duplicate detection. (ii) Algorithms are developed to perform on very large volumes of data in search for duplicates. Well-designed algorithms improve the efficiency of duplicate detection. Finally, we discuss methods to evaluate the success of duplicate detection. Table of Contents: Data Cleansing: Introduction and Motivation / Problem Definition / Similarity Functions / Duplicate Detection Algorithms / Evaluating Detection Success / Conclusion and Outlook / Bibliography With the ever increasing volume of data, data quality problems abound. Table of Contents: Data Cleansing: Introduction and Motivation / Problem Definition / Similarity Functions / Duplicate Detection Algorithms / Evaluating Detection Success / Conclusion and Outlook / Bibliography Shipping may be from our Sydney, NSW warehouse or from our UK or US warehouse, depending on stock availability.…

  • Idioma: Inglés

    Editorial: Springer International Publishing AG, CH, 2010

    3031007077 / 9783031007071

    • Tapa blanda

    Librería: Rarewaves.com UK, London, Reino UnidoRarewaves.com UK

    Vendedor de 5 estrellas
    Contactar con el vendedor

    Condición: Nuevo

    EUR 29,57

    Envío por EUR 76,48 
    Se envía de Reino Unido a Estados Unidos de America

    Cantidad disponible: 1 disponible

    Paperback. Condición: New. With the ever increasing volume of data, data quality problems abound. Multiple, yet different representations of the same real-world objects in data, duplicates, are one of the most intriguing data quality problems. The effects of such duplicates are detrimental; for instance, bank customers can obtain duplicate identities, inventory levels are monitored incorrectly, catalogs are mailed multiple times to the same household, etc. Automatically detecting duplicates is difficult: First, duplicate representations are usually not identical but slightly differ in their values. Second, in principle all pairs of records should be compared, which is infeasible for large volumes of data. This lecture examines closely the two main components to overcome these difficulties: (i) Similarity measures are used to automatically identify duplicates when comparing two records. Well-chosen similarity measures improve the effectiveness of duplicate detection. (ii) Algorithms are developed to perform on very large volumes of data in search for duplicates. Well-designed algorithms improve the efficiency of duplicate detection. Finally, we discuss methods to evaluate the success of duplicate detection. Table of Contents: Data Cleansing: Introduction and Motivation / Problem Definition / Similarity Functions / Duplicate Detection Algorithms / Evaluating Detection Success / Conclusion and Outlook / Bibliography.…

  • Idioma: Inglés

    Editorial: Springer, 2010

    3031007077 / 9783031007071

    • Tapa blanda
    • Impresión bajo demanda

    Librería: Brook Bookstore On Demand, Napoli, NA, ItaliaBrook Bookstore On Demand

    Vendedor de 5 estrellas
    Contactar con el vendedor

    Condición: Nuevo

    EUR 26,21

    Envío por EUR 4,00 
    Se envía de Italia a Estados Unidos de America

    Cantidad disponible: Más de 20 disponibles

    Condición: new. Questo è un articolo print on demand.

  • Idioma: Inglés

    Editorial: Springer-Nature New York Inc, 2010

    3031007077 / 9783031007071

    • Tapa blanda
    • Impresión bajo demanda

    Librería: Revaluation Books, Exeter, Reino UnidoRevaluation Books

    Vendedor de 5 estrellas
    Contactar con el vendedor

    Condición: Nuevo

    EUR 30,35

    Envío por EUR 11,77 
    Se envía de Reino Unido a Estados Unidos de America

    Cantidad disponible: 2 disponibles

    Paperback. Condición: Brand New. 86 pages. 9.25x7.51x9.25 inches. In Stock. This item is printed on demand.

  • Idioma: Inglés

    Editorial: Springer, 2010

    3031007077 / 9783031007071

    • Tapa blanda
    • Impresión bajo demanda

    Librería: Majestic Books, Hounslow, Reino UnidoMajestic Books

    Vendedor de 4 estrellas
    Contactar con el vendedor

    Condición: Nuevo

    EUR 42,61

    Envío por EUR 7,65 
    Se envía de Reino Unido a Estados Unidos de America

    Cantidad disponible: 4 disponibles

    Condición: New. Print on Demand.

  • Idioma: Inglés

    Editorial: Springer International Publishing Mrz 2010, 2010

    3031007077 / 9783031007071

    • Tapa blanda
    • Impresión bajo demanda

    Librería: BuchWeltWeit Ludwig Meier e.K., Bergisch Gladbach, AlemaniaBuchWeltWeit Ludwig Meier e.K.

    Vendedor de 5 estrellas
    Contactar con el vendedor

    Condición: Nuevo

    EUR 26,74

    Envío por EUR 23,00 
    Se envía de Alemania a Estados Unidos de America

    Cantidad disponible: 2 disponibles

    Taschenbuch. Condición: Neu. This item is printed on demand - it takes 3-4 days longer - Neuware -With the ever increasing volume of data, data quality problems abound. Multiple, yet different representations of the same real-world objects in data, duplicates, are one of the most intriguing data quality problems. The effects of such duplicates are detrimental; for instance, bank customers can obtain duplicate identities, inventory levels are monitored incorrectly, catalogs are mailed multiple times to the same household, etc. Automatically detecting duplicates is difficult: First, duplicate representations are usually not identical but slightly differ in their values. Second, in principle all pairs of records should be compared, which is infeasible for large volumes of data. This lecture examines closely the two main components to overcome these difficulties: (i) Similarity measures are used to automatically identify duplicates when comparing two records. Well-chosen similarity measures improve the effectiveness of duplicate detection. (ii) Algorithms are developed to perform on very large volumes of data in search for duplicates. Well-designed algorithms improve the efficiency of duplicate detection. Finally, we discuss methods to evaluate the success of duplicate detection. Table of Contents: Data Cleansing: Introduction and Motivation / Problem Definition / Similarity Functions / Duplicate Detection Algorithms / Evaluating Detection Success / Conclusion and Outlook / Bibliography 88 pp. Englisch.…

  • Idioma: Inglés

    Editorial: Springer, 2010

    3031007077 / 9783031007071

    • Tapa blanda
    • Impresión bajo demanda

    Librería: Biblios, frankfurt am main, HESSE, AlemaniaBiblios

    Vendedor de 4 estrellas
    Contactar con el vendedor

    Condición: Nuevo

    EUR 42,74

    Envío por EUR 9,95 
    Se envía de Alemania a Estados Unidos de America

    Cantidad disponible: 4 disponibles

    Condición: New. PRINT ON DEMAND.

  • Idioma: Inglés

    Editorial: Springer, Springer Mär 2010, 2010

    3031007077 / 9783031007071

    • Tapa blanda
    • Impresión bajo demanda

    Librería: buchversandmimpf2000, Emtmannsberg, BAYE, Alemaniabuchversandmimpf2000

    Vendedor de 5 estrellas
    Contactar con el vendedor

    Condición: Nuevo

    EUR 26,74

    Envío por EUR 60,00 
    Se envía de Alemania a Estados Unidos de America

    Cantidad disponible: 1 disponible

    Taschenbuch. Condición: Neu. This item is printed on demand - Print on Demand Titel. Neuware -With the ever increasing volume of data, data quality problems abound. Multiple, yet different representations of the same real-world objects in data, duplicates, are one of the most intriguing data quality problems. The effects of such duplicates are detrimental; for instance, bank customers can obtain duplicate identities, inventory levels are monitored incorrectly, catalogs are mailed multiple times to the same household, etc. Automatically detecting duplicates is difficult: First, duplicate representations are usually not identical but slightly differ in their values. Second, in principle all pairs of records should be compared, which is infeasible for large volumes of data. This lecture examines closely the two main components to overcome these difficulties: (i) Similarity measures are used to automatically identify duplicates when comparing two records. Well-chosen similarity measures improve the effectiveness of duplicate detection. (ii) Algorithms are developed to perform on very large volumes of data in search for duplicates. Well-designed algorithms improve the efficiency of duplicate detection. Finally, we discuss methods to evaluate the success of duplicate detection. Table of Contents: Data Cleansing: Introduction and Motivation / Problem Definition / Similarity Functions / Duplicate Detection Algorithms / Evaluating Detection Success / Conclusion and Outlook / BibliographySpringer-Verlag KG, Sachsenplatz 4-6, 1201 Wien 88 pp. Englisch.…