000 05318nam a22004575i 4500
999 _c387476
_d387476
001 387476
003 ES-MaUEC
005 20230411113229.0
006 a||||fo|||| 00| 0
007 cr nn 008mamaa
008 230411s2020 sz | s |||| 0|eng d
020 _a9783031023224
024 7 _a10.1007/978-3-031-02322-4
_2doi
040 _aES-MaUEC
_bspa
_cES-MaUEC
_dES-MaUEC
050 4 _aQA76.9.N38
_b2020 EB
100 1 _aFerreira, Anderson A.
_eautor
_4aut
_4http://id.loc.gov/vocabulary/relators/aut
_9688050
245 1 0 _aAutomatic Disambiguation of Author Names in Bibliographic Repositories
_cby Anderson A. Ferreira, Marcos André Gonçalves, Alberto H. F. Laender
250 _a1st edition 2020
264 1 _aCham
_bSpringer International Publishing
_c2020
300 _a1 recurso en línea (XX, 126 páginas)
336 _atexto
_btxt
_2rdacontent
337 _aelectrónico
_bc
_2rdamedia
338 _arecurso electrónico
_bcr
_2rdacarrier
347 _aarchivo de texto
_bPDF
490 0 _aSynthesis Lectures on Information Concepts Retrieval and Services
_x1947-9468
505 0 _aPreface -- Introduction -- The Author Name Disambiguation Task -- Foundations -- Taxonomy -- Heuristic-Based Hierarchical Clustering Disambiguation -- SAND: Self-Training Author Name Disambiguator -- Incremental Author Name Disambiguation -- Additional Methods for Author Name Disambiguation -- Bibliography -- Authors' Biographies.
520 _aThis book deals with a hard problem that is inherent to human language: ambiguity. In particular, we focus on author name ambiguity, a type of ambiguity that exists in digital bibliographic repositories, which occurs when an author publishes works under distinct names or distinct authors publish works under similar names. This problem may be caused by a number of reasons, including the lack of standards and common practices, and the decentralized generation of bibliographic content. As a consequence, the quality of the main services of digital bibliographic repositories such as search, browsing, and recommendation may be severely affected by author name ambiguity. The focal point of the book is on automatic methods, since manual solutions do not scale to the size of the current repositories or the speed in which they are updated. Accordingly, we provide an ample view on the problem of automatic disambiguation of author names, summarizing the results of more than a decade of research on this topic conducted by our group, which were reported in more than a dozen publications that received over 900 citations so far, according to Google Scholar. We start by discussing its motivational issues (Chapter 1). Next, we formally define the author name disambiguation task (Chapter 2) and use this formalization to provide a brief, taxonomically organized, overview of the literature on the topic (Chapter 3). We then organize, summarize and integrate the efforts of our own group on developing solutions for the problem that have historically produced state-of-the-art (by the time of their proposals) results in terms of the quality of the disambiguation results. Thus, Chapter 4 covers HHC - Heuristic-based Clustering, an author name disambiguation method that is based on two specific real-world assumptions regarding scientific authorship. Then, Chapter 5 describes SAND - Self-training Author Name Disambiguator and Chapter 6 presents two incremental author name disambiguation methods, namely INDi - Incremental Unsupervised Name Disambiguation and INC- Incremental Nearest Cluster. Finally, Chapter 7 provides an overview of recent author name disambiguation methods that address new specific approaches such as graph-based representations, alternative predefined similarity functions, visualization facilities and approaches based on artificial neural networks. The chapters are followed by three appendices that cover, respectively: (i) a pattern matching function for comparing proper names and used by some of the methods addressed in this book; (ii) a tool for generating synthetic collections of citation records for distinct experimental tasks; and (iii) a number of datasets commonly used to evaluate author name disambiguation methods. In summary, the book organizes a large body of knowledge and work in the area of author name disambiguation in the last decade, hoping to consolidate a solid basis for future developments in the field.
988 _aSynthesis Collection of Technology_2020
650 7 _2embne
_9158738
_aProceso en lenguaje natural (Informática)
650 7 _2embne
_9669583
_aAnálisis Cluster
650 7 _2embne
_9688051
_aAmbigüedad
700 1 _aGonçalves, Marcos André
_eautor
_4aut
_4http://id.loc.gov/vocabulary/relators/aut
_9686363
700 1 _aLaender, Alberto H. F.,
_eautor
_4aut
_4http://id.loc.gov/vocabulary/relators/aut
_9688052
_d1951-
776 0 8 _iPrinted edition:
_z9783031002298
776 0 8 _iPrinted edition:
_z9783031011948
776 0 8 _iPrinted edition:
_z9783031034503
856 4 0 _uhttps://go.openathens.net/redirector/universidadeuropea.es?url=https://doi.org/10.1007/978-3-031-02322-4
_zAcceso a este recurso digital (usuarios Universidad Europea de Madrid)
942 _2lcc
_cLE
998 _b04/2023
_dz
_eIG
_zSI