Natural Language Processing for Historical Texts (Record no. 387368)

MARC details
000 -CABECERA
campo de control de longitud fija 04083nam a22004215i 4500
001 - NÚMERO DE CONTROL
campo de control 387368
003 - IDENTIFICADOR DEL NÚMERO DE CONTROL
campo de control ES-MaUEC
005 - FECHA Y HORA DE LA ÚLTIMA TRANSACCIÓN
campo de control 20230315174652.0
006 - CÓDIGOS DE INFORMACIÓN DE LONGITUD FIJA--CARACTERÍSTICAS DEL MATERIAL ADICIONAL
campo de control de longitud fija a||||fo|||| 00| 0
007 - CAMPO FIJO DE DESCRIPCIÓN FÍSICA--INFORMACIÓN GENERAL
campo de control de longitud fija cr nn 008mamaa
008 - DATOS DE LONGITUD FIJA--INFORMACIÓN GENERAL
campo de control de longitud fija 220601s2012 sz | s |||| 0|eng d
020 ## - NÚMERO INTERNACIONAL ESTÁNDAR DEL LIBRO
Número Internacional Estándar del Libro 9783031021466
024 7# - IDENTIFICADOR DE OTROS ESTÁNDARES
Número estándar o código 10.1007/978-3-031-02146-6
Fuente del número o código doi
040 ## - FUENTE DE LA CATALOGACIÓN
Centro catalogador/agencia de origen ES-MaUEC
Lengua de catalogación spa
Centro/agencia transcriptor ES-MaUEC
Centro/agencia modificador ES-MaUEC
050 #4 - SIGNATURA TOPOGRÁFICA DE LA BIBLIOTECA DEL CONGRESO
Número de clasificación P140
Número de documento/Ítem 2012 EB
100 1# - ENTRADA PRINCIPAL--NOMBRE DE PERSONA
Nombre de persona Piotrowski, Michael
Término indicativo de función/relación autor
Código de función/relación aut
-- http://id.loc.gov/vocabulary/relators/aut
9 (RLIN) 687391
245 10 - MENCIÓN DE TÍTULO
Título Natural Language Processing for Historical Texts
Mención de responsabilidad, etc. by Michael Piotrowski
250 ## - MENCIÓN DE EDICIÓN
Mención de edición 1st edition 2012
264 #1 - PRODUCCIÓN, PUBLICACIÓN, DISTRIBUCIÓN, FABRICACIÓN Y COPYRIGHT
Producción, publicación, distribución, fabricación y copyright Cham
Nombre del de productor, editor, distribuidor, fabricante Springer International Publishing
Fecha de producción, publicación, distribución, fabricación o copyright 2012
300 ## - DESCRIPCIÓN FÍSICA
Extensión 1 recurso en línea (XII, 145 páginas)
336 ## - TIPO DE CONTENIDO
Término de tipo de contenido texto
Código de tipo de contenido txt
Fuente rdacontent
337 ## - TIPO DE MEDIO
Nombre/término del tipo de medio electrónico
Código del tipo de medio c
Fuente rdamedia
338 ## - TIPO DE SOPORTE
Nombre/término del tipo de soporte recurso electrónico
Código del tipo de soporte cr
Fuente rdacarrier
347 ## - CARACTERÍSTICAS DEL ARCHIVO DIGITAL
Tipo de archivo archivo de texto
Formato de codificación PDF
490 0# - MENCIÓN DE SERIE
Mención de serie Synthesis Lectures on Human Language Technologies
Número Internacional Normalizado para Publicaciones Seriadas 1947-4059
505 0# - NOTA DE CONTENIDO CON FORMATO
Nota de contenido con formato Introduction -- NLP and Digital Humanities -- Spelling in Historical Texts -- Acquiring Historical Texts -- Text Encoding and Annotation Schemes -- Handling Spelling Variation -- NLP Tools for Historical Languages -- Historical Corpora -- Conclusion -- Bibliography.
520 ## - SUMARIO, ETC.
Sumario, etc. More and more historical texts are becoming available in digital form. Digitization of paper documents is motivated by the aim of preserving cultural heritage and making it more accessible, both to laypeople and scholars. As digital images cannot be searched for text, digitization projects increasingly strive to create digital text, which can be searched and otherwise automatically processed, in addition to facsimiles. Indeed, the emerging field of digital humanities heavily relies on the availability of digital text for its studies. Together with the increasing availability of historical texts in digital form, there is a growing interest in applying natural language processing (NLP) methods and tools to historical texts. However, the specific linguistic properties of historical texts -- the lack of standardized orthography, in particular -- pose special challenges for NLP. This book aims to give an introduction to NLP for historical texts and an overview of the state of the art in this field. The book starts with an overview of methods for the acquisition of historical texts (scanning and OCR), discusses text encoding and annotation schemes, and presents examples of corpora of historical texts in a variety of languages. The book then discusses specific methods, such as creating part-of-speech taggers for historical languages or handling spelling variation. A final chapter analyzes the relationship between NLP and the digital humanities. Certain recently emerging textual genres, such as SMS, social media, and chat messages, or newsgroup and forum postings share a number of properties with historical texts, for example, nonstandard orthography and grammar, and profuse use of abbreviations. The methods and techniques required for the effective processing of historical texts are thus also of interest for research in other domains. Table of Contents: Introduction / NLP and Digital Humanities / Spelling in Historical Texts / Acquiring Historical Texts / Text Encoding and Annotation Schemes / Handling Spelling Variation / NLP Tools for Historical Languages / Historical Corpora / Conclusion / Bibliography.
988 ## - NOTA LOCAL 598
Nota local 598 (boletines) Synthesis Collection of Technology_2012
650 #7 - PUNTO DE ACCESO ADICIONAL DE MATERIA--TÉRMINO DE MATERIA
Fuente del encabezamiento o término embne
9 (RLIN) 158738
Término de materia o nombre geográfico como elemento de entrada Proceso en lenguaje natural (Informática)
650 #7 - PUNTO DE ACCESO ADICIONAL DE MATERIA--TÉRMINO DE MATERIA
Fuente del encabezamiento o término embne
9 (RLIN) 683294
Término de materia o nombre geográfico como elemento de entrada Lingüística histórica
650 #7 - PUNTO DE ACCESO ADICIONAL DE MATERIA--TÉRMINO DE MATERIA
Fuente del encabezamiento o término embne
9 (RLIN) 414657
Término de materia o nombre geográfico como elemento de entrada Preservación digital
776 08 - ENTRADA/ENLACE A UN FORMATO FÍSICO ADICIONAL
Información de relación/Frase instructiva de referencia Printed edition:
Número Internacional Estándar del Libro 9783031010187
776 08 - ENTRADA/ENLACE A UN FORMATO FÍSICO ADICIONAL
Información de relación/Frase instructiva de referencia Printed edition:
Número Internacional Estándar del Libro 9783031032745
856 40 - LOCALIZACIÓN Y ACCESO ELECTRÓNICOS
Identificador Uniforme del Recurso https://go.openathens.net/redirector/universidadeuropea.es?url=https://doi.org/10.1007/978-3-031-02146-6
Nota pública Acceso a este recurso digital (usuarios Universidad Europea de Madrid)
942 ## - ELEMENTOS DE PUNTO DE ACCESO ADICIONAL (KOHA)
Fuente del sistema de clasificación o colocación Library of Congress Classification
Tipo de ítem Koha LIBRO-E NO PRÉSTAMO
998 ## - DATOS ESTADÍSTICOS
Fecha de catalogación 03/2023
Tipo de materia E-book
Catalogador Susana Carmen Delgado Sanz
Catalogado
Holdings
Información adicional para el OPAC Código 2 (categoría) Estado de pérdida Fuente del sistema de clasificación o colocación Tipo de material Código 1: Estado físico No se presta Código de colección Estado Localización permanente Ubicación/localización actual Ubicación en estantería Fecha de adquisición Tipo de préstamo Total de préstamos Signatura topográfica completa Código de barras Fecha visto por última vez Precio válido a partir de Tipo de ítem Koha
Acceso concurrente No retirado   Library of Congress Classification E-Libro Buen estado Acceso electrónico Ciencias e Ingeniería Acceso electrónico Madrid Digital Madrid Digital Acceso Electrónico (UEM) 25/11/2022 En línea   P140 2012 EB eBook.01112567 25/11/2022 25/11/2022 LIBRO-E NO PRÉSTAMO