Natural Language Processing for Historical Texts is popular PDF and ePub book, written by Michael Piotrowski in 2012-09-01, it is a fantastic choice for those who relish reading online the Computers genre. Let's immerse ourselves in this engaging Computers book by exploring the summary and details provided below. Remember, Natural Language Processing for Historical Texts can be Read Online from any device for your convenience.

Natural Language Processing for Historical Texts Book PDF Summary

More and more historical texts are becoming available in digital form. Digitization of paper documents is motivated by the aim of preserving cultural heritage and making it more accessible, both to laypeople and scholars. As digital images cannot be searched for text, digitization projects increasingly strive to create digital text, which can be searched and otherwise automatically processed, in addition to facsimiles. Indeed, the emerging field of digital humanities heavily relies on the availability of digital text for its studies. Together with the increasing availability of historical texts in digital form, there is a growing interest in applying natural language processing (NLP) methods and tools to historical texts. However, the specific linguistic properties of historical texts -- the lack of standardized orthography, in particular -- pose special challenges for NLP. This book aims to give an introduction to NLP for historical texts and an overview of the state of the art in this field. The book starts with an overview of methods for the acquisition of historical texts (scanning and OCR), discusses text encoding and annotation schemes, and presents examples of corpora of historical texts in a variety of languages. The book then discusses specific methods, such as creating part-of-speech taggers for historical languages or handling spelling variation. A final chapter analyzes the relationship between NLP and the digital humanities. Certain recently emerging textual genres, such as SMS, social media, and chat messages, or newsgroup and forum postings share a number of properties with historical texts, for example, nonstandard orthography and grammar, and profuse use of abbreviations. The methods and techniques required for the effective processing of historical texts are thus also of interest for research in other domains. Table of Contents: Introduction / NLP and Digital Humanities / Spelling in Historical Texts / Acquiring Historical Texts / Text Encoding and Annotation Schemes / Handling Spelling Variation / NLP Tools for Historical Languages / Historical Corpora / Conclusion / Bibliography

Detail Book of Natural Language Processing for Historical Texts PDF

Natural Language Processing for Historical Texts
  • Author : Michael Piotrowski
  • Release : 01 September 2012
  • Publisher : Morgan & Claypool Publishers
  • ISBN : 9781608459476
  • Genre : Computers
  • Total Page : 159 pages
  • Language : English
  • PDF File Size : 16,5 Mb

If you're still pondering over how to secure a PDF or EPUB version of the book Natural Language Processing for Historical Texts by Michael Piotrowski, don't worry! All you have to do is click the 'Get Book' buttons below to kick off your Download or Read Online journey. Just a friendly reminder: we don't upload or host the files ourselves.

Get Book

Biomedical Natural Language Processing

Biomedical Natural Language Processing Author : Kevin Bretonnel Cohen,Dina Demner-Fushman
Publisher : John Benjamins Publishing Company
File Size : 10,5 Mb
Get Book
Biomedical Natural Language Processing is a comprehensive tour through the classic and current work ...

Turkish Natural Language Processing

Turkish Natural Language Processing Author : Kemal Oflazer,Murat Saraçlar
Publisher : Springer
File Size : 18,6 Mb
Get Book
This book brings together work on Turkish natural language and speech processing over the last 25 ye...