Overview
Part of the book series: The Springer International Series in Engineering and Computer Science (SECS, volume 278)
Access this book
Tax calculation will be finalised at checkout
Other ways to access
About this book
The techniques are tested on twenty different corpora ranging from baseball newsgroups, assassination archives, medical X-ray reports, abstracts on AIDS, to encyclopedia articles on animals, even on the text of the book itself. The corpora range from 40,000 to 6 million characters of text, and results are presented for each in the Appendix.
The methods described in the book have undergone extensive evaluation. Their time and space complexity are shown to be modest. The results are shown to converge to a stable state as the corpus grows. The similarities calculated are compared to those produced by psychological testing. A method of evaluation using Artificial Synonyms is tested. Gold Standards evaluation show that techniques significantly outperform non-linguistic-based techniques for the most important words in corpora.
Explorations in Automatic Thesaurus Discovery includes applications to the fields of information retrieval using established testbeds, existing thesaural enrichment, semantic analysis. Also included are applications showing how to create, implement, and test a first-draft thesaurus.
Similar content being viewed by others
Keywords
Table of contents (6 chapters)
Authors and Affiliations
Bibliographic Information
Book Title: Explorations in Automatic Thesaurus Discovery
Authors: Gregory Grefenstette
Series Title: The Springer International Series in Engineering and Computer Science
DOI: https://doi.org/10.1007/978-1-4615-2710-7
Publisher: Springer New York, NY
-
eBook Packages: Springer Book Archive
Copyright Information: Springer Science+Business Media New York 1994
Hardcover ISBN: 978-0-7923-9468-6Published: 31 July 1994
Softcover ISBN: 978-1-4613-6167-1Published: 21 November 2012
eBook ISBN: 978-1-4615-2710-7Published: 06 December 2012
Series ISSN: 0893-3405
Edition Number: 1
Number of Pages: XIII, 305
Topics: Artificial Intelligence, Natural Language Processing (NLP)