Authors:

Gregory Grefenstette ⁰

Gregory Grefenstette
1. University of Pittsburgh, Pittsburgh, USA
  Rank Xerox Research Centre, Grenoble, France
View author publications

You can also search for this author in PubMed Google Scholar

Part of the book series: The Springer International Series in Engineering and Computer Science (SECS, volume 278)

524 Accesses
186 Citations
10 Altmetric

Buy it now

eBook USD 129.00

Price excludes VAT (USA)

Softcover Book USD 169.99

Price excludes VAT (USA)

Hardcover Book USD 169.99

Price excludes VAT (USA)

Tax calculation will be finalised at checkout

Other ways to access

Licence this eBook for your library

Learn about institutional subscriptions

This is a preview of subscription content, log in via an institution to check for access.

Table of contents (6 chapters)

Front Matter

Pages i-xiii

PDF
Introduction
- Gregory Grefenstette
Pages 1-5
Semantic Extraction
- Gregory Grefenstette
Pages 7-32
Sextant
- Gregory Grefenstette
Pages 33-68
Evaluation
- Gregory Grefenstette
Pages 69-100
Applications
- Gregory Grefenstette
Pages 101-135
Conclusion
- Gregory Grefenstette
Pages 137-148
Back Matter

Pages 149-305

PDF

About this book

Explorations in Automatic Thesaurus Discovery presents an automated method for creating a first-draft thesaurus from raw text. It describes natural processing steps of tokenization, surface syntactic analysis, and syntactic attribute extraction. From these attributes, word and term similarity is calculated and a thesaurus is created showing important common terms and their relation to each other, common verb--noun pairings, common expressions, and word family members.
The techniques are tested on twenty different corpora ranging from baseball newsgroups, assassination archives, medical X-ray reports, abstracts on AIDS, to encyclopedia articles on animals, even on the text of the book itself. The corpora range from 40,000 to 6 million characters of text, and results are presented for each in the Appendix.
The methods described in the book have undergone extensive evaluation. Their time and space complexity are shown to be modest. The results are shown to converge to a stable state as the corpus grows. The similarities calculated are compared to those produced by psychological testing. A method of evaluation using Artificial Synonyms is tested. Gold Standards evaluation show that techniques significantly outperform non-linguistic-based techniques for the most important words in corpora.
Explorations in Automatic Thesaurus Discovery includes applications to the fields of information retrieval using established testbeds, existing thesaural enrichment, semantic analysis. Also included are applications showing how to create, implement, and test a first-draft thesaurus.

Keywords

Authors and Affiliations

University of Pittsburgh, Pittsburgh, USA

Gregory Grefenstette
Rank Xerox Research Centre, Grenoble, France

Gregory Grefenstette

Bibliographic Information

Book Title: Explorations in Automatic Thesaurus Discovery
Authors: Gregory Grefenstette
Series Title: The Springer International Series in Engineering and Computer Science
DOI: https://doi.org/10.1007/978-1-4615-2710-7
Publisher: Springer New York, NY
eBook Packages: Springer Book Archive
Copyright Information: Springer Science+Business Media New York 1994
Hardcover ISBN: 978-0-7923-9468-6Published: 31 July 1994
Softcover ISBN: 978-1-4613-6167-1Published: 21 November 2012
eBook ISBN: 978-1-4615-2710-7Published: 06 December 2012
Series ISSN: 0893-3405
Edition Number: 1
Number of Pages: XIII, 305
Topics: Artificial Intelligence, Natural Language Processing (NLP)

Publish with us

Policies and ethics

Authors:

Sections

Buy it now

Buying options

Other ways to access

Table of contents (6 chapters)

Front Matter

Introduction

Semantic Extraction

Sextant

Evaluation

Applications

Conclusion

Back Matter

About this book

Keywords

Authors and Affiliations

University of Pittsburgh, Pittsburgh, USA

Rank Xerox Research Centre, Grenoble, France

Bibliographic Information

Publish with us

Buy it now

Buying options

Other ways to access

Search

Navigation