Cantitate/Preț
Produs

Explorations in Automatic Thesaurus Discovery: The Springer International Series in Engineering and Computer Science, cartea 278

Autor Gregory Grefenstette
en Limba Engleză Paperback – 21 noi 2012
Explorations in Automatic Thesaurus Discovery presents an automated method for creating a first-draft thesaurus from raw text. It describes natural processing steps of tokenization, surface syntactic analysis, and syntactic attribute extraction. From these attributes, word and term similarity is calculated and a thesaurus is created showing important common terms and their relation to each other, common verb--noun pairings, common expressions, and word family members.
The techniques are tested on twenty different corpora ranging from baseball newsgroups, assassination archives, medical X-ray reports, abstracts on AIDS, to encyclopedia articles on animals, even on the text of the book itself. The corpora range from 40,000 to 6 million characters of text, and results are presented for each in the Appendix.
The methods described in the book have undergone extensive evaluation. Their time and space complexity are shown to be modest. The results are shown to converge to a stable state as the corpus grows. The similarities calculated are compared to those produced by psychological testing. A method of evaluation using Artificial Synonyms is tested. Gold Standards evaluation show that techniques significantly outperform non-linguistic-based techniques for the most important words in corpora.
Explorations in Automatic Thesaurus Discovery includes applications to the fields of information retrieval using established testbeds, existing thesaural enrichment, semantic analysis. Also included are applications showing how to create, implement, and test a first-draft thesaurus.
Citește tot Restrânge

Toate formatele și edițiile

Toate formatele și edițiile Preț Express
Paperback (1) 97165 lei  6-8 săpt.
  Springer Us – 21 noi 2012 97165 lei  6-8 săpt.
Hardback (1) 97685 lei  6-8 săpt.
  Springer Us – 31 iul 1994 97685 lei  6-8 săpt.

Din seria The Springer International Series in Engineering and Computer Science

Preț: 97165 lei

Preț vechi: 121457 lei
-20% Nou

Puncte Express: 1457

Preț estimativ în valută:
18601 19335$ 15423£

Carte tipărită la comandă

Livrare economică 07-21 februarie 25

Preluare comenzi: 021 569.72.76

Specificații

ISBN-13: 9781461361671
ISBN-10: 1461361672
Pagini: 324
Ilustrații: XIII, 305 p.
Dimensiuni: 155 x 235 x 17 mm
Greutate: 0.45 kg
Ediția:Softcover reprint of the original 1st ed. 1994
Editura: Springer Us
Colecția Springer
Seria The Springer International Series in Engineering and Computer Science

Locul publicării:New York, NY, United States

Public țintă

Research

Cuprins

1 INTRODUCTION.- 2 SEMANTIC EXTRACTION.- 2.1 Historical Overview.- 2.2 Cognitive Science Approaches.- 2.3 Recycling Approaches.- 2.4 Knowledge-Poor Approaches.- 3 SEXTANT.- 3.1 Philosophy.- 3.2 Methodology.- 3.3 Other examples.- 3.4 Discussion.- 4 EVALUATION.- 4.1 Deese Antonyms Discovery.- 4.2 Artificial Synonyms.- 4.3 Gold Standards Evaluations.- 4.4 Webster’s 7th.- 4.5 Syntactic vs. Document Co-occurrence.- 4.6 Summary.- 5 APPLICATIONS.- 5.1 Query Expansion.- 5.2 Thesaurus enrichment.- 5.3 Word Meaning Clustering.- 5.4 Automatic Thesaurus Construction.- 5.5 Discussion and Summary.- 6 CONCLUSION.- 6.1 Summary.- 6.2 Criticisms.- 6.3 Future Directions.- 6.4 Vision.- 1 PREPROCESSORS.- 2 WEBSTER STOPWORD LIST.- 3 SIMILARITY LIST.- 4 SEMANTIC CLUSTERING.- 5 AUTOMATIC THESAURUS GENERATION.- 6 CORPORA TREATED.- 6.1 ADI.- 6.2 AI.- 6.3 AIDS.- 6.4 ANIMALS.- 6.5 BASEBALL.- 6.6 BROWN.- 6.7 CACM.- 6.8 CISI.- 6.9 CRAN.- 6.10 HARVARD.- 6.11 JFK.- 6.12 MED.- 6.13 MERGERS.- 6.14 MOBYDICK.- 6.15 NEJM.- 6.16 NPL.- 6.17 SPORTS.- 6.18 TIME.- 6.19 XRAY.- 6.20 THESIS.