Kniha Explorations in Automatic Thesaurus Discovery Gregory Grefenstette

Explorations in Automatic Thesaurus Discovery

Jazyk: Angličtina
Vazba: Pevná
Vydavatel: Springer
Dostupnost: Skladem u dodavatele v malém množství
Odesíláme za 13-18 dnů
3 937
Explorations in Automatic Thesaurus Discovery presents an automated method for creating a first-draf...

Informace o knize

Jazyk
Angličtina
Vazba
Kniha - Pevná
Vydáno
1994
Stránek
305
EAN
9780792394686
ISBN
0792394682
Enbook ID
01398253
Vydavatel
Hmotnost
1390
Rozměry
155 x 235 x 20

Kompletní popis

Explorations in Automatic Thesaurus Discovery presents an automated method for creating a first-draft thesaurus from raw text. It describes natural processing steps of tokenization, surface syntactic analysis, and syntactic attribute extraction. From these attributes, word and term similarity is calculated and a thesaurus is created showing important common terms and their relation to each other, common verb--noun pairings, common expressions, and word family members. The techniques are tested on twenty different corpora ranging from baseball newsgroups, assassination archives, medical X-ray reports, abstracts on AIDS, to encyclopedia articles on animals, even on the text of the book itself. The corpora range from 40,000 to 6 million characters of text, and results are presented for each in the Appendix. The methods described in the book have undergone extensive evaluation. Their time and space complexity are shown to be modest. The results are shown to converge to a stable state as the corpus grows. The similarities calculated are compared to those produced by psychological testing. A method of evaluation using Artificial Synonyms is tested. Gold Standards evaluation show that techniques significantly outperform non-linguistic-based techniques for the most important words in corpora. Explorations in Automatic Thesaurus Discovery includes applications to the fields of information retrieval using established testbeds, existing thesaural enrichment, semantic analysis. Also included are applications showing how to create, implement, and test a first-draft thesaurus.

Mohlo by vás zajímat

4 690
875
2 567

Gender Transgressions

Karen J. Taylor
1 688
170

Landscape Confection

Helen Molesworth
599

The HUUT

Mike Kennedy
331

Writer of the World

Daniel E C Rumbell
331
351
412

Armance

Stendhal
214

Grand Strategy

Peter Layton
391
597

Nature, Volume 26

Um-Medsearch Gateway
706

Neon Darkness

Lauren Shippen
279

Mexico Reader

Timothy J. Henderson
573

Zákaznicí kteří koupili tuto knihu koupili také

Indian Summer

Adalbert Stifter
999
474
1 498
1 495

hans pöbel.

Alexander Hägele
202

Hunter X Hunter 2

Yoshihiro Togashi
175
270
259