CiNii 図書 - Explorations in automatic thesaurus discovery

著者

- Grefenstette, Gregory

書誌事項

Explorations in automatic thesaurus discovery

Gregory Grefenstette

（The Kluwer international series in engineering and computer science, SECS 278 . Natural language processing and machine translation）

Kluwer Academic Publishers, c1994

大学図書館所蔵件 / 全8件

大分大学学術情報拠点(図書館)

025.49||GG511095440

OPAC
大阪大学附属図書館総合図書館

12200043029

OPAC
関西大学図書館図

205973345

OPAC
京都大学大学院情報学研究科

ZH61294039377

OPAC
岐阜大学図書館

OPAC
国立研究開発法人理化学研究所図書館

541||KLU||278H0215708

OPAC
筑波大学附属図書館中央図書館

007.52-G8310023003700

OPAC
福岡教育大学学術情報センター図書館図

401||G822196007893

OPAC
該当する所蔵館はありません
すべての絞り込み条件を解除する

この図書・雑誌をさがす

注記

Includes bibliographical references (p. 295-302) and index

内容説明・目次

内容説明

Explorations in Automatic Thesaurus Discovery presents an automated method for creating a first-draft thesaurus from raw text. It describes natural processing steps of tokenization, surface syntactic analysis, and syntactic attribute extraction. From these attributes, word and term similarity is calculated and a thesaurus is created showing important common terms and their relation to each other, common verb--noun pairings, common expressions, and word family members. The techniques are tested on twenty different corpora ranging from baseball newsgroups, assassination archives, medical X-ray reports, abstracts on AIDS, to encyclopedia articles on animals, even on the text of the book itself. The corpora range from 40,000 to 6 million characters of text, and results are presented for each in the Appendix. The methods described in the book have undergone extensive evaluation. Their time and space complexity are shown to be modest. The results are shown to converge to a stable state as the corpus grows. The similarities calculated are compared to those produced by psychological testing. A method of evaluation using Artificial Synonyms is tested. Gold Standards evaluation show that techniques significantly outperform non-linguistic-based techniques for the most important words in corpora. Explorations in Automatic Thesaurus Discovery includes applications to the fields of information retrieval using established testbeds, existing thesaural enrichment, semantic analysis. Also included are applications showing how to create, implement, and test a first-draft thesaurus.

Preface. 1. Introduction. 2. Semantic Extraction. 3. Sextant. 4. Evaluation. 5. Applications. 6. Conclusion. 1: Preprocesors. 2. Webster Stopword List. 3: Similarity List. 4: Semantic Clustering. 5: Automatic Thesaurus Generation. 6. Corpora Treated. Index.

「Nielsen BookData」より

Explorations in automatic thesaurus discovery

著者

書誌事項

大学図書館所蔵件 / 全8件

この図書・雑誌をさがす

注記

内容説明・目次

関連文献： 1件中 1-1を表示

詳細情報

書き出し

Explorations in automatic thesaurus discovery

著者

書誌事項

大学図書館所蔵 件 / 全8件

この図書・雑誌をさがす

注記

内容説明・目次

関連文献： 1件中 1-1を表示

詳細情報

書き出し

大学図書館所蔵件 / 全8件