intTypePromotion=1
ADSENSE

Báo cáo khoa học: "Efficient Unsupervised Discovery of Word Categories Using Symmetric Patterns and High Frequency Words"

Chia sẻ: Hongvang_1 Hongvang_1 | Ngày: | Loại File: PDF | Số trang:8

40
lượt xem
1
download
 
  Download Vui lòng tải xuống để xem tài liệu đầy đủ

We present a novel approach for discovering word categories, sets of words sharing a significant aspect of their meaning. We utilize meta-patterns of highfrequency words and content words in order to discover pattern candidates. Symmetric patterns are then identified using graph-based measures, and word categories are created based on graph clique sets. Our method is the first pattern-based method that requires no corpus annotation or manually provided seed patterns or words. We evaluate our algorithm on very large corpora in two languages, using both human judgments and WordNetbased evaluation. ...

Chủ đề:
Lưu

Nội dung Text: Báo cáo khoa học: "Efficient Unsupervised Discovery of Word Categories Using Symmetric Patterns and High Frequency Words"

ADSENSE

CÓ THỂ BẠN MUỐN DOWNLOAD


intNumView=40

 

Đồng bộ tài khoản
2=>2