Sciweavers

Free Online Productivity Tools i2Speak i2Symbol i2OCR iTex2Img iWeb2Print iWeb2Shot i2Type iPdf2Split iPdf2Merge i2Bopomofo i2Arabic i2Style i2Image i2PDF iLatex2Rtf Sci2ools

9

ACL
2006

favoriteEmaildiscussreport

103views Computational Linguistics» more ACL 2006»

Discriminative Pruning of Language Models for Chinese Word Segmentation

13 years 5 months ago

Discriminative Pruning of Language Models for Chinese Word Segmentation

Download acl.ldc.upenn.edu

This paper presents a discriminative pruning method of n-gram language model for Chinese word segmentation. To reduce the size of the language model that is used in a Chinese word segmentation system, importance of each bigram is computed in terms of discriminative pruning criterion that is related to the performance loss caused by pruning the bigram. Then we propose a step-by-step growing algorithm to build the language model of desired size. Experimental results show that the discriminative pruning method leads to a much smaller model compared with the model pruned using the state-of-the-art method. At the same Chinese word segmentation F-measure, the number of bigrams in the model can be reduced by up to 90%. Correlation between language model perplexity and word segmentation performance is also discussed.

Jianfeng Li, Haifeng Wang, Dengjun Ren, Guohua Li

Real-time Traffic

ACL 2006 | ACL 2007 | Chinese Word Segmentation | Discriminative Pruning | Language Model |

claim paper

Related Content

» A Chunking Strategy Towards Unknown Word Detection in Chinese Word Segmentation

» Integrating Language Model in Handwritten Chinese Text Recognition

» A Fast Decoder for Joint Word Segmentation and POSTagging Using a Single Discriminative Mo...

» An ErrorDriven WordCharacter Hybrid Model for Joint Chinese Word Segmentation and POS Tagg...

» Improved SourceChannel Models for Chinese Word Segmentation

» A Simple and Efficient Model Pruning Method for Conditional Random Fields

» Adapting Chinese Word Segmentation for Machine Translation Based on Short Units

» SelfSupervised Chinese Word Segmentation

» Chinese Segmentation with a WordBased Perceptron Algorithm

Post Info
More Details (n/a)

Added	30 Oct 2010
Updated	30 Oct 2010
Type	Conference
Year	2006
Where	ACL
Authors	Jianfeng Li, Haifeng Wang, Dengjun Ren, Guohua Li

Comments (0)