Sciweavers

Free Online Productivity Tools i2Speak i2Symbol i2OCR iTex2Img iWeb2Print iWeb2Shot i2Type iPdf2Split iPdf2Merge i2Bopomofo i2Arabic i2Style i2Image i2PDF iLatex2Rtf Sci2ools

37

NLPRS
2001
Springer

favoriteEmaildiscussreport

124views Natural Language Processing» more NLPRS 2001»

Statistical Parsing of Dutch using Maximum Entropy Models with Feature Merging

14 years 1 months ago

Statistical Parsing of Dutch using Maximum Entropy Models with Feature Merging

Download odur.let.rug.nl

In this project report we describe work in statistical parsing using the maximum entropy technique and the Alpino language analysis system for Dutch. A major difﬁculty in this domain is the lack of sufﬁcient corpus data available for training. Among other problems, this sparseness of data increases the danger of the model overﬁtting the training data, making it particularly important that the selection of statistical features upon which to base the model be optimal. To this end we have adapted the notion of feature merging, a means of constructing equivalence classes of statistical features based upon common elements within them. In spite of promising preliminary results, subsequent tests have not enabled us to conclude whether this approach helps the kind of models we are working with.

Tony Mullen, Rob Malouf, Gertjan van Noord

Real-time Traffic

Alpino Language Analysis | Maximum Entropy Technique | Natural Language Processing | NLPRS 2001 | Statistical Features |

claim paper

Related Content

» Semantic and Syntactic Features for Dutch Coreference Resolution

» A Hybrid Japanese Parser with Handcrafted Grammar and Statistics

» Feature Lattices for Maximum Entropy Modelling

» A hybrid approach to NER by MEMM and manual rules

» MARS A Statistical Semantic Parsing and GenerationBased Multilingual Automatic tRanslation...

» Using Maximum Entropy for Automatic Image Annotation

» A maximum entropy web recommendation system combining collaborative and content features

» Domain action classification using a maximum entropy model in a schedule management domain

» A Progressive Feature Selection Algorithm for Ultra Large Feature Spaces

Post Info
More Details (n/a)

Added	30 Jul 2010
Updated	30 Jul 2010
Type	Conference
Year	2001
Where	NLPRS
Authors	Tony Mullen, Rob Malouf, Gertjan van Noord

Comments (0)