Material Detail

Grammatical Inference as a Principal Component Analysis Problem

This video was recorded at 26th International Conference on Machine Learning (ICML), Montreal 2009. One of the main problems in probabilistic grammatical inference consists in inferring a stochastic language, i.e. a probability distribu- tion, in some class of probabilistic models, from a sample of words independently drawn according to a ﬁxed unknown target distribution p. Here we consider the class of rational stochastic languages composed of stochastic languages that can be computed by muliplicity automata, which can be viewed as a generalization of probabilistic automata. Rational stochastic languages p have a useful algebraic characterization: all the mappings up:v-¿p(uv) lie in a ﬁnite dimensional vector subspace Vp of the vector space R(E) composed of all real-valued functions deﬁned over E. Hence, a ﬁrst step in the grammatial inference process can consist in identifying the subspace Vp. In this paper, we study the possibility of using principal component analysis to achieve this task. We provide an inference algorithm which computes an estimate of the target distribution. We prove sometheoreticalpropertiesofthisalgorithmandweprovideresultsfromnumericalsimulationsthatconﬁrm the relevance of our approach.

Keywords:: videolectures, ocwc, oec

Disciplines:

Science and Technology / Computer Science / Programming & Programming Languages

Go to Material

Bookmark / Add to Course ePortfolio

Create a Learning Exercise

Add Accessibility Information

Rate

Add a Comment

Quality

User Rating
Comments
Learning Exercises
Bookmark Collections
Course ePortfolios
Accessibility Info

Report Broken Link
Report as Inappropriate

More about this material

Material Type:: Presentation
Date Added to MERLOT:: February 8, 2015
Date Modified in MERLOT:: February 8, 2015
Author:: Raphaël Bailly, Centre de Mathématiques et Informatique, Aix-Marseille Université
Submitter:: The Open Education Consortium
Primary Audience:: College General Ed, College Lower Division, College Upper Division
Technical Format:: Video

Mobile Compatibility:: Not specified at this time
Language:: English
Cost Involved:: No
Source Code Available:: No
Creative Commons:: This work is licensed under a Attribution-NonCommercial-NoDerivs 3.0 United States