Comments on natural language processing blog: Non-parametric versus model selection/averaging

酒店經紀PRETTY GIRL 台北酒店經紀人 ,禮服店酒店兼差PRETTY GIRL酒店公關酒...

2009-05-12T10:40:00.000-06:00

酒店經紀PRETTY GIRL 台北酒店經紀人 ,禮服店酒店兼差PRETTY GIRL酒店公關酒店小姐彩色爆米花酒店兼職,酒店工作彩色爆米花酒店經紀, 酒店上班,酒店工作 PRETTY GIRL酒店喝酒酒店上班彩色爆米花台北酒店酒店小姐 PRETTY GIRL酒店上班酒店打工PRETTY GIRL酒店打工酒店經紀彩色爆米花

i agree with mark on both counts. it's true that ...

2007-10-25T21:07:00.000-06:00

i agree with mark on both counts.

it's true that nonparametric bayesian models can often be approximated with finite parametric models. the Dirichlet process allows two approximations---one through a finite stick-breaking model and another through a symmetric Dirichlet.

still, even if the finite approximation of the nonparametric model does the trick, it's nice to know what is being approximated. (and this is particularly useful for setting and reasoning about hyperparameters.)

mark's second point is compelling. the promise of NPB models is in moving beyond simply choosing a number of components. NPB models that generate structures, like grammars or trees, allow us to posit complicated combinatorial objects as latent random variables and still hope to infer them from data.

the naysayer might say: this is simply search with an objective function that is the posterior. yes, this is true, at least when you only care about a MAP estimate. but, the NPB posterior gives a nicely regularized objective function trading off what the data imply and a prior preference for simpler (or more complicated) structures.

This is an interesting issue! In fact, a Bayesian...

2007-10-25T18:36:00.000-06:00

This is an interesting issue! In fact, a Bayesian estimator for a parametric model may select a model that only uses a subset of the possible states, particularly if you have a sparse Bayesian prior. Indeed, one way of estimating a non-parametric model is to fit a corresponding parametric model with a state space sufficiently large that not all states will be occupied (or only occupied with very low probability).

I think this makes it clear that non-parametric models aren't necessarily that different to parametric ones.

The great hope (and at this stage I think that's all it is) for non-parametric models is that it will let us formulate and explore models of greater complexity than we could deal with parametrically.

If you'll excuse me patting my own back, I think that the adaptor grammars we presented at NIPS last year are an example of something that would be hard to formulate parametrically. At a very high level, adaptor grammars are an extension of PCFGs that permit an infinite number of possible rules. The possible rules are combinations of other useful rules, and so on recursively. So adaptor grammars are a single framework that integrates the two phases of standard generate-and-prune grammar learning systems (in which a rule-proposal phase is followed by a rule-probability estimation phase that prunes the useless rules).

Comments on natural language processing blog: Non-parametric versus model selection/averaging

酒店經紀PRETTY GIRL 台北酒店經紀人 ,禮服店 酒店兼差PRETTY GIRL酒店公關 酒...

i agree with mark on both counts. it's true that ...

This is an interesting issue! In fact, a Bayesian...

酒店經紀PRETTY GIRL 台北酒店經紀人 ,禮服店酒店兼差PRETTY GIRL酒店公關酒...