Zellner-Tobias-Ryu (1999) Bayesian Method of Moments Analysis of Parametric and Semiparametric Regression Models

bayesianbmomsemiparametric-regressionmoment-conditionsentropy-maximizationmodel-selectionseries-expansionloss-functionces-production-functionregression

Summary

Extends the Bayesian Method of Moments (BMOM) framework to semiparametric regression, where the regression function f(x)f(x) is unknown and approximated by a series expansion (Polynomial, Fourier, or Gallant flexible functional form). The key BMOM idea is to construct a post-data density for parameters by maximizing entropy subject to sample moment constraints, entirely avoiding specification of a likelihood function. Three model-selection tools are developed for choosing the expansion type and number of terms: estimation loss, predictive loss, and BMOM posterior odds. A Monte Carlo experiment with data generated from a constant elasticity of substitution (CES) production function illustrates the methods.

Key Claims

Concepts Introduced or Extended

Entities Mentioned

Quotes

"One of the distinguishing features of the BMOM approach is that it yields post-data densities for models' parameters without use of an assumed likelihood function."

"BMOM gives good answers for the questions it addresses while not purporting to go beyond the information that is really there in the prior and the data." — Laskey (1997), commenting on BMOM

My Take

This paper is primarily a methodological demonstration rather than an econometric application paper. The core BMOM result for regression (MaxEnt posterior = diffuse-prior Bayesian posterior) is elegant but also somewhat deflating: you end up at the same place as standard Bayesian regression, just by a different path. The value is clearest when the likelihood is genuinely unknown — e.g. in semiparametric settings or when error distributions are highly uncertain. The three model-selection tools (estimation loss, predictive loss, posterior odds) are well-motivated and practically useful for series-expansion truncation. The Monte Carlo results are somewhat limited — a single data generating process (DGP; CES) under a single sample size scenario — but illustrate the approach clearly. Connection to the KL divergence in posterior odds is the most durable theoretical insight.