EconPapers    
Economics at your fingertips  
 

Nonparametric Bayesian Policy Learning

Haonan Ye

Papers from arXiv.org

Abstract: I propose Nonparametric Bayesian Policy Learning (NBPL) as a framework for uncertainty-aware treatment choice. The key observation is that welfare is fully determined by a reduced-form distribution, so uncertainty about optimal policies entirely reflects uncertainty about this distribution. NBPL places a Dirichlet process prior on the reduced-form distribution and uses the resulting posterior for both traditional policy choice and inference on optimal welfare and optimal treatment assignments. NBPL is computationally tractable: the default Bayesian-bootstrap implementation requires only exponential reweighting of the observations. I establish two theoretical properties. First, posterior welfare regret converges at the minimax-optimal rate, providing a novel policy-relevant analogue of posterior contraction rates. Second, posterior model selection across policy classes is consistent. I relate NBPL to existing policy learning approaches and illustrate using two empirical applications: the JTPA experiment and the bednet subsidy experiment. In both applications, decision-tree rules tend to yield higher optimal welfare than linear rules.

Date: 2026-05, Revised 2026-09
New Economics Papers: this item is included in nep-ecm
References: View references in EconPapers View complete reference list from CitEc
Citations:

Downloads: (external link)
https://arxiv.org/pdf/2605.17068 Latest version (application/pdf)

Related works:
This item may be available elsewhere in EconPapers: Search for items with the same title.

Export reference: BibTeX RIS (EndNote, ProCite, RefMan) HTML/Text

Persistent link: https://EconPapers.repec.org/RePEc:arx:papers:2605.17068

Access Statistics for this paper

More papers in Papers from arXiv.org
Bibliographic data for series maintained by arXiv administrators ().

 
Page updated 2026-09-18
Handle: RePEc:arx:papers:2605.17068