Off-Policy Confidence Interval Estimation with Confounded Markov Decision Process. (arXiv:2202.10589v3 [stat.ML] UPDATED) | allainews.com

June 24, 2022, 1:11 a.m. | Chengchun Shi, Jin Zhu, Ye Shen, Shikai Luo, Hongtu Zhu, Rui Song

stat.ML updates on arXiv.org arxiv.org

This paper is concerned with constructing a confidence interval for a target
policy's value offline based on a pre-collected observational data in infinite
horizon settings. Most of the existing works assume no unmeasured variables
exist that confound the observed actions. This assumption, however, is likely
to be violated in real applications such as healthcare and technological
industries. In this paper, we show that with some auxiliary variables that
mediate the effect of actions on the system dynamics, the target policy's …

arxiv confidence decision interval markov ml policy process

More from arxiv.org / stat.ML updates on arXiv.org

Simulation-Based Prior Knowledge Elicitation for Parametric Bayesian Models 9 hours ago | arxiv.org

abstract arxiv bayesian domain +17

One-step corrected projected stochastic gradient descent for statistical estimation 9 hours ago | arxiv.org

abstract algorithm arxiv fisher +13

The boosted HP filter is more general than you might think 9 hours ago | arxiv.org

abstract arxiv boosting computational +21

High Probability Bounds for Stochastic Subgradient Schemes with Heavy Tailed Noise 9 hours ago | arxiv.org

abstract arxiv distribution math.oc +9

A Unified Combination Framework for Dependent Tests with Applications to Microbiome Association Studies 9 hours ago | arxiv.org

abstract analysis applications arxiv +17

Causal Inference for Genomic Data with Multiple Heterogeneous Outcomes 9 hours ago | arxiv.org

abstract arxiv become causal +19

Random walks on simplicial complexes 9 hours ago | arxiv.org

abstract arxiv generalized graph +8

Large Language Model for Causal Decision Making 1 day, 9 hours ago | arxiv.org

abstract arxiv capability causal +30

TaCo: Targeted Concept Removal in Output Embeddings for NLP via Information Theory and Explainability 1 day, 9 hours ago | arxiv.org

arxiv concept cs.cl embeddings +7

Data Scientist (m/f/x/d)

@ Symanto Research GmbH & Co. KG | Spain, Germany

View on ai-jobs.net

Sr. Data Science Consultant

@ Blue Yonder | Bengaluru

View on ai-jobs.net

Artificial Intelligence Developer

@ HP | PSR01 - Bengaluru, Pritech Park- SEZ (PSR01)

View on ai-jobs.net

Senior Software Engineer - Cloud Data Extraction

@ Celonis | Munich, Germany

View on ai-jobs.net

Finance Master Data Management

@ Airbus | Lisbon (Airbus Portugal)

View on ai-jobs.net

Imaging Support Associate

@ Lexington Medical Center | West Columbia, SC, US, 29169

View on ai-jobs.net