Robust Counterfactual Policy Optimisation via Nondeterministic Causal Models
Mirrored from arXiv — Machine Learning for archival readability. Support the source by reading on the original site.
Computer Science > Machine Learning
Title:Robust Counterfactual Policy Optimisation via Nondeterministic Causal Models
Abstract:Counterfactual inference approaches for sequential decision-making typically assume deterministic causal models, where all randomness stems from latent variables. However, Markov Decision Processes (MDPs) are inherently stochastic. We address this by formalising counterfactual policy optimisation under probabilistic nondeterministic causal models, which properly separates latent confounding from irreducible stochasticity, and here propose a first practical optimisation problem for identifying robust counterfactual policies under a sensitivity analysis framework. We validate our approach on a sepsis treatment simulator, where diabetes status acts as a hidden global confounder.
| Comments: | Accepted at UAI 2026 Workshop on Causality for Decision Making |
| Subjects: | Machine Learning (cs.LG); Artificial Intelligence (cs.AI) |
| Cite as: | arXiv:2608.02893 [cs.LG] |
| (or arXiv:2608.02893v1 [cs.LG] for this version) | |
| https://doi.org/10.48550/arXiv.2608.02893
arXiv-issued DOI via DataCite (pending registration)
|
Access Paper:
- View PDF
- HTML (experimental)
- TeX Source
References & Citations
Bibliographic and Citation Tools
Code, Data and Media Associated with this Article
Demos
Recommenders and Search Tools
arXivLabs: experimental projects with community collaborators
arXivLabs is a framework that allows collaborators to develop and share new arXiv features directly on our website.
Both individuals and organizations that work with arXivLabs have embraced and accepted our values of openness, community, excellence, and user data privacy. arXiv is committed to these values and only works with partners that adhere to them.
Have an idea for a project that will add value for arXiv's community? Learn more about arXivLabs.
More from arXiv — Machine Learning
-
Transformer Geometry Observatory TGO-IV: Developmental Topology Observatory
Aug 12
-
Uncertainty-Aware Ensemble Deep Randomized Neural Networks for Classification
Aug 12
-
CurveFP: Rational-Radix Logarithmic Datatypes with Closed Products for Language Models
Aug 12
-
Sheaf-Based Federated Representation Learning
Aug 12
Discussion (0)
Sign in to join the discussion. Free account, 30 seconds — email code or GitHub.
Sign in →No comments yet. Sign in and be the first to say something.