Sequential multi-stage intervention policies
Extend Causal Bayesian Optimization to sequential multi-stage intervention policies in which each intervention depends on outcomes observed at previous stages.
References
The main limitation relative to RL is the assumption of a single-step (or few-step) decision problem; extending CBO to sequential multi-stage intervention policies remains an open challenge.
— Causal Bayesian Optimization: Foundations, Methods, and Applications
(2609.24112 - Huang et al., 21 Sep 2026) in Section 5, paragraph “Policy search and reinforcement learning”