Mao Hong, Zhengling Qi, Yanxun Xu: Model-based Reinforcement Learning for Confounded POMDPs. ICML 2024: 18668-18710