mpo maxWe introduce a new algorithm for reinforcement learning called Maximum aposteriori Policy Optimisation (MPO) based on coordinate ascent on a relative entropyMPO has an independent prognostic value overall and most notably in patients tested negative with a hher sensitive cardiac troponin I assay.