The framing of a shared benchmark is really Interesting. However, curious how you're validating that the synthetic simulator's arrival/LOS distributions are realistic enough that a model doing well here actually transfers to real hospital data — is there a planned calibration step against de-identified real occupancy curves, even just for the simulator's parameters