2 August 2026
Random reshuffling as a baseline policy
In an exact quadratic model of learning, some fair choices inside ordinary data shuffling admit a same-baseline value certificate; popular ordering proxies can fail without bound, while global threshold ordering is NP-complete when the stable step is instance-encoded and approaches 2.
availability and archiving · internal replay · 6 dimensions not assessed