Robustness under distribution shift
If I train on one device today, why does it fail on another device tomorrow?
Hardware differences and temporal drift quietly break the i.i.d. assumption almost every architecture rests on. I cancel that divergence at the signal level rather than papering over it with more parameters.


