[D] Why are ML model outputs not tested regarding statistical significance?

Tigmib@alien.top · 1 year ago

[D] Why are ML model outputs not tested regarding statistical significance?

SciGuy42@alien.top · 1 year ago

I review for AAAI, NeuRIPS, etc. If a paper doesn’t report some notion of variance, standard deviation, etc., I have no choice but to reject since it’s impossible to tell whether the proposed approach is actually better. In the rebuttals, the author’s response is typically “well, everyone else also does it this way”. Ideally, I’d like to see an actual test of statistical significance.

iswedlvera@alien.top · 1 year ago

I think op is refering to hypothesis tests between baseline. What’s the point in reporting variance and standard deviation? My outputs on regression tasks are always non-normal. I tend to always plot the cumulative frequency but assigning a number to the distribution such as the variance will have very little meaning.