[D] Why are ML model outputs not tested regarding statistical significance?

Tigmib@alien.top · 1 year ago

[D] Why are ML model outputs not tested regarding statistical significance?

Recent_Ad4998@alien.top · 1 year ago

One thing I have found is that if you have a large dataset, the standard error can become so small that any difference in average performance will be significant. Obviously not always the case, depending on the size of the variance etc but I imagine it might be why it’s often acceptable not to include them.

iswedlvera@alien.top · 1 year ago

This is the reason. People do significance tests when you want to draw conclusions with 20 samples on an entire population. If you have thousands of samples there won’t be much point.

econ1mods1are1cucks@alien.top · 1 year ago

Depends on how big the individual samples are tbh. 1000 samples of 10 people actually sounds like a decent study group

iswedlvera@alien.top · 1 year ago

I see what you mean. Yeah it shouldn’t be by default I don’t do statistical significance tests.