Babe, stop everything! New favorite paper of the year is out! kyutai.org/fid-lottery/ arxiv.org/abs/2606.20536
- Retraining a model with a different seed moves its FID score 3.2x more than resampling from a fixed trained network.
- FID coefficient of variation stays within a 1-2% band even as compute or model size increases.
- The authors recommend treating any FID gap below roughly 1.3% CoV as inconclusive and requiring multi-seed error bars.