#contributionBudget-matched comparison addresses a real confound
Figures 3 and 4 compare multiple GAN objectives under common search budgets and show how apparent rankings change with available tuning. This directly addresses whether reported gains reflect objectives or optimization effort.
↳ Sections 4 and 6; Figures 3–4
#methodological rigourVariance and uncertainty receive explicit treatment
The study retrains selected configurations across 50 random initializations and uses 5,000 bootstrap resamples for budget curves. This shifts attention from isolated best scores toward performance distributions.
↳ Section 6; Figures 3, 5, and 15; Table 2
#positioningLimitations constrain the final conclusion
Section 7 identifies the shared architecture, dataset complexity, optimization dependence, and FID’s inability to detect memorization. Section 8 correspondingly leaves open the possibility of different results under unexplored conditions.
↳ Sections 7–8