Essay

Explain the gap between theory and practice for significance tests on dev set changes.

Question: In a concise analytical response, explain why the source distinguishes the theoretical availability of statistical significance testing from its practical use by teams.

Sample answer: A team can theoretically test whether an algorithm change produces a statistically significant difference on the dev set. In practice, however, most teams do not perform this test. Publishing an academic research paper is the exception identified by the source. For ordinary interim progress, the author usually does not find statistical significance tests useful, so theoretical availability does not make them a routine development practice.

Key points:

  • A significance test on a dev set change is theoretically possible.
  • Most teams do not perform such tests in practice.
  • Publishing academic research papers is the stated exception.
  • The author usually does not find the tests useful for measuring interim progress.

Rubric: A strong response clearly contrasts theoretical possibility with typical practice, identifies academic research publication as the stated exception, and accurately explains the author's view of interim-progress measurement without adding unsupported claims.

0

1

Updated 2026-07-19

Contributors are:

Who are from:

Tags

Machine Learning

Deep Learning

Machine Learning Strategy

Supervised Learning

Dive into Deep Learning @ D2L

Data Science

Machine Learning Yearning @ DeepLearning.AI