Learn Before
Case Study

Select a pipeline component using training-data availability.

Case context: A team is building a multi-step prediction system and is deciding between two modules. It can gather training examples for Module Red with little effort, but collecting training examples for Module Blue would take substantial work.

Question: Using only the stated selection criterion, which module should the team prefer, and what should it conclude about Module Blue?

Sample answer: The team should prefer Module Red because its training data is easy to collect. Under this criterion, Module Blue is the weaker choice because the data needed to train it is hard to obtain.

Key points:

  • Prefer Module Red under the stated criterion.
  • Module Red has easy-to-collect training data.
  • Module Blue is disadvantaged by difficult data collection.
  • The comparison is about a multi-step pipeline, not a fully end-to-end system.

Rubric: The response should select Module Red, state that easy training-data collection is the reason, and describe Module Blue only in terms of its poorer training-data availability.

0

1

Updated 2026-08-12

Contributors are:

Who are from:

Tags

Machine Learning

Deep Learning

Supervised Learning

Dive into Deep Learning @ D2L

Data Science

Machine Learning Strategy

Machine Learning Yearning @ DeepLearning.AI