Essay

How human comparison can still help a stronger-than-average system

Question: A speech recognition model now performs better than the average transcriber on the full development and test sets. Explain how human comparison can still be useful for improving it. State the condition that must hold and describe the benefits human judgment provides.

Sample answer: Human comparison can still be valuable if there is a slice of data where people still do better than the model. On that slice, human labels can be higher quality than the model's predictions, and the human decisions can help the team understand what the system is missing. Those human results also give a realistic target for that difficult slice, even if the overall system is already stronger than the average person.

Key points:

  • There must be a subset where humans outperform the model.
  • Human labels on that subset can be more accurate.
  • Human judgment can explain why the model failed on those cases.
  • Human performance on the subset provides a goal for improvement.

Rubric: Full credit requires stating that a human-better subset must exist, and naming all three benefits: better labels, useful intuition about failures, and a performance target for that subset.

0

1

Updated 2026-08-12

Contributors are:

Who are from:

Tags

Machine Learning

Deep Learning

Supervised Learning

Dive into Deep Learning @ D2L

Data Science

Machine Learning Strategy

Machine Learning Yearning @ DeepLearning.AI