Estimating a Benchmark for Ticket Routing
Scenario: A product team is building a model that assigns incoming support tickets to one of several specialist queues. As part of project planning, the team tries to use experienced agents as a reference point, but the agents often disagree about which queue a ticket should go to.
Question: What does this suggest about estimating the best achievable error rate for the system?
Sample answer: It suggests that estimating the best possible error rate will be difficult, because people do not reliably agree on the correct label for this task.
Key points:
- The task is ambiguous enough that trained people disagree.
- Without a dependable human reference, the ceiling on performance is hard to pin down.
- Therefore, the optimal error rate is difficult to estimate.
Rubric: The response must state that strong disagreement among human judges makes the best achievable error rate hard to estimate.
0
1
Tags
Machine Learning
Deep Learning
Supervised Learning
Dive into Deep Learning @ D2L
Data Science
Machine Learning Strategy
Machine Learning Yearning @ DeepLearning.AI
Related
Why Human Baselines Are Less Helpful on Hard-to-Judge Tasks
Which situation makes it hard to estimate the best achievable error rate for a task?
Recommending products to shoppers is an example of a task that people may find hard to evaluate well.
When human performance is poor, estimating the _____ error rate becomes difficult.
Match each concept to the description that best fits how its error rate can be estimated.
Order the steps for judging whether expert performance can set a target error rate.
Which task is given as an example of a problem that is hard for people to judge reliably?
Human performance is always a dependable estimate of the best achievable error rate, even when people find the task difficult.
Predicting Which _____ to Show a Shopper Can Be Hard for People Too
Match Each Situation to Its Effect on Estimating the Best Possible Error
Order the reasoning steps for a task that is difficult even for people
Why Human Performance May Not Give a Useful Error Target
Estimating a Benchmark for Ticket Routing
Examples where human judgment is a poor guide to optimal error