Measure delivery-time quality — mean, or the tail?
Metric designMediumProblem. You're choosing the metric for a delivery-time experiment. Why is mean delivery time (or mean ETA error) a poor choice, and how do you handle the metric statistically?
Before you reveal: say your answer out loud, as if you were in the real interview — get your reasoning across clearly first. There is no single correct answer: reading what the interviewer is really after and defending your own thinking is what makes an answer strong.