Your Eval Set Is Lying to You
中文摘要
模型在评估集表现良好却在实际应用中失败,说明问题不在于模型,而在于评估集。
English Summary
High evaluation scores don't guarantee real-world performance; the issue often lies with the evaluation set rather than the model itself.
原文节选
A model that scores 94 percent on your eval set and then falls apart on real traffic does not have a model problem. It has an eval set… Continue reading on Python in Plain English »