Back to Home
AI on Medium··Industry Media

Your Eval Set Is Lying to You

中文摘要

模型在评估集表现良好却在实际应用中失败,说明问题不在于模型,而在于评估集。

English Summary

High evaluation scores don't guarantee real-world performance; the issue often lies with the evaluation set rather than the model itself.

Original Excerpt

A model that scores 94 percent on your eval set and then falls apart on real traffic does not have a model problem. It has an eval set… Continue reading on Python in Plain English »