Example

Interpreting 10% Training, 11% Training-Dev, and 12% Development Error

A model with 10% training error, 11% training-dev error, and 12% development error has only a one-percentage-point increase at each comparison. The small training-to-training-dev gap suggests low variance on the training distribution, while the small training-dev-to-development gap suggests little data mismatch. Whether the 10% training error represents high avoidable bias depends on the best achievable, Bayes, or human-level error: high avoidable bias is supported only if that benchmark is substantially lower than 10%.

0

1

Updated 2026-08-30

Contributors are:

Who are from:

Tags

Machine Learning

Deep Learning

Supervised Learning

Dive into Deep Learning @ D2L

Data Science

Machine Learning Strategy

Machine Learning Yearning @ DeepLearning.AI

Related