AI Glossary

Overfitting

Overfitting happens when a model learns training-specific patterns, including noise, that do not carry over well to new examples. Training performance can keep improving while performance on held-out data stalls or worsens.

· Updated · Chain of Thought

Model TrainingAI Evaluation & Reliability

A ticket classifier might learn peculiar wording used by the few agents who labeled its training set. It handles those tickets well but fails on new customers’ wording. This is an illustrative failure, not a measured benchmark: the model has fit details that do not help on the intended task.

Compare training results with validation results on examples that did not fit the model. A falling training loss alongside a rising validation loss is a warning. A gap alone is not proof of its cause: also check for different distributions, label problems or leaked information.

Regularization, less flexible models, representative training examples and early stopping can help. More training is not always the remedy. Nor does success on one test set prove success under changed conditions. The practical goal is generalization, not a perfect score on examples the model has already seen.

Training improves. New examples can get worse.Illustrative loss curves over training: training loss keeps falling, while validation loss falls and then rises. The point near the lowest validation loss is a candidate selected during development; a separate test assesses the final choice. This is a schematic, not measured results. Training improves. New examples can get worse. Overfitting: fitting training details that do not carry overLossMore training →Validation lossTraining lossValidation guides choices; a separate test assesses the result.Illustration only. Lower loss means better fit to that set.
A schematic of the loss pattern explained in Google’s overfitting course, not measured results. Check possible data differences as well as model fit. Download the image

Sources

  • Google: Overfitting — Explains training-specific fit and diverging training and validation loss curves.

Go deeper

From the conversation