Browse Definitions :
Definition

overfitting

Contributor(s): Matthew Haughn

Overfitting is the incorrect optimizing of an artificial intelligence (AI) model, where the seeking of accuracy goes too far and may result in false positives.

Overfitting contrasts with underfitting, which can also result in inaccuracies. Overfitting is often referred to as overtraining and underfitting as undertraining. Overfitting and underfitting both ruin the accuracy of a model by leading to trend observations and predictions that don’t follow the reality of the data.

False positives from overfitting can cause problems with the predictions and assertions made by AI. Underfitting, on the other hand, can miss data that should be included due to omissions resulting from an over-specific model. In unseen data, an overfit model will make errors reflecting those in its training data. This inaccuracy is often a result of the model beginning to try to memorize results instead of accurately predicting previously unseen data.

For example, an AI hunting is for the number 1 in handwritten data. Depending on the clarity of the handwriting, a false positive might be grouping some 7s as 1s, which would be an overfit. Conversely, an omission that might result could be the failure to recognize 1s in some styles of handwriting, which would be an underfit.

Overfitting can be the result of overtraining, a lack of validation, improper validation or adjustment of weights and attempts at optimization after final testing. Overfitting can also be the result of using training data that is too noisy, containing unsuited information.

This was last updated in April 2018

Continue Reading About overfitting

SearchCompliance

SearchSecurity

  • cyber attack

    A cyber attack is any attempt to gain unauthorized access to a computer, computing system or computer network with the intent to ...

  • backdoor (computing)

    A backdoor is a means to access a computer system or encrypted data that bypasses the system's customary security mechanisms.

  • post-quantum cryptography

    Post-quantum cryptography, also called quantum encryption, is the development of cryptographic systems for classical computers ...

SearchHealthIT

SearchDisasterRecovery

  • risk mitigation

    Risk mitigation is a strategy to prepare for and lessen the effects of threats faced by a business.

  • call tree

    A call tree is a layered hierarchical communication model that is used to notify specific individuals of an event and coordinate ...

  • Disaster Recovery as a Service (DRaaS)

    Disaster recovery as a service (DRaaS) is the replication and hosting of physical or virtual servers by a third party to provide ...

SearchStorage

  • cloud SLA (cloud service-level agreement)

    A cloud SLA (cloud service-level agreement) is an agreement between a cloud service provider and a customer that ensures a ...

  • NOR flash memory

    NOR flash memory is one of two types of non-volatile storage technologies.

  • RAM (Random Access Memory)

    RAM (Random Access Memory) is the hardware in a computing device where the operating system (OS), application programs and data ...

Close