Browse Definitions :
Definition

overfitting

Contributor(s): Matthew Haughn

Overfitting is the incorrect optimizing of an artificial intelligence (AI) model, where the seeking of accuracy goes too far and may result in false positives.

Overfitting contrasts with underfitting, which can also result in inaccuracies. Overfitting is often referred to as overtraining and underfitting as undertraining. Overfitting and underfitting both ruin the accuracy of a model by leading to trend observations and predictions that don’t follow the reality of the data.

False positives from overfitting can cause problems with the predictions and assertions made by AI. Underfitting, on the other hand, can miss data that should be included due to omissions resulting from an over-specific model. In unseen data, an overfit model will make errors reflecting those in its training data. This inaccuracy is often a result of the model beginning to try to memorize results instead of accurately predicting previously unseen data.

For example, an AI hunting is for the number 1 in handwritten data. Depending on the clarity of the handwriting, a false positive might be grouping some 7s as 1s, which would be an overfit. Conversely, an omission that might result could be the failure to recognize 1s in some styles of handwriting, which would be an underfit.

Overfitting can be the result of overtraining, a lack of validation, improper validation or adjustment of weights and attempts at optimization after final testing. Overfitting can also be the result of using training data that is too noisy, containing unsuited information.

This was last updated in April 2018

Continue Reading About overfitting

Start the conversation

Send me notifications when other members comment.

Please create a username to comment.

-ADS BY GOOGLE

File Extensions and File Formats

Powered by:

SearchCompliance

  • regulatory compliance

    Regulatory compliance is an organization's adherence to laws, regulations, guidelines and specifications relevant to its business...

  • Whistleblower Protection Act

    The Whistleblower Protection Act of 1989 is a law that protects federal government employees in the United States from ...

  • smart contract

    A smart contract, also known as a cryptocontract, is a computer program that directly controls the transfer of digital currencies...

SearchSecurity

  • RSA algorithm (Rivest-Shamir-Adleman)

    The RSA algorithm is the basis of a cryptosystem -- a suite of cryptographic algorithms that are used for specific security ...

  • remote access

    Remote access is the ability to access a computer or a network remotely through a network connection.

  • IP Spoofing

    IP spoofing is the crafting of Internet Protocol (IP) packets with a source IP address that has been modified to impersonate ...

SearchHealthIT

SearchDisasterRecovery

  • virtual disaster recovery

    Virtual disaster recovery is a type of DR that typically involves replication and allows a user to fail over to virtualized ...

  • tabletop exercise (TTX)

    A tabletop exercise (TTX) is a disaster preparedness activity that takes participants through the process of dealing with a ...

  • risk mitigation

    Risk mitigation is a strategy to prepare for and lessen the effects of threats faced by a data center.

SearchStorage

  • disk array

    A disk array, also called a storage array, is a data storage system used for block-based storage, file-based storage or object ...

  • enterprise storage

    Enterprise storage is a centralized repository for business information that provides common data management, protection and data...

  • optical storage

    Optical storage is any storage type in which data is written and read with a laser. Typically, data is written to optical media, ...

Close