Browse Definitions :
Definition

overfitting

Contributor(s): Matthew Haughn

Overfitting is the incorrect optimizing of an artificial intelligence (AI) model, where the seeking of accuracy goes too far and may result in false positives.

Overfitting contrasts with underfitting, which can also result in inaccuracies. Overfitting is often referred to as overtraining and underfitting as undertraining. Overfitting and underfitting both ruin the accuracy of a model by leading to trend observations and predictions that don’t follow the reality of the data.

False positives from overfitting can cause problems with the predictions and assertions made by AI. Underfitting, on the other hand, can miss data that should be included due to omissions resulting from an over-specific model. In unseen data, an overfit model will make errors reflecting those in its training data. This inaccuracy is often a result of the model beginning to try to memorize results instead of accurately predicting previously unseen data.

For example, an AI hunting is for the number 1 in handwritten data. Depending on the clarity of the handwriting, a false positive might be grouping some 7s as 1s, which would be an overfit. Conversely, an omission that might result could be the failure to recognize 1s in some styles of handwriting, which would be an underfit.

Overfitting can be the result of overtraining, a lack of validation, improper validation or adjustment of weights and attempts at optimization after final testing. Overfitting can also be the result of using training data that is too noisy, containing unsuited information.

This was last updated in April 2018

Continue Reading About overfitting

Start the conversation

Send me notifications when other members comment.

Please create a username to comment.

-ADS BY GOOGLE

File Extensions and File Formats

SearchCompliance

  • risk management

    Risk management is the process of identifying, assessing and controlling threats to an organization's capital and earnings.

  • compliance as a service (CaaS)

    Compliance as a Service (CaaS) is a cloud service service level agreement (SLA) that specified how a managed service provider (...

  • data protection impact assessment (DPIA)

    A data protection impact assessment (DPIA) is a process designed to help organizations determine how data processing systems, ...

SearchSecurity

  • spyware

    Spyware is a type of malicious software -- or malware -- that is installed on a computing device without the end user's knowledge.

  • application whitelisting

    Application whitelisting is the practice of specifying an index of approved software applications or executable files that are ...

  • botnet

    A botnet is a collection of internet-connected devices, which may include PCs, servers, mobile devices and internet of things ...

SearchHealthIT

SearchDisasterRecovery

  • business continuity plan (BCP)

    A business continuity plan (BCP) is a document that consists of the critical information an organization needs to continue ...

  • disaster recovery team

    A disaster recovery team is a group of individuals focused on planning, implementing, maintaining, auditing and testing an ...

  • cloud insurance

    Cloud insurance is any type of financial or data protection obtained by a cloud service provider. 

SearchStorage

  • DRAM (dynamic random access memory)

    Dynamic random access memory (DRAM) is a type of semiconductor memory that is typically used for the data or program code needed ...

  • RAID 10 (RAID 1+0)

    RAID 10, also known as RAID 1+0, is a RAID configuration that combines disk mirroring and disk striping to protect data.

  • PCIe SSD (PCIe solid-state drive)

    A PCIe SSD (PCIe solid-state drive) is a high-speed expansion card that attaches a computer to its peripherals.

Close