Browse Definitions:
Reference

Quick Start Glossary: Big data analytics

Print out our handy glossary of essential big data analytics terminology for a fast reference. Online, each term links to our full definition, which also includes resources for further learning.

advanced analytics -- future-oriented analyses that can be used to help drive changes and improvements in business practices.

big data – a voluminous amount – perhaps petabytes or more -- of structured, semi-structured and unstructured data that has the potential to be mined for information.

big data analytics – the process of examining large data sets containing a variety of data types -- i.e., big data -- to uncover hidden patterns, unknown correlations, market trends, customer preferences and other useful business information. 

big data as a service (BDaaS) -- the delivery of statistical analysis tools or information by an outside provider that helps organizations understand and use insights gained from large information sets in order to gain a competitive advantage.

big data CRM  -- the practice of integrating big data into a company's customer relationship management processes with the goals of improving customer service, calculating return on investment on various initiatives and predicting clientele behavior.

big data management -- the organization, administration and governance of large volumes of both structured and unstructured data.

data analytics (DA) -- the science of examining raw data with the purpose of drawing conclusions about that information. 

data mining -- sorting through data to identify patterns and establish relationships.

data scientist -- a job title for an employee or business intelligence (BI) consultant who excels at analyzing data, particularly large amounts of data, to help a business gain a competitive edge.  

data visualization – representation of data in graphic form to make its information more readily apparent. Patterns, trends and correlations that might go undetected in text-based data can be exposed and recognized easier with data visualization software.

deep analytics -- the application of sophisticated data processing techniques to yield information from large and typically multi-source data sets comprised of both unstructured and semi-structured data.

enterprise data hub, also referred to as a data lake -- a new big data management model for big data that utilizes Hadoop as the central data repository.

Google BigQuery -- a cloud-based big data analytics web service for processing very large read-only data sets. BigQuery was designed for analyzing data on the order of billions of rows, using a SQL-like syntax.

Hadoop -- a free, Java-based programming framework that supports the processing of large data sets in a distributed computing environment.

in-memory analytics -- an approach to querying data when it resides in a computer’s random access memory (RAM), as opposed to querying data that is stored on physical disks. 

predictive analytics -- the branch of data mining concerned with the prediction of future probabilities and trends. 

small data -- data in a volume and format that makes it accessible, informative and actionable. Examples include baseball scores, inventory reports, driving records, sales data, biometric measurements, search histories, weather forecasts and usage alerts.

text mining -- the analysis of data contained in natural language text. Text mining works by transposing words and phrases in unstructured data into numerical values which can then be linked with structured data in a database and analyzed with traditional data mining techniques.

unstructured data -- data that is not contained in a database or some other type of data structure.  

visual analytics -- a form of inquiry in which data that provides insight into solving a problem is displayed in an interactive, graphical manner.

This was last updated in April 2015

Start the conversation

Send me notifications when other members comment.

By submitting you agree to receive email from TechTarget and its partners. If you reside outside of the United States, you consent to having your personal data transferred to and processed in the United States. Privacy

Please create a username to comment.

-ADS BY GOOGLE

File Extensions and File Formats

Powered by:

SearchCompliance

  • internal audit (IA)

    An internal audit (IA) is an organizational initiative to monitor and analyze its own business operations in order to determine ...

  • pure risk (absolute risk)

    Pure risk, also called absolute risk, is a category of threat that is beyond human control and has only one possible outcome if ...

  • risk assessment

    Risk assessment is the identification of hazards that could negatively impact an organization's ability to conduct business.

SearchSecurity

  • computer exploit

    A computer exploit, or exploit, is an attack on a computer system, especially one that takes advantage of a particular ...

  • cyberwarfare

    Cyberwarfare is computer- or network-based conflict involving politically motivated attacks by a nation-state on another ...

  • insider threat

    Insider threat is a generic term for a threat to an organization's security or data that comes from within.

SearchHealthIT

SearchDisasterRecovery

  • business continuity and disaster recovery (BCDR)

    Business continuity and disaster recovery (BCDR) are closely related practices that describe an organization's preparation for ...

  • business continuity plan (BCP)

    A business continuity plan (BCP) is a document that consists of the critical information an organization needs to continue ...

  • call tree

    A call tree -- sometimes referred to as a phone tree -- is a telecommunications chain for notifying specific individuals of an ...

SearchStorage

  • OpenStack Block Storage (Cinder)

    OpenStack Block Storage (Cinder) is open source software designed to create and manage a service that provides persistent data ...

  • SATA Express (SATAe)

    SATA Express (SATAe or Serial ATA Express) is a bus interface to connect storage devices to a computer motherboard, supporting ...

  • DIMM (dual in-line memory module)

    A DIMM (dual in-line memory module) is the standard memory card used in servers and PCs.

SearchSolidStateStorage

  • hybrid flash array

    A hybrid flash array is a solid-state storage system that contains a mix of flash memory drives and hard disk drives.

  • 3D XPoint

    3D XPoint is memory storage technology jointly developed by Intel and Micron Technology Inc.

  • RRAM or ReRAM (resistive RAM)

    RRAM or ReRAM (resistive random access memory) is a form of nonvolatile storage that operates by changing the resistance of a ...

SearchCloudStorage

  • Google Cloud Storage

    Google Cloud Storage is an enterprise public cloud storage platform that can house large unstructured data sets.

  • RESTful API

    A RESTful application program interface breaks down a transaction to create a series of small modules, each of which addresses an...

  • cloud storage infrastructure

    Cloud storage infrastructure is the hardware and software framework that supports the computing requirements of a private or ...

Close