Sponsored Links
-->

Sunday, December 3, 2017

What is Data Virtualization? - YouTube
src: i.ytimg.com

Data virtualization is any approach to data management that allows an application to retrieve and manipulate data without requiring technical details about the data, such as how it is formatted at source, or where it is physically located, and can provide a single customer view (or single view of any other entity) of the overall data.

Unlike the traditional extract, transform, load ("ETL") process, the data remains in place, and real-time access is given to the source system for the data. This reduces the risk of data errors, of the workload moving data around that may never be used, and it does not attempt to impose a single data model on the data (an example of heterogeneous data is a federated database system). The technology also supports the writing of transaction data updates back to the source systems. To resolve differences in source and consumer formats and semantics, various abstraction and transformation techniques are used. This concept and software is a subset of data integration and is commonly used within business intelligence, service-oriented architecture data services, cloud computing, enterprise search, and master data management.


Video Data virtualization



Examples

  • The Phone House--the trading name for the European operations of UK-based mobile phone retail chain Carphone Warehouse--implemented Denodo's data virtualization technology between its Spanish subsidiary's transactional systems and the Web-based systems of mobile operators.
  • Novartis, which implemented a data virtualization tool from Composite Software to enable its researchers to quickly combine data from both internal and external sources into a searchable virtual data store.
  • The storage-agnostic Primary Data data virtualization platform enables applications, servers, and clients to transparently access data while it is intelligently migrated between direct-attached, network-attached, private and public cloud storage. Server flash memory pioneer Fusion-io co-founder David Flynn, now Primary Data CTO, saw the need to move data across storage types to maximize efficiency with data virtualization.
  • Linked Data can use a single hyperlink-based Data Source Name (DSN) to provide a connection to a virtual database layer that is internally connected to a variety of back-end data sources using ODBC, JDBC, OLE DB, ADO.NET, SOA-style services, and/or REST patterns.
  • Database virtualization may use a single ODBC-based DSN to provide a connection to a similar virtual database layer.

Maps Data virtualization



Functionality

Data Virtualization software provides some or all of the following capabilities:

  • Abstraction - Abstract the technical aspects of stored data, such as location, storage structure, API, access language, and storage technology.
  • Virtualized Data Access - Connect to different data sources and make them accessible from a common logical data access point.
  • Transformation - Transform, improve quality, reformat, aggregate etc. source data for consumer use.
  • Data Federation - Combine result sets from across multiple source systems.
  • Data Delivery - Publish result sets as views and/or data services executed by client application or users when requested.

Data virtualization software may include functions for development, operation, and/or management.

Benefits include:

  • Reduce risk of data errors
  • Reduce systems workload through not moving data around
  • Increase speed of access to data on a real-time basis
  • Significantly reduce development and support time
  • Increase governance and reduce risk through the use of policies
  • Reduce data storage required

Drawbacks include:

  • May impact Operational systems response time, particularly if under-scaled to cope with unanticipated user queries or not tuned early on.
  • Does not impose a heterogeneous data model, meaning the user has to interpret the data, unless combined with Data Federation and business understanding of the data
  • Requires a defined Governance approach to avoid budgeting issues with the shared services
  • Not suitable for recording the historic snapshots of data. A data warehouse is better for this
  • Change management "is a huge overhead, as any changes need to be accepted by all applications and users sharing the same virtualization kit"

RED HAT JBOSS DATA VIRTUALIZATION Bill Kemp Sr. Solutions ...
src: slideplayer.com


Technology

Some data virtualization technologies include:

  • Actifio Copy Data Virtualization
  • Capsenta's Ultrawrap Platform
  • Cisco Data Virtualization (formerly Composite Software)
  • Delphix Data Virtualization Platform
  • Denodo Platform
  • DataVirtuality
  • Data Virtualization Platform
  • HiperFabric Data Virtualization and Integration
  • Querona
  • Stone Bond Technologies Enterprise Enabler Data Virtualization Platform - http://www.stonebond.com
  • Red Hat JBoss Enterprise Application Platform Data Virtualization
  • Veritas Provisioning File System / Data Virtualization Veritas_Technologies
  • XAware Data Services

IBM Data Virtualization Manager for zOS - Service Provider - YouTube
src: i.ytimg.com


History

Enterprise information integration (EII) (first coined by Metamatrix), now known as Red Hat JBoss Data Virtualization, and federated database systems are terms used by some vendors to describe a core element of data virtualization: the capability to create relational JOINs in a federated VIEW.


Simplifying Big Data with SAP HANA and SAP Cloud Platform Big Data ...
src: blogs.saphana.com


See also

  • Data integration
  • Enterprise information integration (EII)
  • Master data management
  • Database virtualization
  • Data Federation
  • Disparate system

SimpliVity Data Virtualization Platform Architecture - YouTube
src: i.ytimg.com


References


RED HAT JBOSS DATA VIRTUALIZATION Bill Kemp Sr. Solutions ...
src: slideplayer.com


Further reading

  • Data Virtualization: Going Beyond Traditional Data Integration to Achieve Business Agility, Judith R. Davis and Robert Eve
  • Data Virtualization for Business Intelligence Systems: Revolutionizing Data Integration for Data Warehouses Rick van der Lans
  • Data Integration Blueprint and Modeling: Techniques for a Scalable and Sustainable Architecture Anthony Giordano

Source of article : Wikipedia