Voracity's Features & Benefits

Data discovery, integration, migration, governance, analytics, etc.

You won't believe how much you can do!

IRI Voracity® is a full-stack data lifecycle management platform that seamlessly combines the best of CoSort®, Eclipse™, Hadoop®, and other best-in-class technologies.

Since 1978, IRI CoSort accelerates and expands data transformation work. It's a proven practice: ETL optimizer, BI data preparer, DB load and query accelerator, data validation and quality tool, report generator, data converter, and it can mask PII and create test data.

Today, IRI Voracity leverages the power of CoSort and Hadoop, along with the familiarity and interoperability of Eclipse – plus plug-ins like Erwin Mapping Manager and KNIME – to provide and enhance data discovery, integration, migration, management, and analysis.

Open these tabs to explore Voracity's key data and enterprise information management features. Unless indicated as a third-party option with an asterisk (*), all listed features are included in the base license.

Job Design

Job Design Functions

Details

Advantages

New Job Assistants Automated generation of ETL, migration, masking, test data, and reorg scripts manual step-by-step job creation
Dialog and Form Editors Show and click the job specification in the toolbar or the script. Simpler parameter definition and modification
Job/mapping diagrams Intuitive GUI and Palette for ETL Mapping and Workflow Scripts Quick and easy order creation and review
Syntax-aware editors Write, Modify, and Validate: IRI DDF, Job Scripts, SQL, and Java Code-centric data and ETL architects are excited
Script menus and sketches Drill-Down GUI Views of and Dialog Interactivity with Job Scripts bypass the need to learn 4GL
DDF and SCL Metadata self-documenting, IRI-job, MIMB, and ADSMM capable simple, interoperable, batchable CLI code
EMF and XML Metadata self-documenting, reentrant EMM infrastructure (see below) Modify orders in any user interface and update all
„Gulfstream“ API/SDK documented JARs for IRI stream and metadata to/from applications Simple EAI, simple input and output for IRI XML
Erwin Mapping Manager Spreadsheet, codeless definitions for mapping source to target requirements Leveraging Existing Skills and Repositories
Erwin CatFx and LSC Automated Migration of ETL Tool and SQL Metadata in IRI Voracity Megavendor ETL tools leave sooner
Data Discovery Features Details Advantages
Data classification Definition and Management of Enterprise-Wide Data Class Libraries Application of transformation and protection field rules across multiple sources simultaneously
Structured Data Collection Preview of sequential file/DB layouts, definition of IRI DDFs automated metadata creation
DDF Conversion and Import CPY, CSV, LDIF, ODBC, and XML to DDF Migration in CLI/GUI Automated metadata conversion
DB & File Profiling - Statistics Selection/Report of column or field values, count, length, duplicates, etc. automated, custom table analysis
DB & Date Search finds values in tables or flat files that match the defined Java regex patterns automatic fine-tuning and protection of numerical PII
Database and File String Search Find explicit or dictionary-matching strings (names) in tables or flat files Automatic search and protection of individuals
Database and File Fuzzy Search finds near matches with defined probabilities using multiple algorithms simplify masking, BI, DQ and MDM
Database Integrity Check Compare foreign key with primary key in each defined column Simple maintenance of referential integrity
Cross-DB E-R Diagram Creation Create and customize views of tables and relationships in any database Visualization of multiple DB layouts
Dark Data Discovery Searching/Reporting on pattern-matching values and forensic information in documents Representation of Dark Data and Metadata
Dark Data Structuring Extracting and structuring found values in flat files, creating DDF and EIF integrate/correct unstructured data
Data Integration Features Details Advantages
Multiple source connections Manipulation and merging of structured and unstructured sources correlate internal and external data
Multiple target definitions Updates and bulk loads for tables, files, pipes, procedures, and reports individual I/Os, synchronized data
Fast Extraction (FACT) Parallel unloading of Oracle, DB2, MySQL, SQL Server, Sybase, etc. Faster ETL, reorganization, migration, archiving
CoSort Transformations Engine resource-optimized, single-pass, sort, join, aggregation, etc. Relieves BI/DB tools, prevents additional hardware
CoSort (SortCL) 4GL DDL/DML One-Script / One-Pass: Transformation, Cleansing, Masking, Mapping Consolidates tools, simplifies metadata, saves I/Os
Hadoop Transformations Options for MapReduce 2, Spark, Spark Streaming, Storm, Tez unlimited scalability
Data Mapping Format, Field, Endian, File/Table Mapping, Surrogate Key Supports ETL, federation, replication
Database DDL and Loader Compatibility Automated Table Creation & Load Scripts (pre-CoSorted) Faster table creation & loading
Data Cleaning & Validation find, filter, unify, replace, validate, standardize Higher data quality = reliable ETL & BI
Data migration functions Details Advantages
Data type conversion Conversion between alphanumeric, date/time, and multibyte formats Platform Migration Speed
File format conversion Conversion to/from Fixed/Delimited, LDIF, MF-ISAM, Vision, VB, XML, etc. Application Migration Support
Endian Conversion Big/Little Endian and BOM detection/modification at field and file level Facilitating Mainframe Migration
DB Table/Schema Conversion Profile existing DBs, create new schema & tables, populate mappings Facilitating migration of DB providers
Dark Data Structuring Search & Extract pattern-matched strings from MS documents, PDF, etc. Unlock & Utilize Unstructured Data
Data Federation Process data on-site (LDW) & send mashups to consoles/service apps Ad-hoc values & views without centralization
Data replication Unlocks & copies legacy data into new formats or ETL patterns Facilitates reuse of old data
JCL Data Definition Mainframe Sort Parameter Recognition & Conversion Leave legacy sorting software faster
Data management features Details Advantages
Metadata ManagementEMM) Creation, sharing, and tracking of robust data definition and rule infrastructure in Eclipse Reuse data and manage information
Master Data ManagementMDM) Develop, define, standardize, modify, classify, use, and share master data 360º view of customers and other data
Metadata Management Master and metadata lineage, impact analysis, version control in Git or Erwin Edge. Secure, shared, graphical data lineage
Data masking Search and secure PII with static and dynamic masking functions across 13 categories, automated auditing. Protection and Compliance, Secure BI/DW Operations
Re-ID Risk Scoring Measure and report on the risks of quasi-identifiers and then anonymize them comply with HIPAA EDM, etc.
Database subsetting Define, Create, and Mask Subsets of Production Databases for Testing Fast, agile prototyping of relational data
Data quality Discover, Deduplicate, Filter, Fuzzy Search, Integrity Check, Standardize MDM, ETL, improve analytical reliability
Encryption Key Management Store, rotate, and manage field-level encryption keys with integrated or web/HSM key storage technology improves decryption and recovery security
Test data generation Parse, generate, and load synthetic DB, file, mart, and report targets Faster, compliant EDW/App prototyping
COBIT support support the majority of COBIT goals through data lifecycle management tasks to reconcile data management with risk control
Compliance Services Training on Regulatory Compliance and Collaboration with HIPAA-Qualified Statisticians and Attorneys for Review/Insurance HIPAA and PCI Gap Assessment, Compliance
Analytics Functions Details Advantages
Integrated Reporting Custom detailed and overview BI objectives with math, transformations, masking, etc. Report during transformation
Change Data Capture Find insertions, updates, deletions, no changes, and value deltas from data, not from logs Multi-Input, Custom Output
Slowly Changing Dimensions Update and Report on Data Changes Using Fuzzy Logic Use SCD data in BI, DI
Clickstream and CDR Support Transform, convert, mask, link, and report data in weblog and ASN.1 format. Bypass mediation software
Data segmentation Custom selection and silo data during integration, masking, quality, and reporting CDI, CRM, DLP, MDM
BIRT, KNIME, and Splunk Integrations feed IRI data results directly into plug-ins to enable immediate reporting, deep learning, analytics, and advanced display and actions Faster, free visual BI in Eclipse
BI/Analytics Tool Data Preparation Cleansing raw data and delivering ready-to-use results to accelerate outcomes from BOBJ, Cognos, Microstrategy, Oracle (OBIEE, DV/D, or OAC), Power BI, QlikView, R, Splunk, Spotfire, Tableau, etc. Remove PI from the BI layer
JupiterOne * Analytics Support for real-time visualizations with SQL syntax and Spark processing in Voracity Immediate results, bypass ETL