Voracity's Features & Benefits
Data discovery, integration, migration, governance, analytics, etc.
You won't believe how much you can do!
IRI Voracity® is a full-stack data lifecycle management platform that seamlessly combines the best of CoSort®, Eclipse™, Hadoop®, and other best-in-class technologies.
Since 1978, IRI CoSort accelerates and expands data transformation work. It's a proven practice: ETL optimizer, BI data preparer, DB load and query accelerator, data validation and quality tool, report generator, data converter, and it can mask PII and create test data.
Today, IRI Voracity leverages the power of CoSort and Hadoop, along with the familiarity and interoperability of Eclipse – plus plug-ins like Erwin Mapping Manager and KNIME – to provide and enhance data discovery, integration, migration, management, and analysis.
Open these tabs to explore Voracity's key data and enterprise information management features. Unless indicated as a third-party option with an asterisk (*), all listed features are included in the base license.
Job Design
Job Design Functions |
Details |
Advantages |
| New Job Assistants | Automated generation of ETL, migration, masking, test data, and reorg scripts | manual step-by-step job creation |
| Dialog and Form Editors | Show and click the job specification in the toolbar or the script. | Simpler parameter definition and modification |
| Job/mapping diagrams | Intuitive GUI and Palette for ETL Mapping and Workflow Scripts | Quick and easy order creation and review |
| Syntax-aware editors | Write, Modify, and Validate: IRI DDF, Job Scripts, SQL, and Java | Code-centric data and ETL architects are excited |
| Script menus and sketches | Drill-Down GUI Views of and Dialog Interactivity with Job Scripts | bypass the need to learn 4GL |
| DDF and SCL Metadata | self-documenting, IRI-job, MIMB, and ADSMM capable | simple, interoperable, batchable CLI code |
| EMF and XML Metadata | self-documenting, reentrant EMM infrastructure (see below) | Modify orders in any user interface and update all |
| „Gulfstream“ API/SDK | documented JARs for IRI stream and metadata to/from applications | Simple EAI, simple input and output for IRI XML |
| Erwin Mapping Manager | Spreadsheet, codeless definitions for mapping source to target requirements | Leveraging Existing Skills and Repositories |
| Erwin CatFx and LSC | Automated Migration of ETL Tool and SQL Metadata in IRI Voracity | Megavendor ETL tools leave sooner |
Data discovery
| Data Discovery Features | Details | Advantages |
| Data classification | Definition and Management of Enterprise-Wide Data Class Libraries | Application of transformation and protection field rules across multiple sources simultaneously |
| Structured Data Collection | Preview of sequential file/DB layouts, definition of IRI DDFs | automated metadata creation |
| DDF Conversion and Import | CPY, CSV, LDIF, ODBC, and XML to DDF Migration in CLI/GUI | Automated metadata conversion |
| DB & File Profiling - Statistics | Selection/Report of column or field values, count, length, duplicates, etc. | automated, custom table analysis |
| DB & Date Search | finds values in tables or flat files that match the defined Java regex patterns | automatic fine-tuning and protection of numerical PII |
| Database and File String Search | Find explicit or dictionary-matching strings (names) in tables or flat files | Automatic search and protection of individuals |
| Database and File Fuzzy Search | finds near matches with defined probabilities using multiple algorithms | simplify masking, BI, DQ and MDM |
| Database Integrity Check | Compare foreign key with primary key in each defined column | Simple maintenance of referential integrity |
| Cross-DB E-R Diagram Creation | Create and customize views of tables and relationships in any database | Visualization of multiple DB layouts |
| Dark Data Discovery | Searching/Reporting on pattern-matching values and forensic information in documents | Representation of Dark Data and Metadata |
| Dark Data Structuring | Extracting and structuring found values in flat files, creating DDF and EIF | integrate/correct unstructured data |
Data integration
| Data Integration Features | Details | Advantages |
| Multiple source connections | Manipulation and merging of structured and unstructured sources | correlate internal and external data |
| Multiple target definitions | Updates and bulk loads for tables, files, pipes, procedures, and reports | individual I/Os, synchronized data |
| Fast Extraction (FACT) | Parallel unloading of Oracle, DB2, MySQL, SQL Server, Sybase, etc. | Faster ETL, reorganization, migration, archiving |
| CoSort Transformations Engine | resource-optimized, single-pass, sort, join, aggregation, etc. | Relieves BI/DB tools, prevents additional hardware |
| CoSort (SortCL) 4GL DDL/DML | One-Script / One-Pass: Transformation, Cleansing, Masking, Mapping | Consolidates tools, simplifies metadata, saves I/Os |
| Hadoop Transformations | Options for MapReduce 2, Spark, Spark Streaming, Storm, Tez | unlimited scalability |
| Data Mapping | Format, Field, Endian, File/Table Mapping, Surrogate Key | Supports ETL, federation, replication |
| Database DDL and Loader Compatibility | Automated Table Creation & Load Scripts (pre-CoSorted) | Faster table creation & loading |
| Data Cleaning & Validation | find, filter, unify, replace, validate, standardize | Higher data quality = reliable ETL & BI |
Data migration
| Data migration functions | Details | Advantages |
| Data type conversion | Conversion between alphanumeric, date/time, and multibyte formats | Platform Migration Speed |
| File format conversion | Conversion to/from Fixed/Delimited, LDIF, MF-ISAM, Vision, VB, XML, etc. | Application Migration Support |
| Endian Conversion | Big/Little Endian and BOM detection/modification at field and file level | Facilitating Mainframe Migration |
| DB Table/Schema Conversion | Profile existing DBs, create new schema & tables, populate mappings | Facilitating migration of DB providers |
| Dark Data Structuring | Search & Extract pattern-matched strings from MS documents, PDF, etc. | Unlock & Utilize Unstructured Data |
| Data Federation | Process data on-site (LDW) & send mashups to consoles/service apps | Ad-hoc values & views without centralization |
| Data replication | Unlocks & copies legacy data into new formats or ETL patterns | Facilitates reuse of old data |
| JCL Data Definition | Mainframe Sort Parameter Recognition & Conversion | Leave legacy sorting software faster |
Data management
| Data management features | Details | Advantages |
|---|---|---|
| Metadata ManagementEMM) | Creation, sharing, and tracking of robust data definition and rule infrastructure in Eclipse | Reuse data and manage information |
| Master Data ManagementMDM) | Develop, define, standardize, modify, classify, use, and share master data | 360º view of customers and other data |
| Metadata Management | Master and metadata lineage, impact analysis, version control in Git or Erwin Edge. | Secure, shared, graphical data lineage |
| Data masking | Search and secure PII with static and dynamic masking functions across 13 categories, automated auditing. | Protection and Compliance, Secure BI/DW Operations |
| Re-ID Risk Scoring | Measure and report on the risks of quasi-identifiers and then anonymize them | comply with HIPAA EDM, etc. |
| Database subsetting | Define, Create, and Mask Subsets of Production Databases for Testing | Fast, agile prototyping of relational data |
| Data quality | Discover, Deduplicate, Filter, Fuzzy Search, Integrity Check, Standardize | MDM, ETL, improve analytical reliability |
| Encryption Key Management | Store, rotate, and manage field-level encryption keys with integrated or web/HSM key storage technology | improves decryption and recovery security |
| Test data generation | Parse, generate, and load synthetic DB, file, mart, and report targets | Faster, compliant EDW/App prototyping |
| COBIT support | support the majority of COBIT goals through data lifecycle management tasks | to reconcile data management with risk control |
| Compliance Services | Training on Regulatory Compliance and Collaboration with HIPAA-Qualified Statisticians and Attorneys for Review/Insurance | HIPAA and PCI Gap Assessment, Compliance |
BI & Analytics
| Analytics Functions | Details | Advantages |
|---|---|---|
| Integrated Reporting | Custom detailed and overview BI objectives with math, transformations, masking, etc. | Report during transformation |
| Change Data Capture | Find insertions, updates, deletions, no changes, and value deltas from data, not from logs | Multi-Input, Custom Output |
| Slowly Changing Dimensions | Update and Report on Data Changes Using Fuzzy Logic | Use SCD data in BI, DI |
| Clickstream and CDR Support | Transform, convert, mask, link, and report data in weblog and ASN.1 format. | Bypass mediation software |
| Data segmentation | Custom selection and silo data during integration, masking, quality, and reporting | CDI, CRM, DLP, MDM |
| BIRT, KNIME, and Splunk Integrations | feed IRI data results directly into plug-ins to enable immediate reporting, deep learning, analytics, and advanced display and actions | Faster, free visual BI in Eclipse |
| BI/Analytics Tool Data Preparation | Cleansing raw data and delivering ready-to-use results to accelerate outcomes from BOBJ, Cognos, Microstrategy, Oracle (OBIEE, DV/D, or OAC), Power BI, QlikView, R, Splunk, Spotfire, Tableau, etc. | Remove PI from the BI layer |
| JupiterOne * Analytics | Support for real-time visualizations with SQL syntax and Spark processing in Voracity | Immediate results, bypass ETL |