IData Pipeline

IData Corporation

3.0
Free Trial
Product Overview

The IData Pipeline is a no-code AWS data pipeline. No coding required, just configure, ingest and consume. Using a simple JSON configuration for each dataset, you can inform the Pipeline what tasks should be performed upon a dataset. The Pipeline can perform deduplication of data, data quality checks, transformation with Javascript, transformation into S3 parquet in your data lake, transformation into Redshift or Snowflake and much more. When data lands in its final destination, automatic notifications are fired via SNS, enabling downstream systems and users to be notified immediately. The Pipeline also has a user-interface for real-time monitoring of data flows and automatic notifications to support teams when warnings or errors occur. Other features include Lakehouse support using Apache Iceberg open-source technology. Apache Iceberg enables upsert and time-travel queries against your data lake. Snowflake and Redshift support. If you need to move datasets directly into Redshift or Snowflake, this is automated with a simple JSON configuration. Dataset validation using data quality configuration rules and data deduplication capabilities. Data transformation. Associate Javascript with a dataset to perform your own powerful custom transformations. Infrastructure as code (IaC). The IData Pipeline can be spun up in an hour or less in your AWS account using our IAC code written entirely in Terraform. Structured, semi-structured, and unstructured data support. The Pipeline can consume and process basically unlimited types of data (csv, delimited, JSON, XML, PDF, images, video, etc). The Pipeline uses AWS Glue which automatically enables a large number of AWS services to consume datasets downstream.
Version: v2.3.2

By: IData Corporation
Categories: Migration
Data Preparation
ELT/ETL

Operating System
Linux
Delivery Methods
Container

Supported servicesAmazon EKSAmazon ECSAmazon ECS AnywhereAmazon EKS AnywhereSelf-managed Kubernetes

Highlights: No-code data pipeline - There's no coding, just simple JSON configuration, Data quality rules, Data transformation using Javascript, Automatic conversion into S3 using parquet, Lakehouse Technology - The pipeline uses open-source Apache Iceberg technology which provide upsert capabilities and time travel queries on top of S3 object store, Snowflake and Redshift integration including automatic table creation, Support for any data type, structured, semi-structured and unstructured

Category: Infrastructure Software

Delivery Method: Container

Company Information

Company Name: IData Corporation

About Company: IData is a data-driven cloud computing consulting company. IData has specific knowledge that crosses several areas of financial services including Asset Managers, Hedge Funds, Prime Brokerage, Investment Banking & Brokerage, Wealth Management, Mutual Funds and Retail Banking.

Our client list includes Goldman Sachs, Bridgewater, Canada Pension Plan Investment Board, Deutsche Bank, Bank of America, JP Morgan, UBS Painewebber, Credit Suisse, Citibank and a number of leading hedge funds.

We started working with cloud technologies in 2007 and we are a Select AWS Partner.

You might also like these products

Back to Product List

 

Order Product Data List

Our experts are here to help you.

Contact Us

Let our Expert team find the right Products for you