Average Ratings 0 Ratings
Average Ratings 0 Ratings
Description
Hudi serves as a robust platform for constructing streaming data lakes equipped with incremental data pipelines, all while utilizing a self-managing database layer that is finely tuned for lake engines and conventional batch processing. It effectively keeps a timeline of every action taken on the table at various moments, enabling immediate views of the data while also facilitating the efficient retrieval of records in the order they were received. Each Hudi instant is composed of several essential components, allowing for streamlined operations. The platform excels in performing efficient upserts by consistently linking a specific hoodie key to a corresponding file ID through an indexing system. This relationship between record key and file group or file ID remains constant once the initial version of a record is written to a file, ensuring stability in data management. Consequently, the designated file group encompasses all iterations of a collection of records, allowing for seamless data versioning and retrieval. This design enhances both the reliability and efficiency of data operations within the Hudi ecosystem.
Description
The newest iteration of DBIntegrate, version 3.0.3.7, is now accessible for download. This update features improvements to Change Data Capture (CDC) and introduces new functionalities for data de-duplication, facilitating users in identifying duplicates more efficiently. Notably, CDC can now output to a flat-text file when disconnected from the message queue, which is subsequently read back into the message queue upon reconnection, ensuring that messages are delivered to the target data source in the correct order. Additionally, the flat-text file option may serve as the default for CDC, allowing for seamless overnight batch imports into other systems. A log loader mechanism accompanies this release, permitting the loading of files through a command line utility. Moreover, DBIntegrate now enables the recording of de-duplication merge scores in the DBI_WORK temporary tables, and the master record can be displayed in a new column labeled DBI_RecordMerged. This update marks a significant advancement in the software's capabilities, streamlining the data integration process for its users.
API Access
Has API
API Access
Has API
Integrations
AWS Marketplace
Actian Data Observability
Alluxio
Amazon Athena
Amazon Redshift
Apache Cassandra
Apache Doris
Apache Kafka
Apache Spark
Azure Data Lake
Integrations
AWS Marketplace
Actian Data Observability
Alluxio
Amazon Athena
Amazon Redshift
Apache Cassandra
Apache Doris
Apache Kafka
Apache Spark
Azure Data Lake
Pricing Details
No price information available.
Free Trial
Free Version
Pricing Details
No price information available.
Free Trial
Free Version
Deployment
Web-Based
On-Premises
iPhone App
iPad App
Android App
Windows
Mac
Linux
Chromebook
Deployment
Web-Based
On-Premises
iPhone App
iPad App
Android App
Windows
Mac
Linux
Chromebook
Customer Support
Business Hours
Live Rep (24/7)
Online Support
Customer Support
Business Hours
Live Rep (24/7)
Online Support
Types of Training
Training Docs
Webinars
Live Training (Online)
In Person
Types of Training
Training Docs
Webinars
Live Training (Online)
In Person
Vendor Details
Company Name
Apache Corporation
Founded
1954
Country
United States
Website
hudi.apache.org
Vendor Details
Company Name
Transoft
Founded
1983
Country
United Kingdom
Website
transoftdev.wordpress.com/2012/07/19/transoft-dbintegrate-v-3-0-3-7-release/
Product Features
Data Warehouse
Ad hoc Query
Analytics
Data Integration
Data Migration
Data Quality Control
ETL - Extract / Transfer / Load
In-Memory Processing
Match & Merge
Product Features
Database
Backup and Recovery
Creation / Development
Data Migration
Data Replication
Data Search
Data Security
Database Conversion
Mobile Access
Monitoring
NOSQL
Performance Analysis
Queries
Relational Interface
Virtualization