Average Ratings 0 Ratings

Total
ease
features
design
support

No User Reviews. Be the first to provide a review:

Write a Review

Average Ratings 0 Ratings

Total
ease
features
design
support

No User Reviews. Be the first to provide a review:

Write a Review

Description

Apache DataFusion is a versatile and efficient query engine crafted in Rust, leveraging Apache Arrow for its in-memory data representation. It caters to developers engaged in creating data-focused systems, including databases, data frames, machine learning models, and real-time streaming applications. With its SQL and DataFrame APIs, DataFusion features a vectorized, multi-threaded execution engine that processes data streams efficiently and supports various partitioned data sources. It is compatible with several native formats such as CSV, Parquet, JSON, and Avro, and facilitates smooth integration with popular object storage solutions like AWS S3, Azure Blob Storage, and Google Cloud Storage. The architecture includes a robust query planner and an advanced optimizer that boasts capabilities such as expression coercion, simplification, and optimizations that consider distribution and sorting, along with automatic reordering of joins. Furthermore, DataFusion allows for extensive customization, enabling developers to incorporate user-defined scalar, aggregate, and window functions along with custom data sources and query languages, making it a powerful tool for diverse data processing needs. This adaptability ensures that developers can tailor the engine to fit their unique use cases effectively.

Description

A streaming database is specifically designed to efficiently ingest, store, process, and analyze large volumes of data streams. This advanced data infrastructure integrates messaging, stream processing, and storage to enable real-time value extraction from your data. It continuously handles vast amounts of data generated by diverse sources, including sensors from IoT devices. Data streams are securely stored in a dedicated distributed streaming data storage cluster that can manage millions of streams. By subscribing to topics in HStreamDB, users can access and consume data streams in real-time at speeds comparable to Kafka. The system also allows for permanent storage of data streams, enabling users to replay and analyze them whenever needed. With a familiar SQL syntax, you can process these data streams based on event-time, similar to querying data in a traditional relational database. This functionality enables users to filter, transform, aggregate, and even join multiple streams seamlessly, enhancing the overall data analysis experience. Ultimately, the integration of these features ensures that organizations can leverage their data effectively and make timely decisions.

API Access

Has API Yes 

API Access

Has API No 

Screenshots View All

Screenshots View All

Integrations

Apache Arrow Yes 
Apache Avro Yes 
Apache Parquet Yes 
Apache Spark No 
C Yes 
Elastic Cloud No 
Google Cloud Storage Yes 
Google Sheets Yes 
Microsoft Excel Yes 
MongoDB No 
Oracle Fusion Cloud ERP No 
Presto No 
PyTorch No 
Python Yes 
Rust Yes 
SAP ERP No 
SDF Yes 
SQL Yes 
Snowflake No 
TensorFlow No 

Integrations

Apache Arrow No 
Apache Avro No 
Apache Parquet No 
Apache Spark Yes 
C No 
Elastic Cloud Yes 
Google Cloud Storage No 
Google Sheets No 
Microsoft Excel No 
MongoDB Yes 
Oracle Fusion Cloud ERP Yes 
Presto Yes 
PyTorch Yes 
Python No 
Rust No 
SAP ERP Yes 
SDF No 
SQL No 
Snowflake Yes 
TensorFlow Yes 

Pricing Details

Free
Free Trial No 
Free Version Yes 

Pricing Details

Free
Free Trial No 
Free Version Yes 

Deployment

Web-Based Yes 
On-Premises No 
iPhone App No 
iPad App No 
Android App No 
Windows No 
Mac No 
Linux No 
Chromebook No 

Deployment

Web-Based No 
On-Premises No 
iPhone App No 
iPad App No 
Android App No 
Windows Yes 
Mac Yes 
Linux Yes 
Chromebook No 

Customer Support

Business Hours Yes 
Live Rep (24/7) No 
Online Support Yes 

Customer Support

Business Hours No 
Live Rep (24/7) No 
Online Support Yes 

Types of Training

Training Docs Yes 
Webinars No 
Live Training (Online) No 
In Person Yes 

Types of Training

Training Docs Yes 
Webinars No 
Live Training (Online) No 
In Person No 

Vendor Details

Company Name

Apache Software Foundation

Founded

2019

Country

United States

Website

datafusion.apache.org

Vendor Details

Company Name

EMQ

Founded

2013

Country

United States

Website

hstream.io

Product Features

Database

Backup and Recovery No 
Creation / Development No 
Data Migration No 
Data Replication No 
Data Search No 
Data Security No 
Database Conversion No 
Mobile Access No 
Monitoring No 
NOSQL No 
Performance Analysis No 
Queries No 
Relational Interface No 
Virtualization No 

Product Features

Database

Backup and Recovery No 
Creation / Development No 
Data Migration No 
Data Replication No 
Data Search No 
Data Security No 
Database Conversion No 
Mobile Access No 
Monitoring No 
NOSQL No 
Performance Analysis No 
Queries No 
Relational Interface No 
Virtualization No 

Alternatives

Alternatives

ksqlDB Reviews

ksqlDB

Confluent