Average Ratings 0 Ratings
Average Ratings 0 Ratings
Description
The Apache Hadoop software library serves as a framework for the distributed processing of extensive data sets across computer clusters, utilizing straightforward programming models. It is built to scale from individual servers to thousands of machines, each providing local computation and storage capabilities. Instead of depending on hardware for high availability, the library is engineered to identify and manage failures within the application layer, ensuring that a highly available service can run on a cluster of machines that may be susceptible to disruptions. Numerous companies and organizations leverage Hadoop for both research initiatives and production environments. Users are invited to join the Hadoop PoweredBy wiki page to showcase their usage. The latest version, Apache Hadoop 3.3.4, introduces several notable improvements compared to the earlier major release, hadoop-3.2, enhancing its overall performance and functionality. This continuous evolution of Hadoop reflects the growing need for efficient data processing solutions in today's data-driven landscape.
Description
Kylo serves as an open-source platform designed for effective management of enterprise-level data lakes, facilitating self-service data ingestion and preparation while also incorporating robust metadata management, governance, security, and best practices derived from Think Big's extensive experience with over 150 big data implementation projects. It allows users to perform self-service data ingestion complemented by features for data cleansing, validation, and automatic profiling. Users can manipulate data effortlessly using visual SQL and an interactive transformation interface that is easy to navigate. The platform enables users to search and explore both data and metadata, examine data lineage, and access profiling statistics. Additionally, it provides tools to monitor the health of data feeds and services within the data lake, allowing users to track service level agreements (SLAs) and address performance issues effectively. Users can also create batch or streaming pipeline templates using Apache NiFi and register them with Kylo, thereby empowering self-service capabilities. Despite organizations investing substantial engineering resources to transfer data into Hadoop, they often face challenges in maintaining governance and ensuring data quality, but Kylo significantly eases the data ingestion process by allowing data owners to take control through its intuitive guided user interface. This innovative approach not only enhances operational efficiency but also fosters a culture of data ownership within organizations.
API Access
Has API
No
API Access
Has API
Yes
Integrations
Apache Spark
Yes
OpenText Data Privacy & Protection Foundation
Yes
Apache Impala
Yes
Apache Mahout
Yes
Apache Ranger
Yes
Baidu Palo
Yes
BiG EVAL
No
Cleo Integration Cloud
Yes
Deeplearning4j
Yes
Elasticsearch
No
Integrations
Apache Spark
Yes
OpenText Data Privacy & Protection Foundation
Yes
Apache Impala
No
Apache Mahout
No
Apache Ranger
No
Baidu Palo
No
BiG EVAL
Yes
Cleo Integration Cloud
No
Deeplearning4j
No
Elasticsearch
Yes
Pricing Details
No price information available.
Free Trial
No
Free Version
Yes
Pricing Details
No price information available.
Free Trial
No
Free Version
No
Deployment
Web-Based
Yes
On-Premises
No
iPhone App
No
iPad App
No
Android App
No
Windows
No
Mac
No
Linux
No
Chromebook
No
Deployment
Web-Based
Yes
On-Premises
No
iPhone App
No
iPad App
No
Android App
No
Windows
No
Mac
No
Linux
No
Chromebook
No
Customer Support
Business Hours
No
Live Rep (24/7)
No
Online Support
No
Customer Support
Business Hours
No
Live Rep (24/7)
No
Online Support
Yes
Types of Training
Training Docs
Yes
Webinars
No
Live Training (Online)
No
In Person
No
Types of Training
Training Docs
Yes
Webinars
Yes
Live Training (Online)
Yes
In Person
No
Vendor Details
Company Name
Apache Software Foundation
Founded
1999
Country
United States
Website
hadoop.apache.org
Vendor Details
Company Name
Teradata
Founded
1979
Country
United States
Website
kylo.io
Product Features
Product Features
Data Governance
Access Control
No
Data Discovery
No
Data Mapping
No
Data Profiling
No
Deletion Management
No
Email Management
No
Policy Management
No
Process Management
No
Roles Management
No
Storage Management
No
Data Lineage
Database Change Impact Analysis
No
Filter Lineage Links
No
Implicit Connection Discovery
No
Lineage Object Filtering
No
Object Lineage Tracing
No
Point-in-Time Visibility
No
User/Client/Target Connection Visibility
No
Visual & Text Lineage View
No
Data Preparation
Collaboration Tools
No
Data Access
No
Data Blending
No
Data Cleansing
No
Data Governance
No
Data Mashup
No
Data Modeling
No
Data Transformation
No
Machine Learning
No
Visual User Interface
No