Average Ratings 0 Ratings
Average Ratings 0 Ratings
Description
Parquet was developed to provide the benefits of efficient, compressed columnar data representation to all projects within the Hadoop ecosystem. Designed with a focus on accommodating complex nested data structures, Parquet employs the record shredding and assembly technique outlined in the Dremel paper, which we consider to be a more effective strategy than merely flattening nested namespaces. This format supports highly efficient compression and encoding methods, and various projects have shown the significant performance improvements that arise from utilizing appropriate compression and encoding strategies for their datasets. Furthermore, Parquet enables the specification of compression schemes at the column level, ensuring its adaptability for future developments in encoding technologies. It is crafted to be accessible for any user, as the Hadoop ecosystem comprises a diverse range of data processing frameworks, and we aim to remain neutral in our support for these different initiatives. Ultimately, our goal is to empower users with a flexible and robust tool that enhances their data management capabilities across various applications.
Description
Experience serverless and interactive data querying with IBM Cloud Object Storage, enabling you to analyze your data directly at its source without the need for ETL processes, databases, or infrastructure management. IBM Cloud SQL Query leverages Apache Spark, a high-performance, open-source data processing engine designed for quick and flexible analysis, allowing SQL queries without requiring ETL or schema definitions. You can easily perform data analysis on your IBM Cloud Object Storage via our intuitive query editor and REST API. With a pay-per-query pricing model, you only incur costs for the data that is scanned, providing a cost-effective solution that allows for unlimited queries. To enhance both savings and performance, consider compressing or partitioning your data. Furthermore, IBM Cloud SQL Query ensures high availability by executing queries across compute resources located in various facilities. Supporting multiple data formats, including CSV, JSON, and Parquet, it also accommodates standard ANSI SQL for your querying needs, making it a versatile tool for data analysis. This capability empowers organizations to make data-driven decisions more efficiently than ever before.
API Access
Has API
No
API Access
Has API
Yes
Integrations
Autymate
Yes
Data Sentinel
Yes
Amazon Data Firehose
Yes
Apache DataFusion
Yes
Apache Spark
No
Blotout
Yes
CSViewer
Yes
Ficstar
Yes
Flyte
Yes
Gravity Data
Yes
Integrations
Autymate
Yes
Data Sentinel
Yes
Amazon Data Firehose
No
Apache DataFusion
No
Apache Spark
Yes
Blotout
No
CSViewer
No
Ficstar
No
Flyte
No
Gravity Data
No
Pricing Details
No price information available.
Free Trial
No
Free Version
No
Pricing Details
$5.00/Terabyte-Month
Free Trial
Yes
Free Version
Yes
Deployment
Web-Based
No
On-Premises
No
iPhone App
No
iPad App
No
Android App
No
Windows
Yes
Mac
Yes
Linux
Yes
Chromebook
No
Deployment
Web-Based
Yes
On-Premises
No
iPhone App
No
iPad App
No
Android App
No
Windows
Yes
Mac
Yes
Linux
Yes
Chromebook
No
Customer Support
Business Hours
No
Live Rep (24/7)
No
Online Support
Yes
Customer Support
Business Hours
Yes
Live Rep (24/7)
Yes
Online Support
Yes
Types of Training
Training Docs
Yes
Webinars
Yes
Live Training (Online)
No
In Person
No
Types of Training
Training Docs
Yes
Webinars
Yes
Live Training (Online)
Yes
In Person
Yes
Vendor Details
Company Name
The Apache Software Foundation
Founded
1999
Country
United States
Website
parquet.apache.org
Vendor Details
Company Name
IBM
Founded
1911
Country
United States
Website
www.ibm.com/cloud/sql-query
Product Features
Product Features
Relational Database
ACID Compliance
Yes
Data Failure Recovery
Yes
Multi-Platform
Yes
Referential Integrity
Yes
SQL DDL Support
Yes
SQL DML Support
Yes
System Catalog
Yes
Unicode Support
Yes