Average Ratings 0 Ratings
Average Ratings 0 Ratings
Description
The data refinery tool, which can be accessed through IBM Watson® Studio and Watson™ Knowledge Catalog, significantly reduces the time spent on data preparation by swiftly converting extensive volumes of raw data into high-quality, usable information suitable for analytics. Users can interactively discover, clean, and transform their data using more than 100 pre-built operations without needing any coding expertise. Gain insights into the quality and distribution of your data with a variety of integrated charts, graphs, and statistical tools. The tool automatically identifies data types and business classifications, ensuring accuracy and relevance. It also allows easy access to and exploration of data from diverse sources, whether on-premises or cloud-based. Data governance policies set by professionals are automatically enforced within the tool, providing an added layer of compliance. Users can schedule data flow executions for consistent results and easily monitor those results while receiving timely notifications. Furthermore, the solution enables seamless scaling through Apache Spark, allowing transformation recipes to be applied to complete datasets without the burden of managing Apache Spark clusters. This feature enhances efficiency and effectiveness in data processing, making it a valuable asset for organizations looking to optimize their data analytics capabilities.
Description
Pepperdata autonomous, application-level cost optimization delivers 30-47% greater cost savings for data-intensive workloads such as Apache Spark on Amazon EMR and Amazon EKS with no application changes. Using patented algorithms, Pepperdata Capacity Optimizer autonomously optimizes CPU and memory in real time with no application code changes.
Pepperdata automatically analyzes resource usage in real time, identifying where more work can be done, enabling the scheduler to add tasks to nodes with available resources and spin up new nodes only when existing nodes are fully utilized. The result: CPU and memory are autonomously and continuously optimized, without delay and without the need for recommendations to be applied, and the need for ongoing manual tuning is safely eliminated.
Pepperdata pays for itself, immediately decreasing instance hours/waste, increasing Spark utilization, and freeing developers from manual tuning to focus on innovation.
API Access
Has API
Yes
API Access
Has API
No
Integrations
Apache Spark
Yes
AWS Marketplace
No
Amazon EKS
No
Amazon EMR
No
Google Cloud Managed Service for Apache Spark
No
IBM Cloud
Yes
IBM Cloud Pak for Watson AIOps
Yes
IBM Watson
Yes
IBM Watson Discovery
Yes
IBM Watson Language Translator
Yes
Integrations
Apache Spark
Yes
AWS Marketplace
Yes
Amazon EKS
Yes
Amazon EMR
Yes
Google Cloud Managed Service for Apache Spark
Yes
IBM Cloud
No
IBM Cloud Pak for Watson AIOps
No
IBM Watson
No
IBM Watson Discovery
No
IBM Watson Language Translator
No
Pricing Details
No price information available.
Free Trial
Yes
Free Version
No
Pricing Details
No price information available.
Free Trial
Yes
Free Version
No
Deployment
Web-Based
Yes
On-Premises
No
iPhone App
No
iPad App
No
Android App
No
Windows
No
Mac
No
Linux
No
Chromebook
No
Deployment
Web-Based
Yes
On-Premises
No
iPhone App
No
iPad App
No
Android App
No
Windows
Yes
Mac
Yes
Linux
No
Chromebook
No
Customer Support
Business Hours
Yes
Live Rep (24/7)
Yes
Online Support
Yes
Customer Support
Business Hours
Yes
Live Rep (24/7)
No
Online Support
No
Types of Training
Training Docs
Yes
Webinars
Yes
Live Training (Online)
No
In Person
No
Types of Training
Training Docs
Yes
Webinars
Yes
Live Training (Online)
No
In Person
Yes
Vendor Details
Company Name
IBM
Founded
1911
Country
United States
Website
www.ibm.com/products/data-refinery
Vendor Details
Company Name
Pepperdata, Inc.
Founded
2012
Country
United States
Website
www.pepperdata.com
Product Features
Data Preparation
Collaboration Tools
No
Data Access
No
Data Blending
No
Data Cleansing
No
Data Governance
No
Data Mashup
No
Data Modeling
No
Data Transformation
No
Machine Learning
No
Visual User Interface
No
Product Features
Application Performance Monitoring (APM)
Baseline Manager
Yes
Diagnostic Tools
Yes
Full Transaction Diagnostics
Yes
Performance Control
Yes
Resource Management
Yes
Root-Cause Diagnosis
Yes
Server Performance
Yes
Trace Individual Transactions
Yes
Cloud Cost Management
Cost Reduction Optimization
Yes
Dashboard
Yes
Data Import/Export
No
Data Storage
No
Data Visualization
Yes
Resource Usage Reporting
Yes
Roles / Permissions
Yes
Spend and Cost Reporting
Yes