Software Alternatives, Accelerators & Startups

Kettle Pentaho VS Dataiku

Compare Kettle Pentaho VS Dataiku and see what are their differences

Kettle Pentaho logo Kettle Pentaho

Pentaho Data Integration ( ETL ) a.k.a Kettle

Dataiku logo Dataiku

Dataiku is the developer of DSS, the integrated development platform for data professionals to turn raw data into predictions.
  • Kettle Pentaho Landing page
    Landing page //
    2023-09-22
  • Dataiku Landing page
    Landing page //
    2023-08-17

Kettle Pentaho videos

No Kettle Pentaho videos yet. You could help us improve this page by suggesting one.

+ Add video

Dataiku videos

AutoML with Dataiku: And End-to-End Demo

More videos:

  • Review - Dataiku: For Everyone in the Data-Powered Organization
  • Tutorial - Dataiku DSS Tutorial 101: Your very first steps

Category Popularity

0-100% (relative to Kettle Pentaho and Dataiku)
Data Integration
100 100%
0% 0
Data Science And Machine Learning
Web Service Automation
100 100%
0% 0
Data Science Tools
0 0%
100% 100

User comments

Share your experience with using Kettle Pentaho and Dataiku. For example, how are they different and which one is better?
Log in or Post with

Reviews

These are some of the external sources and on-site user reviews we've used to compare Kettle Pentaho and Dataiku

Kettle Pentaho Reviews

10 Best Open Source ETL Tools for Data Integration
The best ETL tool is the one that aligns with your demands and provides the solution that you are looking for. Perhaps, you can choose Keboola, Pentaho Kettle, CloverDX, Logstash, and Apache Kafka. However, you must go for Scriptella or Talend Open Studio if your team wants to save time manually creating and connecting data pipelines. These tools are perfect for technically...
Source: testsigma.com
11 Best FREE Open-Source ETL Tools in 2024
Pentaho Kettle is now a part of the Hitachi Vantara Community and provides ETL capabilities using a metadata-driven approach. This tool allows users to create their own data manipulation jobs without writing a single line of code. Hitachi Vantara also offers Open-Source BI tools for reporting and Data Mining that work seamlessly with Pentaho Kettle.
Source: hevodata.com
Top 10 Popular Open-Source ETL Tools for 2021
Pentaho Kettle is now a part of the Hitachi Vantara Community and provides ETL capabilities using a metadata-driven approach. It has a graphical drag and drop UI and standard architecture. This tool allows users to create their own data manipulation jobs without writing a single line of code. Hitachi Vantara also offers Open-Source BI tools for reporting and Data Mining that...
Source: hevodata.com

Dataiku Reviews

15 data science tools to consider using in 2021
Some platforms are also available in free open source or community editions -- examples include Dataiku and H2O. Knime combines an open source analytics platform with a commercial Knime Server software package that supports team-based collaboration and workflow automation, deployment and management.
The 16 Best Data Science and Machine Learning Platforms for 2021
Description: Dataiku offers an advanced analytics solution that allows organizations to create their own data tools. The company’s flagship product features a team-based user interface for both data analysts and data scientists. Dataiku’s unified framework for development and deployment provides immediate access to all the features needed to design data tools from scratch....

What are some alternatives?

When comparing Kettle Pentaho and Dataiku, you can also consider the following products

Oracle Data Integrator - Oracle Data Integrator is a data integration platform that covers batch loads, to trickle-feed integration processes.

Scikit-learn - scikit-learn (formerly scikits.learn) is an open source machine learning library for the Python programming language.

Talend - Talend Cloud delivers a single, open platform for data integration across cloud and on-premises environments. Put more data to work for your business faster with Talend.

Pandas - Pandas is an open source library providing high-performance, easy-to-use data structures and data analysis tools for the Python.

Apache Airflow - Airflow is a platform to programmaticaly author, schedule and monitor data pipelines.

NumPy - NumPy is the fundamental package for scientific computing with Python