This section of the Stackable documentation contains information about the individual operators that make up the Stackable Data Platform. You can find an overview over the product operators as well as internal operators below.
This section also contains an overview over the supported product versions and how to enable monitoring in all operators.
Airflow is a workflow engine and your replacement should you be using Apache Oozie.
Read more
HDFS is a distributed file system that provides high-throughput access to application data.
Read more
The Apache Hive data warehouse software facilitates reading, writing, and managing large datasets residing in distributed storage using SQL. We support the Hive Metastore.
Read more
Apache Kafka is an open-source distributed event streaming platform used by thousands of companies for high-performance data pipelines, streaming analytics, data integration, and mission-critical applications.
Read more
Apache Spark is a multi-language engine for executing data engineering, data science, and machine learning on single-node machines or clusters.
Read more
Fast distributed SQL query engine for big data analytics that helps you explore your data universe.
Read more
The commons operator supplies shared CustomResourceDefinitions for all other operators.
Read more
The secret operator is responsible for handling secrets as well as certificates and auto-renewing them.
Read more