Stackable Docs Hub

Stackable

Stackable

Managed Data Lakehouse: sovereign & open source | Stackable

Managed Stackable
Data Lakehouse

The enterprise data lakehouse as a managed service, built on the established tools of the Stackable Data Platform (SDP):

Sovereign · Compliant · Open Source

general illustration

Stackable - your sovereign data platform

Use Cases

From data silos to a central data platform

The Managed Stackable Data Lakehouse is designed for analytical enterprise workloads: sovereign data management independent of industry and company size.

Central data platform

Consolidate data from different source systems (ERP, CRM, machine data, Excel lists) into a single lakehouse, for consistent reporting and self-service analytics instead of manual Excel consolidation.

Data foundation for AI and ML projects

Create a clean, consolidated data foundation on which machine-learning models or AI applications can be built, without expensive parallel infrastructure.

Compliance and auditability

Keep historical data states traceable and audit-proof (for example for GDPR requests, financial audits or certifications) through the time-travel capabilities of the lakehouse.

Customer and sales analytics

Bring together order histories, customer behaviour and sales data to spot cross-selling potential and customer churn early.

Predictive maintenance

Analyse machine and sensor data to reduce downtime and maintenance costs, developing and applying specialised models, including in real time.

Supply chain transparency

Combine data from procurement, logistics and production to identify bottlenecks, supplier dependencies and data silos early.

Build on the solid Stackable Data Platform foundation

The open source data lakehouse

All modules of the Managed Stackable Data Lakehouse are stable, performant and fully managed – future-proof by design.

Trino logo

Trino

Fast, distributed SQL query engine for interactive ad-hoc analytics directly on your data lakes, with no data movement.

Apache Iceberg

Open, high-performance table format for ACID transactions, time travel and cross-schema data consistency on S3 storage.

Supercast logo

Apache Superset

Interactive dashboards and ad-hoc data exploration for business users.

nifi logo

Apache NiFi

Visual dataflow design for ingestion, transformation and routing.

Apache Spark logo

Apache Spark

The proven, high-speed compute engine for demanding batch processing, ETL pipelines and machine learning.

Apache Airflow logo

Apache Airflow

Scalable workflow manager for orchestration, scheduling and complete monitoring of all data streams.

Open policy agent logo

OPA & OIDC

Fine-grained authorization policies and single sign-on across all components.

Keycloak & OpenID

Single sign-on (SSO) and secure user authentication.

The Infrastructure

Configure your Managed Service

The Stackable Data Lakehouse in different sizes: choose your infrastructure with the right service partner.

Step 1 · Choose infrastructure Kubernetes & Object Storage
Stackable Stackable Data Platform · Trino · Apache Spark · Apache Iceberg · Apache Airflow · Apache Superset · Apache NiFi
Step 2 · Choose package Size & price
Red HatAdd Red Hat OpenShift Optional: enterprise container orchestration & access management.

Step 3 · Request this combination
Package

* The first 30 days are free. From day 31 onward, 449 € per month (without support and SLA).

+ Red Hat OpenShift () – price on request

Request this combination
Your selection + + Custom

Frequently asked questions

Here you will find quick answers to the most important questions about our flexible multi-cloud combinations and managed infrastructures.

What is a managed data lakehouse and who is it for?

A managed data lakehouse is a ready-to-run data platform that combines a data lake and a data warehouse, with a partner running it for you. It is particularly suited to organizations with large data volumes that do not want to build their own platform ops capacity.

It avoids vendor lock-in through open formats (Apache Iceberg on S3) and a platform defined as code, portable via GitOps and built from open source modules. Cloud, managed partner and components can all be swapped.

Yes. Choose the infrastructure yourself, for example a cloud service provider under German jurisdiction, a private cloud or your own data centre. All hosting partners and data centres are ISO 27001 certified and operate in full compliance with the European General Data Protection Regulation (GDPR). Secure access control via Open Policy Agent and Keycloak, together with an audit-proof history through Apache Iceberg, supports regulated industries (for example DORA in the financial sector).

No. In the managed model, the partner handles provisioning, operations, monitoring and updates on Kubernetes, so you can focus on your data applications.

No. Business users work through the Stackable Cockpit without any contact with the underlying infrastructure, using Apache Superset for dashboards and ad-hoc analysis, Apache Airflow for orchestration, Apache NiFi for data pipelines, plus a Stackable query editor for Trino and an S3 manager.

The monthly price bundles the SDP subscription, the managed service and the cloud infrastructure with SLA support. This includes running the platform, covering configuration, monitoring, patches and updates.

The packages differ in compute and storage capacity (nodes, vCPUs, RAM, S3), while the feature set stays the same. If the standard packages are not enough, “Individual” sizes the platform to your specific needs. The XS package is the entry point with minimal resources for tests, proofs of concept and small workloads, without support or SLA. For production use, choose a paid package.

Yes, we agree binding service level agreements tailored precisely to your business-critical workloads. This includes availability guarantees of up to 99.99 % and dedicated support (8×5), and around the clock on request (24×7 at additional cost).

A switch is straightforward thanks to standardised APIs and containerised setup structures. We actively support you through the transition to guarantee minimal downtime and maximum data security.

You pay a fixed monthly price for clearly defined resources, meaning no variable, usage-based fees. So there are no extra costs driven by query or data volume. You can also adjust your package to your needs at the end of each month (scale up or down).

Build your combination

Choose your preferred cloud, and together we build the technology bridge for your enterprise workloads. Let us clarify the details in a short call.