Managed Data Lakehouse: sovereign & open source | Stackable
Managed Stackable
Data Lakehouse
The enterprise data lakehouse as a managed service, built on the established tools of the Stackable Data Platform (SDP):
Sovereign · Compliant · Open Source
- The end of data silos
- All the features for your data management
- Full digital sovereignty
- One foundation, many use cases
- Fully managed, no operational overhead
Stackable - your sovereign data platform
From data silos to a central data platform
The Managed Stackable Data Lakehouse is designed for analytical enterprise workloads: sovereign data management independent of industry and company size.
Central data platform
Consolidate data from different source systems (ERP, CRM, machine data, Excel lists) into a single lakehouse, for consistent reporting and self-service analytics instead of manual Excel consolidation.
Data foundation for AI and ML projects
Create a clean, consolidated data foundation on which machine-learning models or AI applications can be built, without expensive parallel infrastructure.
Compliance and auditability
Keep historical data states traceable and audit-proof (for example for GDPR requests, financial audits or certifications) through the time-travel capabilities of the lakehouse.
Customer and sales analytics
Bring together order histories, customer behaviour and sales data to spot cross-selling potential and customer churn early.
Predictive maintenance
Analyse machine and sensor data to reduce downtime and maintenance costs, developing and applying specialised models, including in real time.
Supply chain transparency
Combine data from procurement, logistics and production to identify bottlenecks, supplier dependencies and data silos early.
The open source data lakehouse
All modules of the Managed Stackable Data Lakehouse are stable, performant and fully managed – future-proof by design.
Trino
Fast, distributed SQL query engine for interactive ad-hoc analytics directly on your data lakes, with no data movement.
Apache Iceberg
Open, high-performance table format for ACID transactions, time travel and cross-schema data consistency on S3 storage.
Apache Superset
Interactive dashboards and ad-hoc data exploration for business users.
Apache NiFi
Visual dataflow design for ingestion, transformation and routing.
Apache Spark
The proven, high-speed compute engine for demanding batch processing, ETL pipelines and machine learning.
Apache Airflow
Scalable workflow manager for orchestration, scheduling and complete monitoring of all data streams.
OPA & OIDC
Fine-grained authorization policies and single sign-on across all components.
Keycloak & OpenID
Single sign-on (SSO) and secure user authentication.
Configure your Managed Service
The Stackable Data Lakehouse in different sizes: choose your infrastructure with the right service partner.
Stackable Data Platform · Trino · Apache Spark · Apache Iceberg · Apache Airflow · Apache Superset · Apache NiFi
Add Red Hat OpenShift
Optional: enterprise container orchestration & access management.
* The first 30 days are free. From day 31 onward, 449 € per month (without support and SLA).
+
+
Custom
Frequently asked questions
Here you will find quick answers to the most important questions about our flexible multi-cloud combinations and managed infrastructures.
What is a managed data lakehouse and who is it for?
A managed data lakehouse is a ready-to-run data platform that combines a data lake and a data warehouse, with a partner running it for you. It is particularly suited to organizations with large data volumes that do not want to build their own platform ops capacity.
How does an open source data lakehouse avoid vendor lock-in?
It avoids vendor lock-in through open formats (Apache Iceberg on S3) and a platform defined as code, portable via GitOps and built from open source modules. Cloud, managed partner and components can all be swapped.
Does my data stay GDPR-compliant and in Europe?
Yes. Choose the infrastructure yourself, for example a cloud service provider under German jurisdiction, a private cloud or your own data centre. All hosting partners and data centres are ISO 27001 certified and operate in full compliance with the European General Data Protection Regulation (GDPR). Secure access control via Open Policy Agent and Keycloak, together with an audit-proof history through Apache Iceberg, supports regulated industries (for example DORA in the financial sector).
Do I need my own operations team for a managed data lakehouse?
No. In the managed model, the partner handles provisioning, operations, monitoring and updates on Kubernetes, so you can focus on your data applications.
Do business users need to know Kubernetes for data analytics?
No. Business users work through the Stackable Cockpit without any contact with the underlying infrastructure, using Apache Superset for dashboards and ad-hoc analysis, Apache Airflow for orchestration, Apache NiFi for data pipelines, plus a Stackable query editor for Trino and an S3 manager.
What is included in the price?
The monthly price bundles the SDP subscription, the managed service and the cloud infrastructure with SLA support. This includes running the platform, covering configuration, monitoring, patches and updates.
How do the XS, S, M and Individual packages differ?
The packages differ in compute and storage capacity (nodes, vCPUs, RAM, S3), while the feature set stays the same. If the standard packages are not enough, “Individual” sizes the platform to your specific needs. The XS package is the entry point with minimal resources for tests, proofs of concept and small workloads, without support or SLA. For production use, choose a paid package.
Are there guaranteed service level agreements (SLAs)?
Yes, we agree binding service level agreements tailored precisely to your business-critical workloads. This includes availability guarantees of up to 99.99 % and dedicated support (8×5), and around the clock on request (24×7 at additional cost).
Can I switch partner or cloud later?
A switch is straightforward thanks to standardised APIs and containerised setup structures. We actively support you through the transition to guarantee minimal downtime and maximum data security.
How predictable are costs compared to Databricks or Snowflake?
You pay a fixed monthly price for clearly defined resources, meaning no variable, usage-based fees. So there are no extra costs driven by query or data volume. You can also adjust your package to your needs at the end of each month (scale up or down).
Build your combination
Choose your preferred cloud, and together we build the technology bridge for your enterprise workloads. Let us clarify the details in a short call.