Cross-source analytics
Join your managed PostgreSQL with your Iceberg tables in a single query
Multiple sources, one SQL query
Break down organizational silos
Transform your scattered data sources into unified analytics power. Built on Trino and Apache Iceberg, Data Hub federates your managed PostgreSQL and your Iceberg tables on S3 object storage into a single queryable base, and is fed through the Open Data, Google Drive and Email connectors.
Join your managed PostgreSQL with your Iceberg tables in a single query
Multiple sources, one SQL query
Advanced analytics on distributed datasets
No ETL pipeline to maintain
Open data into Iceberg, Google Drive files and email attachments into S3: connectors feed your queryable lakehouse
No ingestion code to write
An analyst needs to cross customer data stored in PostgreSQL with product events stored in Iceberg tables
A multi-source join with no prior extraction and no intermediate copy
A finance team consolidates its monthly reporting from its PostgreSQL database and its Iceberg tables
Consolidated reporting, queryable in SQL, with no pipeline to maintain
A data team feeds its lakehouse with open data datasets and collects its business files from Google Drive and an email inbox
A lakehouse fed from the console, with no ingestion code to write
Distributed and stateless SQL engine. Query your connected sources with standard SQL, without moving your data.
Open source table format for data lakes: ACID transactions, schema evolution, compaction and snapshot expiry driven from the console.
Open data, files and emails feed your Iceberg tables and S3 buckets with no pipeline to write. More federation sources are coming soon.
Discover how Data Hub can eliminate your data silos.