Amazon Aurora PostgreSQL now supports direct querying of Apache Iceberg and Parquet data in your data lake
AWS News Blog
This article announces that Amazon Aurora PostgreSQL now supports direct querying of Apache Iceberg and Parquet data stored in data lakes, eliminating the need for ETL pipelines.
- Query operational data in Aurora alongside data lake data in a single query using PostgreSQL syntax
- Supported on Aurora PostgreSQL versions 17.11+ and 18.6+; enable via aurora_analytics extension
- DuckDB embedded in Aurora handles analytical scans of data lake files efficiently
- Supports querying Iceberg tables from AWS Glue Data Catalog and IRC-compatible catalogs
- Automatic schema inference from Parquet and Iceberg metadata eliminates manual column definition
- Query optimizations include predicate pushdown, column pruning, and caching of frequently accessed data
- Available in all commercial AWS Regions and GovCloud at no additional charge
Aurora PostgreSQL users can now combine live transactional data with historical data lake records through a single familiar interface, simplifying application development and reducing infrastructure complexity.
The AWS News Feed is currently looking for gold sponsors. If you want to support the AWS community and reach a large audience of AWS professionals, consider sponsoring the AWS News Feed.
Related articles
2026
2025
2026
2026
The AWS News Feed is currently looking for silver sponsors. If you want to support the AWS community and reach a large audience of AWS professionals, consider sponsoring the AWS News Feed.