Cloudflare announced the general availability of Cloudflare Basin, a serverless data analytics platform built on Apache Iceberg and R2 Object Storage. Basin includes Pipelines for data ingestion and transformation, Catalog for metadata management, and SQL for querying, enabling developers to collect and analyze data from multiple sources with speed, openness, and cost efficiency.
Snowflake open sourced pg_lake, a set of PostgreSQL extensions that enable Postgres to query, manage, and write to Iceberg tables and data lakehouse files using standard SQL. The technology was developed by Crunchy Data over several years and is now available under an Apache license to benefit the broader Postgres community.
Amazon Aurora PostgreSQL now supports direct querying of Apache Iceberg and Parquet data in data lakes without ETL pipelines, using embedded DuckDB technology. This capability enables applications to combine operational data with data lake records through familiar PostgreSQL syntax and tools, reducing infrastructure complexity and supporting use cases like real-time dashboards and AI agents.
Aurora PostgreSQL now enables direct querying of Apache Iceberg and Parquet data stored in data lakes without ETL pipelines, using DuckDB's query engine embedded in PostgreSQL. Users can create foreign tables referencing data in Amazon S3 or AWS Glue Data Catalog, allowing existing PostgreSQL applications to access data lake information while optionally materializing it into native Aurora tables.
Amazon Aurora PostgreSQL now supports direct querying of Apache Iceberg and Parquet data stored in data lakes, eliminating the need for ETL pipelines. DuckDB has been embedded into Aurora to enable single queries combining live operational data with historical data lake records using familiar PostgreSQL syntax. The feature is available on Aurora PostgreSQL 17.11+ and 18.6+.