Tag:Apache Iceberg
Articles tagged "Apache Iceberg", page 2.
- 30 MIN READ•Sep 10, 2026
What to Assert When You Test an Iceberg Pipeline
Fixtures, in-memory catalogs, and golden metadata: the assertions that catch wrong rows, unsafe reruns, schema drift, and concurrent-write corruption in CI.
Apache IcebergtestingCI - 31 MIN READ•Sep 10, 2026
The 2026 Iceberg REST Catalog Compatibility Report
A repeatable test for what an Iceberg REST catalog actually serves, a scoring scheme that separates design from breakage, and the 2026 evidence across seven catalogs.
Iceberg REST catalogcatalog compatibilityApache Polaris - 31 MIN READ•Sep 10, 2026
Serving Iceberg Tables From Two Regions
Three multi-region topologies that work and one that mostly does not, what an Iceberg commit costs across regions, and where the catalog has to live.
Apache Icebergmulti-regionreplication - 31 MIN READ•Sep 10, 2026
How Iceberg Catalogs Hand Engines Storage Access
Credential vending end to end: the wire protocol, scoped access on each cloud, remote signing, credential lifetime on long jobs, and failures that look like bugs.
credential vendingApache IcebergApache Polaris - 31 MIN READ•Sep 10, 2026
Kafka Connect to Iceberg: How the Commit Actually Works
Exactly-once semantics in the Iceberg sink connector: the coordinator, the control topic, offsets stored inside Iceberg snapshots, and where duplicates still get in.
Apache KafkaKafka ConnectApache Iceberg - 31 MIN READ•Sep 10, 2026
The Open Lakehouse Explained, Then Built on Your Laptop with Dremio and MinIO
The five layers of the open lakehouse explained, then a lab: Parquet, Iceberg, Polaris, Arrow, and Ossie running in two containers on your own machine.
open lakehouseApache IcebergApache Polaris - 31 MIN READ•Sep 10, 2026
Where Lock-In Went
The format war ended and exit cost did not: where lock-in relocated after open tables won, how to measure it, and which costs are worth keeping down.
vendor lock-inopen formatsApache Iceberg - 30 MIN READ•Sep 2, 2026
dbt on Iceberg: Incremental Models on Open Tables
How dbt incremental materializations map to Iceberg operations, and the configuration, predicates, and maintenance that keep them healthy.
dbtApache IcebergIncremental Models - 30 MIN READ•Sep 2, 2026
Disaster Recovery for Iceberg Tables: Replication, Backup, and Restore
Disaster recovery for Iceberg across four tiers: snapshots, object versioning, catalog backup, and cross-region replication.
Apache IcebergDisaster RecoveryReplication - 30 MIN READ•Sep 2, 2026
Deleting User Data From an Immutable Lakehouse: GDPR Hard Deletes on Iceberg
How to turn a logical delete on immutable Iceberg into a physical erasure across snapshots, versions, replicas, and downstream copies.
GDPRApache IcebergData Privacy - 32 MIN READ•Sep 2, 2026
Geospatial Data in Apache Iceberg: Geometry, Geography, and GeoParquet
How Iceberg v3 geometry and geography types, bounding boxes, and native Parquet types give spatial data first-class standing.
Apache IcebergGeospatialGeoParquet - 31 MIN READ•Sep 2, 2026
Default Column Values and Field IDs: How Iceberg Schema Evolution Works at the Spec Level
How field IDs and initial and write defaults let Iceberg change schemas on large tables without rewriting data, at the spec level.
Apache IcebergSchema EvolutionField IDs