682 tutorials · free to read
A working handbook for data and software engineering.
Hands-on tutorials by Alex Merced on Apache Iceberg, the data lakehouse, pipelines, agentic AI, and the languages and tools that hold it all together.
Browse by topic
All topics- Data Lakehouse201
- Data Engineering177
- Apache Iceberg142
- Dremio65
- AI Agents49
- developer tools33
- AI coding tools31
- agentic development31
- Databases28
- MCP25
- Data Architecture24
- AI23
- Apache Polaris22
- connectors21
Latest tutorials
Page 35 of 114- 01What Iceberg V3 Advances Mean for CDC PipelinesApache Iceberg V3 brings deletion vectors and row lineage that reshape CDC pipeline design. Learn what these features mean for your streaming data architecture.Apache IcebergMay 24, 2026
- 02Kafka 4.0 Changes Streaming Platform OperationsKafka 4.0 removes ZooKeeper and ships KRaft and KIP-848 by default. Learn what those changes mean for platform operations, upgrades, and client configurations.Data EngineeringMay 24, 2026
- 03Lance and Iceberg for Multimodal AI DataLanceDB and Apache Iceberg serve complementary roles in a multimodal AI lakehouse. Learn when to use Lance for embeddings and random access, and Iceberg for structured metadata and SQL analytics.Apache IcebergMay 24, 2026
- 04Bringing MLflow and Data Pipelines Closer TogetherMLflow 3 extends observability from classic ML experiments to GenAI tracing and data pipeline lineage. Learn how to connect data quality monitoring with model performance tracking.Data EngineeringMay 24, 2026
- 05Modern Feature Stores Beyond Batch PipelinesFeature stores like Feast now support streaming feature views from Kafka and Kinesis alongside batch pipelines. Learn how to build real-time features that maintain training-serving consistency.Data LakehouseMay 24, 2026
- 06OpenLineage as the Spine of Data ObservabilityOpenLineage provides a standard API for collecting pipeline lineage across Airflow, Spark, Flink, and dbt. Learn how it powers blast radius analysis and incident triage.Data LakehouseMay 24, 2026