Curriculum planned

Stage 10 · Lakehouse Table Formats

Apache Hudi

A comprehensive Apache Hudi course covering timeline-based table state, Copy-on-Write and Merge-on-Read, record keys and preCombine semantics, indexing, metadata table, upserts/deletes, incremental and CDC queries, compaction, clustering, cleaning, concurrency control, Spark/Flink streaming, table services, catalogs, object storage, AI/multimodal types, Lance/vector capabilities, security, observability, migration, and production operations.

33planned chapters
165reserved lesson paths
Advancedlearning level
Plannedcourse state
Coverage baselineApache Hudi 1.2.0 baseline (May 2026) with table version 9, Copy-on-Write and Merge-on-Read, timeline and instant semantics, record-level and secondary indexes, metadata table, concurrency control, incremental and CDC queries, Spark/Flink integrations, table services, clustering/compaction/cleaning, VECTOR/VARIANT/BLOB types, Lance support, vector search, and production lakehouse operations

Course brief

Understand Hudi as a transactional mutation and incremental-processing system for data lakes: model records, timeline instants, file groups, indexes, table services, concurrency, and incremental consumption together rather than treating it as merely another storage format.

A comprehensive Apache Hudi course covering timeline-based table state, Copy-on-Write and Merge-on-Read, record keys and preCombine semantics, indexing, metadata table, upserts/deletes, incremental and CDC queries, compaction, clustering, cleaning, concurrency control, Spark/Flink streaming, table services, catalogs, object storage, AI/multimodal types, Lance/vector capabilities, security, observability, migration, and production operations.

This syllabus deliberately separates foundations, data/model semantics, internals, reliability, security, performance, operations, and production design so advanced material is not compressed into generic catch-all chapters.

By the end

You will be able to

  • Explain Hudi timeline/instants, file groups/slices, base and log files, Copy-on-Write vs Merge-on-Read, record keys, ordering/preCombine, and table-version semantics
  • Design scalable upsert/delete workloads with bucket/record-level/secondary indexes, metadata table, concurrency control, compaction, clustering and cleaning
  • Build batch and streaming ingestion with Spark and Flink and consume snapshot, read-optimized, incremental and CDC query modes correctly
  • Use Hudi 1.2 capabilities including table version 9, native VECTOR/VARIANT/BLOB types, Lance integration and vector-search concepts while respecting engine compatibility
  • Operate secure production Hudi tables with table services, observability, backup/recovery, upgrades, migration, cost/performance benchmarks, and failure drills

Complete planned syllabus

33 chapters · 165 lesson paths.

Every lesson path is reserved now but intentionally not linked until its lesson HTML is actually published. The sequence moves from foundations through advanced implementation, architecture, operations, reliability, security, tuning, and a production capstone.

01

Chapter 1

Hudi Foundations: Mutable Lakehouse Tables, Apache Hudi 1.2.0, Table Version 9, Workload Fit, and Local Spark/Flink Lab

5 lessons
01
Hudi Foundations: Mutable Lakehouse Tables, Apache Hudi 1.2.0, Table Version 9, Workload Fit, and Local Spark/Flink Lab: Concepts, Terminology, Architecture, and Mental ModelPlanned lesson · reserved path Chapter01/Lesson1.html
Planned
02
Hudi Foundations: Mutable Lakehouse Tables, Apache Hudi 1.2.0, Table Version 9, Workload Fit, and Local Spark/Flink Lab: Guided Hands-On Workflow, Commands, APIs, and ConfigurationPlanned lesson · reserved path Chapter01/Lesson2.html
Planned
03
Hudi Foundations: Mutable Lakehouse Tables, Apache Hudi 1.2.0, Table Version 9, Workload Fit, and Local Spark/Flink Lab: Design Choices, Scaling Behavior, Compatibility, and TradeoffsPlanned lesson · reserved path Chapter01/Lesson3.html
Planned
04
Hudi Foundations: Mutable Lakehouse Tables, Apache Hudi 1.2.0, Table Version 9, Workload Fit, and Local Spark/Flink Lab: Failure Modes, Diagnostics, Security, Reliability, and PerformancePlanned lesson · reserved path Chapter01/Lesson4.html
Planned
05
Checkpoint Lab — Hudi Foundations: Mutable Lakehouse Tables, Apache Hudi 1.2.0, Table Version 9, Workload Fit, and Local Spark/Flink Lab: Build, Test, Measure, Troubleshoot, and Explain the ResultPlanned lesson · reserved path Chapter01/Lesson5.html
Planned
02

Chapter 2

Timeline and Instant Semantics: Requested/In-Flight/Completed States, Commit Types, Timeline Ordering, Archival, and Table State

5 lessons
01
Timeline and Instant Semantics: Requested/In-Flight/Completed States, Commit Types, Timeline Ordering, Archival, and Table State: Concepts, Terminology, Architecture, and Mental ModelPlanned lesson · reserved path Chapter02/Lesson1.html
Planned
02
Timeline and Instant Semantics: Requested/In-Flight/Completed States, Commit Types, Timeline Ordering, Archival, and Table State: Guided Hands-On Workflow, Commands, APIs, and ConfigurationPlanned lesson · reserved path Chapter02/Lesson2.html
Planned
03
Timeline and Instant Semantics: Requested/In-Flight/Completed States, Commit Types, Timeline Ordering, Archival, and Table State: Design Choices, Scaling Behavior, Compatibility, and TradeoffsPlanned lesson · reserved path Chapter02/Lesson3.html
Planned
04
Timeline and Instant Semantics: Requested/In-Flight/Completed States, Commit Types, Timeline Ordering, Archival, and Table State: Failure Modes, Diagnostics, Security, Reliability, and PerformancePlanned lesson · reserved path Chapter02/Lesson4.html
Planned
05
Checkpoint Lab — Timeline and Instant Semantics: Requested/In-Flight/Completed States, Commit Types, Timeline Ordering, Archival, and Table State: Build, Test, Measure, Troubleshoot, and Explain the ResultPlanned lesson · reserved path Chapter02/Lesson5.html
Planned
03

Chapter 3

File Groups and File Slices: File IDs, Base Files, Log Files, Versions, File-System Layout, and How Mutations Accumulate

5 lessons
01
File Groups and File Slices: File IDs, Base Files, Log Files, Versions, File-System Layout, and How Mutations Accumulate: Concepts, Terminology, Architecture, and Mental ModelPlanned lesson · reserved path Chapter03/Lesson1.html
Planned
02
File Groups and File Slices: File IDs, Base Files, Log Files, Versions, File-System Layout, and How Mutations Accumulate: Guided Hands-On Workflow, Commands, APIs, and ConfigurationPlanned lesson · reserved path Chapter03/Lesson2.html
Planned
03
File Groups and File Slices: File IDs, Base Files, Log Files, Versions, File-System Layout, and How Mutations Accumulate: Design Choices, Scaling Behavior, Compatibility, and TradeoffsPlanned lesson · reserved path Chapter03/Lesson3.html
Planned
04
File Groups and File Slices: File IDs, Base Files, Log Files, Versions, File-System Layout, and How Mutations Accumulate: Failure Modes, Diagnostics, Security, Reliability, and PerformancePlanned lesson · reserved path Chapter03/Lesson4.html
Planned
05
Checkpoint Lab — File Groups and File Slices: File IDs, Base Files, Log Files, Versions, File-System Layout, and How Mutations Accumulate: Build, Test, Measure, Troubleshoot, and Explain the ResultPlanned lesson · reserved path Chapter03/Lesson5.html
Planned
04

Chapter 4

Copy-on-Write Tables: Write Path, File Rewrites, Read Simplicity, Update Amplification, Latency, and Workload Selection

5 lessons
01
Copy-on-Write Tables: Write Path, File Rewrites, Read Simplicity, Update Amplification, Latency, and Workload Selection: Concepts, Terminology, Architecture, and Mental ModelPlanned lesson · reserved path Chapter04/Lesson1.html
Planned
02
Copy-on-Write Tables: Write Path, File Rewrites, Read Simplicity, Update Amplification, Latency, and Workload Selection: Guided Hands-On Workflow, Commands, APIs, and ConfigurationPlanned lesson · reserved path Chapter04/Lesson2.html
Planned
03
Copy-on-Write Tables: Write Path, File Rewrites, Read Simplicity, Update Amplification, Latency, and Workload Selection: Design Choices, Scaling Behavior, Compatibility, and TradeoffsPlanned lesson · reserved path Chapter04/Lesson3.html
Planned
04
Copy-on-Write Tables: Write Path, File Rewrites, Read Simplicity, Update Amplification, Latency, and Workload Selection: Failure Modes, Diagnostics, Security, Reliability, and PerformancePlanned lesson · reserved path Chapter04/Lesson4.html
Planned
05
Checkpoint Lab — Copy-on-Write Tables: Write Path, File Rewrites, Read Simplicity, Update Amplification, Latency, and Workload Selection: Build, Test, Measure, Troubleshoot, and Explain the ResultPlanned lesson · reserved path Chapter04/Lesson5.html
Planned
05

Chapter 5

Merge-on-Read Tables: Delta Commits, Log Blocks, Merging, Compaction, Read-Optimized vs Snapshot Views, and Freshness Tradeoffs

5 lessons
01
Merge-on-Read Tables: Delta Commits, Log Blocks, Merging, Compaction, Read-Optimized vs Snapshot Views, and Freshness Tradeoffs: Concepts, Terminology, Architecture, and Mental ModelPlanned lesson · reserved path Chapter05/Lesson1.html
Planned
02
Merge-on-Read Tables: Delta Commits, Log Blocks, Merging, Compaction, Read-Optimized vs Snapshot Views, and Freshness Tradeoffs: Guided Hands-On Workflow, Commands, APIs, and ConfigurationPlanned lesson · reserved path Chapter05/Lesson2.html
Planned
03
Merge-on-Read Tables: Delta Commits, Log Blocks, Merging, Compaction, Read-Optimized vs Snapshot Views, and Freshness Tradeoffs: Design Choices, Scaling Behavior, Compatibility, and TradeoffsPlanned lesson · reserved path Chapter05/Lesson3.html
Planned
04
Merge-on-Read Tables: Delta Commits, Log Blocks, Merging, Compaction, Read-Optimized vs Snapshot Views, and Freshness Tradeoffs: Failure Modes, Diagnostics, Security, Reliability, and PerformancePlanned lesson · reserved path Chapter05/Lesson4.html
Planned
05
Checkpoint Lab — Merge-on-Read Tables: Delta Commits, Log Blocks, Merging, Compaction, Read-Optimized vs Snapshot Views, and Freshness Tradeoffs: Build, Test, Measure, Troubleshoot, and Explain the ResultPlanned lesson · reserved path Chapter05/Lesson5.html
Planned
06

Chapter 6

Record Keys, Partition Paths, and Ordering: Primary Identity, Key Generators, Composite Keys, preCombine/Ordering Fields, and Duplicate Resolution

5 lessons
01
Record Keys, Partition Paths, and Ordering: Primary Identity, Key Generators, Composite Keys, preCombine/Ordering Fields, and Duplicate Resolution: Concepts, Terminology, Architecture, and Mental ModelPlanned lesson · reserved path Chapter06/Lesson1.html
Planned
02
Record Keys, Partition Paths, and Ordering: Primary Identity, Key Generators, Composite Keys, preCombine/Ordering Fields, and Duplicate Resolution: Guided Hands-On Workflow, Commands, APIs, and ConfigurationPlanned lesson · reserved path Chapter06/Lesson2.html
Planned
03
Record Keys, Partition Paths, and Ordering: Primary Identity, Key Generators, Composite Keys, preCombine/Ordering Fields, and Duplicate Resolution: Design Choices, Scaling Behavior, Compatibility, and TradeoffsPlanned lesson · reserved path Chapter06/Lesson3.html
Planned
04
Record Keys, Partition Paths, and Ordering: Primary Identity, Key Generators, Composite Keys, preCombine/Ordering Fields, and Duplicate Resolution: Failure Modes, Diagnostics, Security, Reliability, and PerformancePlanned lesson · reserved path Chapter06/Lesson4.html
Planned
05
Checkpoint Lab — Record Keys, Partition Paths, and Ordering: Primary Identity, Key Generators, Composite Keys, preCombine/Ordering Fields, and Duplicate Resolution: Build, Test, Measure, Troubleshoot, and Explain the ResultPlanned lesson · reserved path Chapter06/Lesson5.html
Planned
07

Chapter 7

Write Operations: Insert, Bulk Insert, Upsert, Delete, Insert Overwrite, Delete Partition, Operation Semantics, and Idempotency

5 lessons
01
Write Operations: Insert, Bulk Insert, Upsert, Delete, Insert Overwrite, Delete Partition, Operation Semantics, and Idempotency: Concepts, Terminology, Architecture, and Mental ModelPlanned lesson · reserved path Chapter07/Lesson1.html
Planned
02
Write Operations: Insert, Bulk Insert, Upsert, Delete, Insert Overwrite, Delete Partition, Operation Semantics, and Idempotency: Guided Hands-On Workflow, Commands, APIs, and ConfigurationPlanned lesson · reserved path Chapter07/Lesson2.html
Planned
03
Write Operations: Insert, Bulk Insert, Upsert, Delete, Insert Overwrite, Delete Partition, Operation Semantics, and Idempotency: Design Choices, Scaling Behavior, Compatibility, and TradeoffsPlanned lesson · reserved path Chapter07/Lesson3.html
Planned
04
Write Operations: Insert, Bulk Insert, Upsert, Delete, Insert Overwrite, Delete Partition, Operation Semantics, and Idempotency: Failure Modes, Diagnostics, Security, Reliability, and PerformancePlanned lesson · reserved path Chapter07/Lesson4.html
Planned
05
Checkpoint Lab — Write Operations: Insert, Bulk Insert, Upsert, Delete, Insert Overwrite, Delete Partition, Operation Semantics, and Idempotency: Build, Test, Measure, Troubleshoot, and Explain the ResultPlanned lesson · reserved path Chapter07/Lesson5.html
Planned
08

Chapter 8

Indexing Foundations: Global vs Non-Global, Simple/Bloom/Bucket Concepts, Lookup Cost, Partition Movement, and Choosing an Index

5 lessons
01
Indexing Foundations: Global vs Non-Global, Simple/Bloom/Bucket Concepts, Lookup Cost, Partition Movement, and Choosing an Index: Concepts, Terminology, Architecture, and Mental ModelPlanned lesson · reserved path Chapter08/Lesson1.html
Planned
02
Indexing Foundations: Global vs Non-Global, Simple/Bloom/Bucket Concepts, Lookup Cost, Partition Movement, and Choosing an Index: Guided Hands-On Workflow, Commands, APIs, and ConfigurationPlanned lesson · reserved path Chapter08/Lesson2.html
Planned
03
Indexing Foundations: Global vs Non-Global, Simple/Bloom/Bucket Concepts, Lookup Cost, Partition Movement, and Choosing an Index: Design Choices, Scaling Behavior, Compatibility, and TradeoffsPlanned lesson · reserved path Chapter08/Lesson3.html
Planned
04
Indexing Foundations: Global vs Non-Global, Simple/Bloom/Bucket Concepts, Lookup Cost, Partition Movement, and Choosing an Index: Failure Modes, Diagnostics, Security, Reliability, and PerformancePlanned lesson · reserved path Chapter08/Lesson4.html
Planned
05
Checkpoint Lab — Indexing Foundations: Global vs Non-Global, Simple/Bloom/Bucket Concepts, Lookup Cost, Partition Movement, and Choosing an Index: Build, Test, Measure, Troubleshoot, and Explain the ResultPlanned lesson · reserved path Chapter08/Lesson5.html
Planned
09

Chapter 9

Record-Level Index: Metadata-Backed Global Location, Spark/Flink Use, Streaming Global Upserts, Scale Characteristics, and Recovery

5 lessons
01
Record-Level Index: Metadata-Backed Global Location, Spark/Flink Use, Streaming Global Upserts, Scale Characteristics, and Recovery: Concepts, Terminology, Architecture, and Mental ModelPlanned lesson · reserved path Chapter09/Lesson1.html
Planned
02
Record-Level Index: Metadata-Backed Global Location, Spark/Flink Use, Streaming Global Upserts, Scale Characteristics, and Recovery: Guided Hands-On Workflow, Commands, APIs, and ConfigurationPlanned lesson · reserved path Chapter09/Lesson2.html
Planned
03
Record-Level Index: Metadata-Backed Global Location, Spark/Flink Use, Streaming Global Upserts, Scale Characteristics, and Recovery: Design Choices, Scaling Behavior, Compatibility, and TradeoffsPlanned lesson · reserved path Chapter09/Lesson3.html
Planned
04
Record-Level Index: Metadata-Backed Global Location, Spark/Flink Use, Streaming Global Upserts, Scale Characteristics, and Recovery: Failure Modes, Diagnostics, Security, Reliability, and PerformancePlanned lesson · reserved path Chapter09/Lesson4.html
Planned
05
Checkpoint Lab — Record-Level Index: Metadata-Backed Global Location, Spark/Flink Use, Streaming Global Upserts, Scale Characteristics, and Recovery: Build, Test, Measure, Troubleshoot, and Explain the ResultPlanned lesson · reserved path Chapter09/Lesson5.html
Planned
10

Chapter 10

Secondary Indexes: Column/Expression-Like Lookup Concepts, Metadata Encoding, Query/Write Tradeoffs, Lifecycle, and Table Version 9 Considerations

5 lessons
01
Secondary Indexes: Column/Expression-Like Lookup Concepts, Metadata Encoding, Query/Write Tradeoffs, Lifecycle, and Table Version 9 Considerations: Concepts, Terminology, Architecture, and Mental ModelPlanned lesson · reserved path Chapter10/Lesson1.html
Planned
02
Secondary Indexes: Column/Expression-Like Lookup Concepts, Metadata Encoding, Query/Write Tradeoffs, Lifecycle, and Table Version 9 Considerations: Guided Hands-On Workflow, Commands, APIs, and ConfigurationPlanned lesson · reserved path Chapter10/Lesson2.html
Planned
03
Secondary Indexes: Column/Expression-Like Lookup Concepts, Metadata Encoding, Query/Write Tradeoffs, Lifecycle, and Table Version 9 Considerations: Design Choices, Scaling Behavior, Compatibility, and TradeoffsPlanned lesson · reserved path Chapter10/Lesson3.html
Planned
04
Secondary Indexes: Column/Expression-Like Lookup Concepts, Metadata Encoding, Query/Write Tradeoffs, Lifecycle, and Table Version 9 Considerations: Failure Modes, Diagnostics, Security, Reliability, and PerformancePlanned lesson · reserved path Chapter10/Lesson4.html
Planned
05
Checkpoint Lab — Secondary Indexes: Column/Expression-Like Lookup Concepts, Metadata Encoding, Query/Write Tradeoffs, Lifecycle, and Table Version 9 Considerations: Build, Test, Measure, Troubleshoot, and Explain the ResultPlanned lesson · reserved path Chapter10/Lesson5.html
Planned
11

Chapter 11

Metadata Table: File Listings, Column Stats, Bloom Filters, Record Index, Secondary Metadata, Sync, Failure Recovery, and Scale

5 lessons
01
Metadata Table: File Listings, Column Stats, Bloom Filters, Record Index, Secondary Metadata, Sync, Failure Recovery, and Scale: Concepts, Terminology, Architecture, and Mental ModelPlanned lesson · reserved path Chapter11/Lesson1.html
Planned
02
Metadata Table: File Listings, Column Stats, Bloom Filters, Record Index, Secondary Metadata, Sync, Failure Recovery, and Scale: Guided Hands-On Workflow, Commands, APIs, and ConfigurationPlanned lesson · reserved path Chapter11/Lesson2.html
Planned
03
Metadata Table: File Listings, Column Stats, Bloom Filters, Record Index, Secondary Metadata, Sync, Failure Recovery, and Scale: Design Choices, Scaling Behavior, Compatibility, and TradeoffsPlanned lesson · reserved path Chapter11/Lesson3.html
Planned
04
Metadata Table: File Listings, Column Stats, Bloom Filters, Record Index, Secondary Metadata, Sync, Failure Recovery, and Scale: Failure Modes, Diagnostics, Security, Reliability, and PerformancePlanned lesson · reserved path Chapter11/Lesson4.html
Planned
05
Checkpoint Lab — Metadata Table: File Listings, Column Stats, Bloom Filters, Record Index, Secondary Metadata, Sync, Failure Recovery, and Scale: Build, Test, Measure, Troubleshoot, and Explain the ResultPlanned lesson · reserved path Chapter11/Lesson5.html
Planned
12

Chapter 12

Schema Management: Avro-Oriented Schema Evolution, Nullability, Type Promotion, Reconciliation, Internal Schema, Compatibility, and Rollout Strategy

5 lessons
01
Schema Management: Avro-Oriented Schema Evolution, Nullability, Type Promotion, Reconciliation, Internal Schema, Compatibility, and Rollout Strategy: Concepts, Terminology, Architecture, and Mental ModelPlanned lesson · reserved path Chapter12/Lesson1.html
Planned
02
Schema Management: Avro-Oriented Schema Evolution, Nullability, Type Promotion, Reconciliation, Internal Schema, Compatibility, and Rollout Strategy: Guided Hands-On Workflow, Commands, APIs, and ConfigurationPlanned lesson · reserved path Chapter12/Lesson2.html
Planned
03
Schema Management: Avro-Oriented Schema Evolution, Nullability, Type Promotion, Reconciliation, Internal Schema, Compatibility, and Rollout Strategy: Design Choices, Scaling Behavior, Compatibility, and TradeoffsPlanned lesson · reserved path Chapter12/Lesson3.html
Planned
04
Schema Management: Avro-Oriented Schema Evolution, Nullability, Type Promotion, Reconciliation, Internal Schema, Compatibility, and Rollout Strategy: Failure Modes, Diagnostics, Security, Reliability, and PerformancePlanned lesson · reserved path Chapter12/Lesson4.html
Planned
05
Checkpoint Lab — Schema Management: Avro-Oriented Schema Evolution, Nullability, Type Promotion, Reconciliation, Internal Schema, Compatibility, and Rollout Strategy: Build, Test, Measure, Troubleshoot, and Explain the ResultPlanned lesson · reserved path Chapter12/Lesson5.html
Planned
13

Chapter 13

Table Version 8 vs 9: Storage Format Capabilities, Compatibility, Upgrades, New Types/Indexes, Reader Constraints, and Safe Fleet Migration

5 lessons
01
Table Version 8 vs 9: Storage Format Capabilities, Compatibility, Upgrades, New Types/Indexes, Reader Constraints, and Safe Fleet Migration: Concepts, Terminology, Architecture, and Mental ModelPlanned lesson · reserved path Chapter13/Lesson1.html
Planned
02
Table Version 8 vs 9: Storage Format Capabilities, Compatibility, Upgrades, New Types/Indexes, Reader Constraints, and Safe Fleet Migration: Guided Hands-On Workflow, Commands, APIs, and ConfigurationPlanned lesson · reserved path Chapter13/Lesson2.html
Planned
03
Table Version 8 vs 9: Storage Format Capabilities, Compatibility, Upgrades, New Types/Indexes, Reader Constraints, and Safe Fleet Migration: Design Choices, Scaling Behavior, Compatibility, and TradeoffsPlanned lesson · reserved path Chapter13/Lesson3.html
Planned
04
Table Version 8 vs 9: Storage Format Capabilities, Compatibility, Upgrades, New Types/Indexes, Reader Constraints, and Safe Fleet Migration: Failure Modes, Diagnostics, Security, Reliability, and PerformancePlanned lesson · reserved path Chapter13/Lesson4.html
Planned
05
Checkpoint Lab — Table Version 8 vs 9: Storage Format Capabilities, Compatibility, Upgrades, New Types/Indexes, Reader Constraints, and Safe Fleet Migration: Build, Test, Measure, Troubleshoot, and Explain the ResultPlanned lesson · reserved path Chapter13/Lesson5.html
Planned
14

Chapter 14

Snapshot and Read-Optimized Queries: Latest Table State, MOR Merge Cost, Query Paths, Engine Semantics, and Analytical Tradeoffs

5 lessons
01
Snapshot and Read-Optimized Queries: Latest Table State, MOR Merge Cost, Query Paths, Engine Semantics, and Analytical Tradeoffs: Concepts, Terminology, Architecture, and Mental ModelPlanned lesson · reserved path Chapter14/Lesson1.html
Planned
02
Snapshot and Read-Optimized Queries: Latest Table State, MOR Merge Cost, Query Paths, Engine Semantics, and Analytical Tradeoffs: Guided Hands-On Workflow, Commands, APIs, and ConfigurationPlanned lesson · reserved path Chapter14/Lesson2.html
Planned
03
Snapshot and Read-Optimized Queries: Latest Table State, MOR Merge Cost, Query Paths, Engine Semantics, and Analytical Tradeoffs: Design Choices, Scaling Behavior, Compatibility, and TradeoffsPlanned lesson · reserved path Chapter14/Lesson3.html
Planned
04
Snapshot and Read-Optimized Queries: Latest Table State, MOR Merge Cost, Query Paths, Engine Semantics, and Analytical Tradeoffs: Failure Modes, Diagnostics, Security, Reliability, and PerformancePlanned lesson · reserved path Chapter14/Lesson4.html
Planned
05
Checkpoint Lab — Snapshot and Read-Optimized Queries: Latest Table State, MOR Merge Cost, Query Paths, Engine Semantics, and Analytical Tradeoffs: Build, Test, Measure, Troubleshoot, and Explain the ResultPlanned lesson · reserved path Chapter14/Lesson5.html
Planned
15

Chapter 15

Incremental Queries: Begin/End Instants, Changed Records, Hollow Commits, Consumer Checkpoints, Reprocessing, and Downstream Pipelines

5 lessons
01
Incremental Queries: Begin/End Instants, Changed Records, Hollow Commits, Consumer Checkpoints, Reprocessing, and Downstream Pipelines: Concepts, Terminology, Architecture, and Mental ModelPlanned lesson · reserved path Chapter15/Lesson1.html
Planned
02
Incremental Queries: Begin/End Instants, Changed Records, Hollow Commits, Consumer Checkpoints, Reprocessing, and Downstream Pipelines: Guided Hands-On Workflow, Commands, APIs, and ConfigurationPlanned lesson · reserved path Chapter15/Lesson2.html
Planned
03
Incremental Queries: Begin/End Instants, Changed Records, Hollow Commits, Consumer Checkpoints, Reprocessing, and Downstream Pipelines: Design Choices, Scaling Behavior, Compatibility, and TradeoffsPlanned lesson · reserved path Chapter15/Lesson3.html
Planned
04
Incremental Queries: Begin/End Instants, Changed Records, Hollow Commits, Consumer Checkpoints, Reprocessing, and Downstream Pipelines: Failure Modes, Diagnostics, Security, Reliability, and PerformancePlanned lesson · reserved path Chapter15/Lesson4.html
Planned
05
Checkpoint Lab — Incremental Queries: Begin/End Instants, Changed Records, Hollow Commits, Consumer Checkpoints, Reprocessing, and Downstream Pipelines: Build, Test, Measure, Troubleshoot, and Explain the ResultPlanned lesson · reserved path Chapter15/Lesson5.html
Planned
16

Chapter 16

Change Data Capture: CDC Logging, Before/After Semantics, Supplemental Logging Concepts, Incremental Consumers, Retention, and Replay

5 lessons
01
Change Data Capture: CDC Logging, Before/After Semantics, Supplemental Logging Concepts, Incremental Consumers, Retention, and Replay: Concepts, Terminology, Architecture, and Mental ModelPlanned lesson · reserved path Chapter16/Lesson1.html
Planned
02
Change Data Capture: CDC Logging, Before/After Semantics, Supplemental Logging Concepts, Incremental Consumers, Retention, and Replay: Guided Hands-On Workflow, Commands, APIs, and ConfigurationPlanned lesson · reserved path Chapter16/Lesson2.html
Planned
03
Change Data Capture: CDC Logging, Before/After Semantics, Supplemental Logging Concepts, Incremental Consumers, Retention, and Replay: Design Choices, Scaling Behavior, Compatibility, and TradeoffsPlanned lesson · reserved path Chapter16/Lesson3.html
Planned
04
Change Data Capture: CDC Logging, Before/After Semantics, Supplemental Logging Concepts, Incremental Consumers, Retention, and Replay: Failure Modes, Diagnostics, Security, Reliability, and PerformancePlanned lesson · reserved path Chapter16/Lesson4.html
Planned
05
Checkpoint Lab — Change Data Capture: CDC Logging, Before/After Semantics, Supplemental Logging Concepts, Incremental Consumers, Retention, and Replay: Build, Test, Measure, Troubleshoot, and Explain the ResultPlanned lesson · reserved path Chapter16/Lesson5.html
Planned
17

Chapter 17

Compaction: Scheduling vs Execution, Inline/Async Modes, Strategies, Backlog, Read/Write Amplification, Failure Recovery, and SLAs

5 lessons
01
Compaction: Scheduling vs Execution, Inline/Async Modes, Strategies, Backlog, Read/Write Amplification, Failure Recovery, and SLAs: Concepts, Terminology, Architecture, and Mental ModelPlanned lesson · reserved path Chapter17/Lesson1.html
Planned
02
Compaction: Scheduling vs Execution, Inline/Async Modes, Strategies, Backlog, Read/Write Amplification, Failure Recovery, and SLAs: Guided Hands-On Workflow, Commands, APIs, and ConfigurationPlanned lesson · reserved path Chapter17/Lesson2.html
Planned
03
Compaction: Scheduling vs Execution, Inline/Async Modes, Strategies, Backlog, Read/Write Amplification, Failure Recovery, and SLAs: Design Choices, Scaling Behavior, Compatibility, and TradeoffsPlanned lesson · reserved path Chapter17/Lesson3.html
Planned
04
Compaction: Scheduling vs Execution, Inline/Async Modes, Strategies, Backlog, Read/Write Amplification, Failure Recovery, and SLAs: Failure Modes, Diagnostics, Security, Reliability, and PerformancePlanned lesson · reserved path Chapter17/Lesson4.html
Planned
05
Checkpoint Lab — Compaction: Scheduling vs Execution, Inline/Async Modes, Strategies, Backlog, Read/Write Amplification, Failure Recovery, and SLAs: Build, Test, Measure, Troubleshoot, and Explain the ResultPlanned lesson · reserved path Chapter17/Lesson5.html
Planned
18

Chapter 18

Clustering: File Reorganization, Sorting, Small-File Handling, Async Services, Concurrent Updates, Non-Blocking Improvements, and Layout Evolution

5 lessons
01
Clustering: File Reorganization, Sorting, Small-File Handling, Async Services, Concurrent Updates, Non-Blocking Improvements, and Layout Evolution: Concepts, Terminology, Architecture, and Mental ModelPlanned lesson · reserved path Chapter18/Lesson1.html
Planned
02
Clustering: File Reorganization, Sorting, Small-File Handling, Async Services, Concurrent Updates, Non-Blocking Improvements, and Layout Evolution: Guided Hands-On Workflow, Commands, APIs, and ConfigurationPlanned lesson · reserved path Chapter18/Lesson2.html
Planned
03
Clustering: File Reorganization, Sorting, Small-File Handling, Async Services, Concurrent Updates, Non-Blocking Improvements, and Layout Evolution: Design Choices, Scaling Behavior, Compatibility, and TradeoffsPlanned lesson · reserved path Chapter18/Lesson3.html
Planned
04
Clustering: File Reorganization, Sorting, Small-File Handling, Async Services, Concurrent Updates, Non-Blocking Improvements, and Layout Evolution: Failure Modes, Diagnostics, Security, Reliability, and PerformancePlanned lesson · reserved path Chapter18/Lesson4.html
Planned
05
Checkpoint Lab — Clustering: File Reorganization, Sorting, Small-File Handling, Async Services, Concurrent Updates, Non-Blocking Improvements, and Layout Evolution: Build, Test, Measure, Troubleshoot, and Explain the ResultPlanned lesson · reserved path Chapter18/Lesson5.html
Planned
19

Chapter 19

Cleaning and Archival: Commit Retention, File-Version Cleanup, Timeline Archival, Incremental Consumer Safety, Object Costs, and Recovery Windows

5 lessons
01
Cleaning and Archival: Commit Retention, File-Version Cleanup, Timeline Archival, Incremental Consumer Safety, Object Costs, and Recovery Windows: Concepts, Terminology, Architecture, and Mental ModelPlanned lesson · reserved path Chapter19/Lesson1.html
Planned
02
Cleaning and Archival: Commit Retention, File-Version Cleanup, Timeline Archival, Incremental Consumer Safety, Object Costs, and Recovery Windows: Guided Hands-On Workflow, Commands, APIs, and ConfigurationPlanned lesson · reserved path Chapter19/Lesson2.html
Planned
03
Cleaning and Archival: Commit Retention, File-Version Cleanup, Timeline Archival, Incremental Consumer Safety, Object Costs, and Recovery Windows: Design Choices, Scaling Behavior, Compatibility, and TradeoffsPlanned lesson · reserved path Chapter19/Lesson3.html
Planned
04
Cleaning and Archival: Commit Retention, File-Version Cleanup, Timeline Archival, Incremental Consumer Safety, Object Costs, and Recovery Windows: Failure Modes, Diagnostics, Security, Reliability, and PerformancePlanned lesson · reserved path Chapter19/Lesson4.html
Planned
05
Checkpoint Lab — Cleaning and Archival: Commit Retention, File-Version Cleanup, Timeline Archival, Incremental Consumer Safety, Object Costs, and Recovery Windows: Build, Test, Measure, Troubleshoot, and Explain the ResultPlanned lesson · reserved path Chapter19/Lesson5.html
Planned
20

Chapter 20

Concurrency Control: Single Writer, Optimistic Concurrency, Non-Blocking Concurrency Concepts, Locks, Conflict Detection, and Multi-Writer Design

5 lessons
01
Concurrency Control: Single Writer, Optimistic Concurrency, Non-Blocking Concurrency Concepts, Locks, Conflict Detection, and Multi-Writer Design: Concepts, Terminology, Architecture, and Mental ModelPlanned lesson · reserved path Chapter20/Lesson1.html
Planned
02
Concurrency Control: Single Writer, Optimistic Concurrency, Non-Blocking Concurrency Concepts, Locks, Conflict Detection, and Multi-Writer Design: Guided Hands-On Workflow, Commands, APIs, and ConfigurationPlanned lesson · reserved path Chapter20/Lesson2.html
Planned
03
Concurrency Control: Single Writer, Optimistic Concurrency, Non-Blocking Concurrency Concepts, Locks, Conflict Detection, and Multi-Writer Design: Design Choices, Scaling Behavior, Compatibility, and TradeoffsPlanned lesson · reserved path Chapter20/Lesson3.html
Planned
04
Concurrency Control: Single Writer, Optimistic Concurrency, Non-Blocking Concurrency Concepts, Locks, Conflict Detection, and Multi-Writer Design: Failure Modes, Diagnostics, Security, Reliability, and PerformancePlanned lesson · reserved path Chapter20/Lesson4.html
Planned
05
Checkpoint Lab — Concurrency Control: Single Writer, Optimistic Concurrency, Non-Blocking Concurrency Concepts, Locks, Conflict Detection, and Multi-Writer Design: Build, Test, Measure, Troubleshoot, and Explain the ResultPlanned lesson · reserved path Chapter20/Lesson5.html
Planned
21

Chapter 21

Spark Integration: DataSource/SQL APIs, Catalogs, Procedures, Streaming, Hive Sync, Upserts, Table Services, and Spark 4 Compatibility

5 lessons
01
Spark Integration: DataSource/SQL APIs, Catalogs, Procedures, Streaming, Hive Sync, Upserts, Table Services, and Spark 4 Compatibility: Concepts, Terminology, Architecture, and Mental ModelPlanned lesson · reserved path Chapter21/Lesson1.html
Planned
02
Spark Integration: DataSource/SQL APIs, Catalogs, Procedures, Streaming, Hive Sync, Upserts, Table Services, and Spark 4 Compatibility: Guided Hands-On Workflow, Commands, APIs, and ConfigurationPlanned lesson · reserved path Chapter21/Lesson2.html
Planned
03
Spark Integration: DataSource/SQL APIs, Catalogs, Procedures, Streaming, Hive Sync, Upserts, Table Services, and Spark 4 Compatibility: Design Choices, Scaling Behavior, Compatibility, and TradeoffsPlanned lesson · reserved path Chapter21/Lesson3.html
Planned
04
Spark Integration: DataSource/SQL APIs, Catalogs, Procedures, Streaming, Hive Sync, Upserts, Table Services, and Spark 4 Compatibility: Failure Modes, Diagnostics, Security, Reliability, and PerformancePlanned lesson · reserved path Chapter21/Lesson4.html
Planned
05
Checkpoint Lab — Spark Integration: DataSource/SQL APIs, Catalogs, Procedures, Streaming, Hive Sync, Upserts, Table Services, and Spark 4 Compatibility: Build, Test, Measure, Troubleshoot, and Explain the ResultPlanned lesson · reserved path Chapter21/Lesson5.html
Planned
22

Chapter 22

Flink Integration: Streaming Writes, Checkpoints, Bucket/Record-Level Indexing, Async Compaction, Global Upserts, FLIP-27 Source Concepts, and Recovery

5 lessons
01
Flink Integration: Streaming Writes, Checkpoints, Bucket/Record-Level Indexing, Async Compaction, Global Upserts, FLIP-27 Source Concepts, and Recovery: Concepts, Terminology, Architecture, and Mental ModelPlanned lesson · reserved path Chapter22/Lesson1.html
Planned
02
Flink Integration: Streaming Writes, Checkpoints, Bucket/Record-Level Indexing, Async Compaction, Global Upserts, FLIP-27 Source Concepts, and Recovery: Guided Hands-On Workflow, Commands, APIs, and ConfigurationPlanned lesson · reserved path Chapter22/Lesson2.html
Planned
03
Flink Integration: Streaming Writes, Checkpoints, Bucket/Record-Level Indexing, Async Compaction, Global Upserts, FLIP-27 Source Concepts, and Recovery: Design Choices, Scaling Behavior, Compatibility, and TradeoffsPlanned lesson · reserved path Chapter22/Lesson3.html
Planned
04
Flink Integration: Streaming Writes, Checkpoints, Bucket/Record-Level Indexing, Async Compaction, Global Upserts, FLIP-27 Source Concepts, and Recovery: Failure Modes, Diagnostics, Security, Reliability, and PerformancePlanned lesson · reserved path Chapter22/Lesson4.html
Planned
05
Checkpoint Lab — Flink Integration: Streaming Writes, Checkpoints, Bucket/Record-Level Indexing, Async Compaction, Global Upserts, FLIP-27 Source Concepts, and Recovery: Build, Test, Measure, Troubleshoot, and Explain the ResultPlanned lesson · reserved path Chapter22/Lesson5.html
Planned
23

Chapter 23

Ingestion Frameworks: DeltaStreamer/Streamer Concepts, Kafka/Object Sources, Transformations, Checkpointing, Multi-Table Pipelines, and Operational Boundaries

5 lessons
01
Ingestion Frameworks: DeltaStreamer/Streamer Concepts, Kafka/Object Sources, Transformations, Checkpointing, Multi-Table Pipelines, and Operational Boundaries: Concepts, Terminology, Architecture, and Mental ModelPlanned lesson · reserved path Chapter23/Lesson1.html
Planned
02
Ingestion Frameworks: DeltaStreamer/Streamer Concepts, Kafka/Object Sources, Transformations, Checkpointing, Multi-Table Pipelines, and Operational Boundaries: Guided Hands-On Workflow, Commands, APIs, and ConfigurationPlanned lesson · reserved path Chapter23/Lesson2.html
Planned
03
Ingestion Frameworks: DeltaStreamer/Streamer Concepts, Kafka/Object Sources, Transformations, Checkpointing, Multi-Table Pipelines, and Operational Boundaries: Design Choices, Scaling Behavior, Compatibility, and TradeoffsPlanned lesson · reserved path Chapter23/Lesson3.html
Planned
04
Ingestion Frameworks: DeltaStreamer/Streamer Concepts, Kafka/Object Sources, Transformations, Checkpointing, Multi-Table Pipelines, and Operational Boundaries: Failure Modes, Diagnostics, Security, Reliability, and PerformancePlanned lesson · reserved path Chapter23/Lesson4.html
Planned
05
Checkpoint Lab — Ingestion Frameworks: DeltaStreamer/Streamer Concepts, Kafka/Object Sources, Transformations, Checkpointing, Multi-Table Pipelines, and Operational Boundaries: Build, Test, Measure, Troubleshoot, and Explain the ResultPlanned lesson · reserved path Chapter23/Lesson5.html
Planned
24

Chapter 24

Catalog and Metastore Integration: Hive Metastore, Glue-Style Catalogs, Sync Tools, Partition Metadata, Schema Propagation, and Multi-Engine Reads

5 lessons
01
Catalog and Metastore Integration: Hive Metastore, Glue-Style Catalogs, Sync Tools, Partition Metadata, Schema Propagation, and Multi-Engine Reads: Concepts, Terminology, Architecture, and Mental ModelPlanned lesson · reserved path Chapter24/Lesson1.html
Planned
02
Catalog and Metastore Integration: Hive Metastore, Glue-Style Catalogs, Sync Tools, Partition Metadata, Schema Propagation, and Multi-Engine Reads: Guided Hands-On Workflow, Commands, APIs, and ConfigurationPlanned lesson · reserved path Chapter24/Lesson2.html
Planned
03
Catalog and Metastore Integration: Hive Metastore, Glue-Style Catalogs, Sync Tools, Partition Metadata, Schema Propagation, and Multi-Engine Reads: Design Choices, Scaling Behavior, Compatibility, and TradeoffsPlanned lesson · reserved path Chapter24/Lesson3.html
Planned
04
Catalog and Metastore Integration: Hive Metastore, Glue-Style Catalogs, Sync Tools, Partition Metadata, Schema Propagation, and Multi-Engine Reads: Failure Modes, Diagnostics, Security, Reliability, and PerformancePlanned lesson · reserved path Chapter24/Lesson4.html
Planned
05
Checkpoint Lab — Catalog and Metastore Integration: Hive Metastore, Glue-Style Catalogs, Sync Tools, Partition Metadata, Schema Propagation, and Multi-Engine Reads: Build, Test, Measure, Troubleshoot, and Explain the ResultPlanned lesson · reserved path Chapter24/Lesson5.html
Planned
25

Chapter 25

Object Storage Semantics: S3/GCS/Azure Concepts, Listing, Consistency, Multipart I/O, Markers, Credentials, Encryption, and Storage Cost

5 lessons
01
Object Storage Semantics: S3/GCS/Azure Concepts, Listing, Consistency, Multipart I/O, Markers, Credentials, Encryption, and Storage Cost: Concepts, Terminology, Architecture, and Mental ModelPlanned lesson · reserved path Chapter25/Lesson1.html
Planned
02
Object Storage Semantics: S3/GCS/Azure Concepts, Listing, Consistency, Multipart I/O, Markers, Credentials, Encryption, and Storage Cost: Guided Hands-On Workflow, Commands, APIs, and ConfigurationPlanned lesson · reserved path Chapter25/Lesson2.html
Planned
03
Object Storage Semantics: S3/GCS/Azure Concepts, Listing, Consistency, Multipart I/O, Markers, Credentials, Encryption, and Storage Cost: Design Choices, Scaling Behavior, Compatibility, and TradeoffsPlanned lesson · reserved path Chapter25/Lesson3.html
Planned
04
Object Storage Semantics: S3/GCS/Azure Concepts, Listing, Consistency, Multipart I/O, Markers, Credentials, Encryption, and Storage Cost: Failure Modes, Diagnostics, Security, Reliability, and PerformancePlanned lesson · reserved path Chapter25/Lesson4.html
Planned
05
Checkpoint Lab — Object Storage Semantics: S3/GCS/Azure Concepts, Listing, Consistency, Multipart I/O, Markers, Credentials, Encryption, and Storage Cost: Build, Test, Measure, Troubleshoot, and Explain the ResultPlanned lesson · reserved path Chapter25/Lesson5.html
Planned
26

Chapter 26

AI and Multimodal Data in Hudi 1.2: VECTOR, VARIANT, BLOB Types, Embeddings, Semi-Structured/Binary Data, and Compatibility Boundaries

5 lessons
01
AI and Multimodal Data in Hudi 1.2: VECTOR, VARIANT, BLOB Types, Embeddings, Semi-Structured/Binary Data, and Compatibility Boundaries: Concepts, Terminology, Architecture, and Mental ModelPlanned lesson · reserved path Chapter26/Lesson1.html
Planned
02
AI and Multimodal Data in Hudi 1.2: VECTOR, VARIANT, BLOB Types, Embeddings, Semi-Structured/Binary Data, and Compatibility Boundaries: Guided Hands-On Workflow, Commands, APIs, and ConfigurationPlanned lesson · reserved path Chapter26/Lesson2.html
Planned
03
AI and Multimodal Data in Hudi 1.2: VECTOR, VARIANT, BLOB Types, Embeddings, Semi-Structured/Binary Data, and Compatibility Boundaries: Design Choices, Scaling Behavior, Compatibility, and TradeoffsPlanned lesson · reserved path Chapter26/Lesson3.html
Planned
04
AI and Multimodal Data in Hudi 1.2: VECTOR, VARIANT, BLOB Types, Embeddings, Semi-Structured/Binary Data, and Compatibility Boundaries: Failure Modes, Diagnostics, Security, Reliability, and PerformancePlanned lesson · reserved path Chapter26/Lesson4.html
Planned
05
Checkpoint Lab — AI and Multimodal Data in Hudi 1.2: VECTOR, VARIANT, BLOB Types, Embeddings, Semi-Structured/Binary Data, and Compatibility Boundaries: Build, Test, Measure, Troubleshoot, and Explain the ResultPlanned lesson · reserved path Chapter26/Lesson5.html
Planned
27

Chapter 27

Lance File Format and Vector Search: Lance Base Files, hudi_vector_search, AI/ML Access Patterns, Hybrid Table Design, and Experimental/Workload Tradeoffs

5 lessons
01
Lance File Format and Vector Search: Lance Base Files, hudi_vector_search, AI/ML Access Patterns, Hybrid Table Design, and Experimental/Workload Tradeoffs: Concepts, Terminology, Architecture, and Mental ModelPlanned lesson · reserved path Chapter27/Lesson1.html
Planned
02
Lance File Format and Vector Search: Lance Base Files, hudi_vector_search, AI/ML Access Patterns, Hybrid Table Design, and Experimental/Workload Tradeoffs: Guided Hands-On Workflow, Commands, APIs, and ConfigurationPlanned lesson · reserved path Chapter27/Lesson2.html
Planned
03
Lance File Format and Vector Search: Lance Base Files, hudi_vector_search, AI/ML Access Patterns, Hybrid Table Design, and Experimental/Workload Tradeoffs: Design Choices, Scaling Behavior, Compatibility, and TradeoffsPlanned lesson · reserved path Chapter27/Lesson3.html
Planned
04
Lance File Format and Vector Search: Lance Base Files, hudi_vector_search, AI/ML Access Patterns, Hybrid Table Design, and Experimental/Workload Tradeoffs: Failure Modes, Diagnostics, Security, Reliability, and PerformancePlanned lesson · reserved path Chapter27/Lesson4.html
Planned
05
Checkpoint Lab — Lance File Format and Vector Search: Lance Base Files, hudi_vector_search, AI/ML Access Patterns, Hybrid Table Design, and Experimental/Workload Tradeoffs: Build, Test, Measure, Troubleshoot, and Explain the ResultPlanned lesson · reserved path Chapter27/Lesson5.html
Planned
28

Chapter 28

Performance Engineering: File Sizing, Index Selection, Metadata Table, Compaction/Clustering Cadence, Write Parallelism, Read Amplification, and Benchmarking

5 lessons
01
Performance Engineering: File Sizing, Index Selection, Metadata Table, Compaction/Clustering Cadence, Write Parallelism, Read Amplification, and Benchmarking: Concepts, Terminology, Architecture, and Mental ModelPlanned lesson · reserved path Chapter28/Lesson1.html
Planned
02
Performance Engineering: File Sizing, Index Selection, Metadata Table, Compaction/Clustering Cadence, Write Parallelism, Read Amplification, and Benchmarking: Guided Hands-On Workflow, Commands, APIs, and ConfigurationPlanned lesson · reserved path Chapter28/Lesson2.html
Planned
03
Performance Engineering: File Sizing, Index Selection, Metadata Table, Compaction/Clustering Cadence, Write Parallelism, Read Amplification, and Benchmarking: Design Choices, Scaling Behavior, Compatibility, and TradeoffsPlanned lesson · reserved path Chapter28/Lesson3.html
Planned
04
Performance Engineering: File Sizing, Index Selection, Metadata Table, Compaction/Clustering Cadence, Write Parallelism, Read Amplification, and Benchmarking: Failure Modes, Diagnostics, Security, Reliability, and PerformancePlanned lesson · reserved path Chapter28/Lesson4.html
Planned
05
Checkpoint Lab — Performance Engineering: File Sizing, Index Selection, Metadata Table, Compaction/Clustering Cadence, Write Parallelism, Read Amplification, and Benchmarking: Build, Test, Measure, Troubleshoot, and Explain the ResultPlanned lesson · reserved path Chapter28/Lesson5.html
Planned
29

Chapter 29

Security and Governance: Storage IAM, Catalog Authorization, Encryption, Secrets, Audit, PII Handling, Data Retention, and Least Privilege

5 lessons
01
Security and Governance: Storage IAM, Catalog Authorization, Encryption, Secrets, Audit, PII Handling, Data Retention, and Least Privilege: Concepts, Terminology, Architecture, and Mental ModelPlanned lesson · reserved path Chapter29/Lesson1.html
Planned
02
Security and Governance: Storage IAM, Catalog Authorization, Encryption, Secrets, Audit, PII Handling, Data Retention, and Least Privilege: Guided Hands-On Workflow, Commands, APIs, and ConfigurationPlanned lesson · reserved path Chapter29/Lesson2.html
Planned
03
Security and Governance: Storage IAM, Catalog Authorization, Encryption, Secrets, Audit, PII Handling, Data Retention, and Least Privilege: Design Choices, Scaling Behavior, Compatibility, and TradeoffsPlanned lesson · reserved path Chapter29/Lesson3.html
Planned
04
Security and Governance: Storage IAM, Catalog Authorization, Encryption, Secrets, Audit, PII Handling, Data Retention, and Least Privilege: Failure Modes, Diagnostics, Security, Reliability, and PerformancePlanned lesson · reserved path Chapter29/Lesson4.html
Planned
05
Checkpoint Lab — Security and Governance: Storage IAM, Catalog Authorization, Encryption, Secrets, Audit, PII Handling, Data Retention, and Least Privilege: Build, Test, Measure, Troubleshoot, and Explain the ResultPlanned lesson · reserved path Chapter29/Lesson5.html
Planned
30

Chapter 30

Observability and Troubleshooting: Timeline Inspection, CLI/Procedures, Metrics, Failed Writes, Rollbacks, Index Drift, Compaction Backlog, and Corruption Triage

5 lessons
01
Observability and Troubleshooting: Timeline Inspection, CLI/Procedures, Metrics, Failed Writes, Rollbacks, Index Drift, Compaction Backlog, and Corruption Triage: Concepts, Terminology, Architecture, and Mental ModelPlanned lesson · reserved path Chapter30/Lesson1.html
Planned
02
Observability and Troubleshooting: Timeline Inspection, CLI/Procedures, Metrics, Failed Writes, Rollbacks, Index Drift, Compaction Backlog, and Corruption Triage: Guided Hands-On Workflow, Commands, APIs, and ConfigurationPlanned lesson · reserved path Chapter30/Lesson2.html
Planned
03
Observability and Troubleshooting: Timeline Inspection, CLI/Procedures, Metrics, Failed Writes, Rollbacks, Index Drift, Compaction Backlog, and Corruption Triage: Design Choices, Scaling Behavior, Compatibility, and TradeoffsPlanned lesson · reserved path Chapter30/Lesson3.html
Planned
04
Observability and Troubleshooting: Timeline Inspection, CLI/Procedures, Metrics, Failed Writes, Rollbacks, Index Drift, Compaction Backlog, and Corruption Triage: Failure Modes, Diagnostics, Security, Reliability, and PerformancePlanned lesson · reserved path Chapter30/Lesson4.html
Planned
05
Checkpoint Lab — Observability and Troubleshooting: Timeline Inspection, CLI/Procedures, Metrics, Failed Writes, Rollbacks, Index Drift, Compaction Backlog, and Corruption Triage: Build, Test, Measure, Troubleshoot, and Explain the ResultPlanned lesson · reserved path Chapter30/Lesson5.html
Planned
31

Chapter 31

Migration, Upgrade, and Interoperability: Hudi Version/Table-Version Upgrades, Existing Lake Conversion, XTable Awareness, Engine Compatibility, and Rollback Gates

5 lessons
01
Migration, Upgrade, and Interoperability: Hudi Version/Table-Version Upgrades, Existing Lake Conversion, XTable Awareness, Engine Compatibility, and Rollback Gates: Concepts, Terminology, Architecture, and Mental ModelPlanned lesson · reserved path Chapter31/Lesson1.html
Planned
02
Migration, Upgrade, and Interoperability: Hudi Version/Table-Version Upgrades, Existing Lake Conversion, XTable Awareness, Engine Compatibility, and Rollback Gates: Guided Hands-On Workflow, Commands, APIs, and ConfigurationPlanned lesson · reserved path Chapter31/Lesson2.html
Planned
03
Migration, Upgrade, and Interoperability: Hudi Version/Table-Version Upgrades, Existing Lake Conversion, XTable Awareness, Engine Compatibility, and Rollback Gates: Design Choices, Scaling Behavior, Compatibility, and TradeoffsPlanned lesson · reserved path Chapter31/Lesson3.html
Planned
04
Migration, Upgrade, and Interoperability: Hudi Version/Table-Version Upgrades, Existing Lake Conversion, XTable Awareness, Engine Compatibility, and Rollback Gates: Failure Modes, Diagnostics, Security, Reliability, and PerformancePlanned lesson · reserved path Chapter31/Lesson4.html
Planned
05
Checkpoint Lab — Migration, Upgrade, and Interoperability: Hudi Version/Table-Version Upgrades, Existing Lake Conversion, XTable Awareness, Engine Compatibility, and Rollback Gates: Build, Test, Measure, Troubleshoot, and Explain the ResultPlanned lesson · reserved path Chapter31/Lesson5.html
Planned
32

Chapter 32

Backup and Disaster Recovery: Timeline/Object Backups, Cross-Region Copies, Catalog Recovery, Restore Validation, RPO/RTO, and Failure Drills

5 lessons
01
Backup and Disaster Recovery: Timeline/Object Backups, Cross-Region Copies, Catalog Recovery, Restore Validation, RPO/RTO, and Failure Drills: Concepts, Terminology, Architecture, and Mental ModelPlanned lesson · reserved path Chapter32/Lesson1.html
Planned
02
Backup and Disaster Recovery: Timeline/Object Backups, Cross-Region Copies, Catalog Recovery, Restore Validation, RPO/RTO, and Failure Drills: Guided Hands-On Workflow, Commands, APIs, and ConfigurationPlanned lesson · reserved path Chapter32/Lesson2.html
Planned
03
Backup and Disaster Recovery: Timeline/Object Backups, Cross-Region Copies, Catalog Recovery, Restore Validation, RPO/RTO, and Failure Drills: Design Choices, Scaling Behavior, Compatibility, and TradeoffsPlanned lesson · reserved path Chapter32/Lesson3.html
Planned
04
Backup and Disaster Recovery: Timeline/Object Backups, Cross-Region Copies, Catalog Recovery, Restore Validation, RPO/RTO, and Failure Drills: Failure Modes, Diagnostics, Security, Reliability, and PerformancePlanned lesson · reserved path Chapter32/Lesson4.html
Planned
05
Checkpoint Lab — Backup and Disaster Recovery: Timeline/Object Backups, Cross-Region Copies, Catalog Recovery, Restore Validation, RPO/RTO, and Failure Drills: Build, Test, Measure, Troubleshoot, and Explain the ResultPlanned lesson · reserved path Chapter32/Lesson5.html
Planned
33

Chapter 33

Production Capstone: Ingest, Upsert, Index, Compact, Cluster, Stream, Query Incrementally, Secure, Observe, Fail, Recover, and Tune Hudi 1.2

5 lessons
01
Production Capstone: Ingest, Upsert, Index, Compact, Cluster, Stream, Query Incrementally, Secure, Observe, Fail, Recover, and Tune Hudi 1.2: Concepts, Terminology, Architecture, and Mental ModelPlanned lesson · reserved path Chapter33/Lesson1.html
Planned
02
Production Capstone: Ingest, Upsert, Index, Compact, Cluster, Stream, Query Incrementally, Secure, Observe, Fail, Recover, and Tune Hudi 1.2: Guided Hands-On Workflow, Commands, APIs, and ConfigurationPlanned lesson · reserved path Chapter33/Lesson2.html
Planned
03
Production Capstone: Ingest, Upsert, Index, Compact, Cluster, Stream, Query Incrementally, Secure, Observe, Fail, Recover, and Tune Hudi 1.2: Design Choices, Scaling Behavior, Compatibility, and TradeoffsPlanned lesson · reserved path Chapter33/Lesson3.html
Planned
04
Production Capstone: Ingest, Upsert, Index, Compact, Cluster, Stream, Query Incrementally, Secure, Observe, Fail, Recover, and Tune Hudi 1.2: Failure Modes, Diagnostics, Security, Reliability, and PerformancePlanned lesson · reserved path Chapter33/Lesson4.html
Planned
05
Checkpoint Lab — Production Capstone: Ingest, Upsert, Index, Compact, Cluster, Stream, Query Incrementally, Secure, Observe, Fail, Recover, and Tune Hudi 1.2: Build, Test, Measure, Troubleshoot, and Explain the ResultPlanned lesson · reserved path Chapter33/Lesson5.html
Planned