Paths Subjects Questions Quizzes Pricing Search
Data Engineering Intermediate Pro

Data Warehouses & Lakehouses

OLTP vs OLAP, columnar storage, MPP architecture, and the lakehouse's bet on open table formats — the systems layer every data engineering interview eventually asks you to reason about

25 min read 8 views

A practitioner's tour of the storage and compute architectures that sit underneath every analytics stack: why OLTP systems and analytical workloads want fundamentally different engines, what columnar storage buys you and why it's the single biggest lever behind fast aggregate queries, the MPP conceptual model that Snowflake, BigQuery, and Redshift all implement in different ways, and the lakehouse's core bet — that open table formats (Iceberg, Delta Lake, Hudi) can bring warehouse-grade ACID transactions, time travel, and schema evolution to plain object storage. Closes with a decision framework for warehouse vs lakehouse vs plain data lake, and a comparison of how Snowflake, BigQuery, Redshift, and Databricks actually position against each other.

Practice questions (5)

  • Diagnosing a Slow Analytics Replica

    Intermediate · Free
    View →
  • A Warehouse Bill Doubled Overnight — Diagnose the MPP Cost Spike

    Intermediate
    View →
  • Debugging a Concurrent-Write Conflict on an Iceberg Table

    Intermediate
    View →
  • Choosing an Architecture for a Growing Data Platform

    Intermediate
    View →
  • Evaluating a Proposed Redshift-to-Snowflake Migration

    Intermediate
    View →

We use cookies for product analytics to improve OmniAtlas. See our Privacy Policy.