60 Lakehouse & Architecture questions from the Big Data bank, written for Indian campus drives and tech interviews. Every question has a verified answer and an AI-tutor explanation on placd.
Free to start: the 2-minute IT readiness check — six questions and a result.
A.an architecture that adds data-warehouse features like ACID transactions and governance directly on low-cost data-lake storage
B.a layered design organizing data into bronze raw, silver cleaned, and gold curated tables for progressive refinement
C.a data-warehouse pattern for tracking changes to dimension attributes over time, such as keeping history with type-2 rows
D.a dimensional modeling pattern for tracking attribute history, with Type 1 overwriting and Type 2 adding new versioned rows
Answer + AI explanation with Pro
2. Which term means: "an architecture that adds data-warehouse features like ACID transactions and governance directly on low-cost data-lake storage"?
Junior
A.Schema-on-read
B.Slowly changing dimension
C.data contract
D.Lakehouse
Answer + AI explanation with Pro
3. Which statement is correct?
Junior
A.Lakehouse — the practice of moving curated warehouse data back into operational systems like CRMs to activate analytics
B.Lakehouse — a central repository that stores raw structured and unstructured data at scale on low-cost object storage
C.Lakehouse — a data-quality pattern that writes to a hidden branch, runs validations, then publishes only if checks pass
D.Lakehouse — an architecture that adds data-warehouse features like ACID transactions and governance directly on low-cost data-lake storage
Answer + AI explanation with Pro
4. What is Medallion architecture?
Junior
A.a design that processes everything as a stream, replaying the log for reprocessing instead of maintaining a separate batch layer
B.a lakehouse table continuously appended from a streaming source and incrementally processed, often the bronze layer of a pipeline
C.a data-warehouse pattern for tracking changes to dimension attributes over time, such as keeping history with type-2 rows
D.a layered design organizing data into bronze raw, silver cleaned, and gold curated tables for progressive refinement
Answer + AI explanation with Pro
5. Which term means: "a layered design organizing data into bronze raw, silver cleaned, and gold curated tables for progressive refinement"?
Junior
A.data lineage
B.Medallion architecture
C.Reverse ETL
D.Data contract
Answer + AI explanation with Pro
6. Which statement is correct?
Junior
A.Medallion architecture — a lakehouse table continuously appended from a streaming source and incrementally processed, often the bronze layer of a pipeline
B.Medallion architecture — a self-contained, governed, and documented dataset owned by a domain team and treated as a first-class deliverable in data mesh
C.Medallion architecture — a design that processes everything as a stream, replaying the log for reprocessing instead of maintaining a separate batch layer
D.Medallion architecture — a layered design organizing data into bronze raw, silver cleaned, and gold curated tables for progressive refinement
Answer + AI explanation with Pro
7. What is Data mesh?
Junior
A.an architecture that adds data-warehouse features like ACID transactions and governance directly on low-cost data-lake storage
B.a central repository that stores raw structured and unstructured data at scale on low-cost object storage
C.the layered lakehouse design refining data through bronze (raw), silver (cleaned), and gold (aggregated) tables
D.a decentralized approach where domain teams own their data as products served through a self-serve platform with federated governance
Answer + AI explanation with Pro
8. Which term means: "a decentralized approach where domain teams own their data as products served through a self-serve platform with federated governance"?
Junior
A.Schema-on-read
B.data contract
C.change data capture
D.Data mesh
Answer + AI explanation with Pro
9. Which statement is correct?
Junior
A.Data mesh — a layered design organizing data into bronze raw, silver cleaned, and gold curated tables for progressive refinement
B.Data mesh — the pattern of capturing inserts, updates, and deletes from a source database's log and streaming them into a lakehouse for near-real-time replication
C.Data mesh — a precomputed, stored query result that is incrementally refreshed so downstream reads avoid recomputing expensive aggregations
D.Data mesh — a decentralized approach where domain teams own their data as products served through a self-serve platform with federated governance
Answer + AI explanation with Pro
10. What is Lambda architecture?
Mid
A.the layered lakehouse design refining data through bronze (raw), silver (cleaned), and gold (aggregated) tables
B.a data-quality pattern that writes to a hidden branch, runs validations, then publishes only if checks pass
C.a design that runs parallel batch and speed layers and merges their outputs at query time to balance accuracy with low latency
D.a central repository that stores raw structured and unstructured data at scale on low-cost object storage
Answer + AI explanation with Pro
11. Which term means: "a design that runs parallel batch and speed layers and merges their outputs at query time to balance accuracy with low latency"?
Mid
A.Write-audit-publish
B.CDC
C.Lambda architecture
D.data product
Answer + AI explanation with Pro
12. Which statement is correct?
Mid
A.Lambda architecture — a design that runs parallel batch and speed layers and merges their outputs at query time to balance accuracy with low latency
B.Lambda architecture — the layered lakehouse design refining data through bronze (raw), silver (cleaned), and gold (aggregated) tables
C.Lambda architecture — a layered design organizing data into bronze raw, silver cleaned, and gold curated tables for progressive refinement
D.Lambda architecture — an architecture that adds data-warehouse features like ACID transactions and governance directly on low-cost data-lake storage
Answer + AI explanation with Pro
13. What is Kappa architecture?
Mid
A.the pattern of capturing inserts, updates, and deletes from a source database's log and streaming them into a lakehouse for near-real-time replication
B.the tracked path of data from sources through transformations to outputs, used for impact analysis, debugging, and governance
C.a precomputed, stored query result that is incrementally refreshed so downstream reads avoid recomputing expensive aggregations
D.a design that processes everything as a stream, replaying the log for reprocessing instead of maintaining a separate batch layer
Answer + AI explanation with Pro
14. Which term means: "a design that processes everything as a stream, replaying the log for reprocessing instead of maintaining a separate batch layer"?
Mid
A.Write-audit-publish
B.Kappa architecture
C.Data mesh
D.Data lake
Answer + AI explanation with Pro
15. Which statement is correct?
Mid
A.Kappa architecture — a design that processes everything as a stream, replaying the log for reprocessing instead of maintaining a separate batch layer
B.Kappa architecture — an architecture that adds data-warehouse features like ACID transactions and governance directly on low-cost data-lake storage
C.Kappa architecture — a design that runs parallel batch and speed layers and merges their outputs at query time to balance accuracy with low latency
D.Kappa architecture — a central repository that stores raw structured and unstructured data at scale on low-cost object storage
Answer + AI explanation with Pro
16. What is Schema-on-read?
Mid
A.change data capture, the technique of streaming row-level inserts, updates, and deletes from a source database into a pipeline
B.a dimensional modeling pattern for tracking attribute history, with Type 1 overwriting and Type 2 adding new versioned rows
C.the approach of storing raw data as-is and applying structure only when it is queried, common in data lakes
D.an architecture that adds data-warehouse features like ACID transactions and governance directly on low-cost data-lake storage
Answer + AI explanation with Pro
17. Which term means: "the approach of storing raw data as-is and applying structure only when it is queried, common in data lakes"?
Mid
A.Data mesh
B.Schema-on-read
C.CDC
D.Lakehouse
Answer + AI explanation with Pro
18. Which statement is correct?
Mid
A.Schema-on-read — a layered design organizing data into bronze raw, silver cleaned, and gold curated tables for progressive refinement
B.Schema-on-read — the approach of storing raw data as-is and applying structure only when it is queried, common in data lakes
C.Schema-on-read — a self-contained, governed, and documented dataset owned by a domain team and treated as a first-class deliverable in data mesh
D.Schema-on-read — the layered lakehouse design refining data through bronze (raw), silver (cleaned), and gold (aggregated) tables
Answer + AI explanation with Pro
19. What is Slowly changing dimension?
Mid
A.a central repository that stores raw structured and unstructured data at scale on low-cost object storage
B.a data-warehouse pattern for tracking changes to dimension attributes over time, such as keeping history with type-2 rows
C.a lakehouse table continuously appended from a streaming source and incrementally processed, often the bronze layer of a pipeline
D.a data-quality pattern that writes to a hidden branch, runs validations, then publishes only if checks pass
Answer + AI explanation with Pro
20. Which term means: "a data-warehouse pattern for tracking changes to dimension attributes over time, such as keeping history with type-2 rows"?
Mid
A.change data capture
B.Schema-on-read
C.Slowly changing dimension
D.data product
Answer + AI explanation with Pro
21. Which statement is correct?
Mid
A.Slowly changing dimension — a decentralized approach where domain teams own their data as products served through a self-serve platform with federated governance
B.Slowly changing dimension — a design that runs parallel batch and speed layers and merges their outputs at query time to balance accuracy with low latency
C.Slowly changing dimension — a data-warehouse pattern for tracking changes to dimension attributes over time, such as keeping history with type-2 rows
D.Slowly changing dimension — a precomputed, stored query result that is incrementally refreshed so downstream reads avoid recomputing expensive aggregations
Answer + AI explanation with Pro
22. What is Data contract?
Senior
A.the tracked path of data from sources through transformations to outputs, used for impact analysis, debugging, and governance
B.a formal agreement on schema, semantics, and quality between data producers and consumers to keep pipelines stable
C.a lakehouse table continuously appended from a streaming source and incrementally processed, often the bronze layer of a pipeline
D.a layered design organizing data into bronze raw, silver cleaned, and gold curated tables for progressive refinement
Answer + AI explanation with Pro
23. Which term means: "a formal agreement on schema, semantics, and quality between data producers and consumers to keep pipelines stable"?
Senior
A.Data contract
B.medallion architecture
C.Data mesh
D.streaming table
Answer + AI explanation with Pro
24. Which statement is correct?
Senior
A.Data contract — a formal agreement on schema, semantics, and quality between data producers and consumers to keep pipelines stable
B.Data contract — an agreed, versioned specification of a dataset's schema, semantics, and quality guarantees between producers and consumers
C.Data contract — a dimensional modeling pattern for tracking attribute history, with Type 1 overwriting and Type 2 adding new versioned rows
D.Data contract — the approach of storing raw data as-is and applying structure only when it is queried, common in data lakes
Answer + AI explanation with Pro
25. What is Reverse ETL?
Senior
A.a self-contained, governed, and documented dataset owned by a domain team and treated as a first-class deliverable in data mesh
B.the practice of moving curated warehouse data back into operational systems like CRMs to activate analytics
C.the approach of storing raw data as-is and applying structure only when it is queried, common in data lakes
D.a lakehouse table continuously appended from a streaming source and incrementally processed, often the bronze layer of a pipeline
Answer + AI explanation with Pro
26. Which term means: "the practice of moving curated warehouse data back into operational systems like CRMs to activate analytics"?
Senior
A.CDC
B.materialized view
C.Reverse ETL
D.data lineage
Answer + AI explanation with Pro
27. Which statement is correct?
Senior
A.Reverse ETL — a design that processes everything as a stream, replaying the log for reprocessing instead of maintaining a separate batch layer
B.Reverse ETL — change data capture, the technique of streaming row-level inserts, updates, and deletes from a source database into a pipeline
C.Reverse ETL — the practice of moving curated warehouse data back into operational systems like CRMs to activate analytics
D.Reverse ETL — the tracked path of data from sources through transformations to outputs, used for impact analysis, debugging, and governance
Answer + AI explanation with Pro
28. What is Write-audit-publish?
Senior
A.the practice of moving curated warehouse data back into operational systems like CRMs to activate analytics
B.a formal agreement on schema, semantics, and quality between data producers and consumers to keep pipelines stable
C.a data-quality pattern that writes to a hidden branch, runs validations, then publishes only if checks pass
D.the layered lakehouse design refining data through bronze (raw), silver (cleaned), and gold (aggregated) tables
Answer + AI explanation with Pro
29. Which term means: "a data-quality pattern that writes to a hidden branch, runs validations, then publishes only if checks pass"?
Senior
A.Write-audit-publish
B.Schema-on-read
C.Kappa architecture
D.data contract
Answer + AI explanation with Pro
30. Which statement is correct?
Senior
A.Write-audit-publish — a data-quality pattern that writes to a hidden branch, runs validations, then publishes only if checks pass
B.Write-audit-publish — a data-warehouse pattern for tracking changes to dimension attributes over time, such as keeping history with type-2 rows
C.Write-audit-publish — a central repository that stores raw structured and unstructured data at scale on low-cost object storage
D.Write-audit-publish — a design that processes everything as a stream, replaying the log for reprocessing instead of maintaining a separate batch layer
Answer + AI explanation with Pro
Showing 30 of 60 Lakehouse & Architecture questions — the full set, with answers, explanations and an AI tutor on every question, is inside.
Free to start
Start with a free readiness check
Sign up free for the 2-minute IT readiness check and a scored result. Answers, explanations and the AI tutor on every Lakehouse & Architecture question come with Pro.