96 Iceberg questions from the Big Data bank, written for Indian campus drives and tech interviews. Every question has a verified answer and an AI-tutor explanation on placd.
Free to start: the 2-minute IT readiness check — six questions and a result.
A.an Iceberg spec v3 feature assigning each row a stable _row_id and sequence number to support change tracking
B.an Iceberg feature where partition values are derived from columns via transforms, so queries need not filter on a separate partition column
C.an Iceberg merge-on-read delete file marking removed rows by their file path and row position, efficient to apply but tied to specific data files
D.a named reference to a specific Iceberg snapshot, often used to mark a release or retain a snapshot for a retention period
Answer + AI explanation with Pro
2. Which term means: "an Iceberg feature where partition values are derived from columns via transforms, so queries need not filter on a separate partition column"?
Junior
A.rewrite manifests
B.sort order
C.Partition evolution
D.Hidden partitioning
Answer + AI explanation with Pro
3. Which statement is correct?
Junior
A.Hidden partitioning — an Iceberg virtual table (such as files, snapshots, history, or partitions) queried to inspect a table's structure and evolution
B.Hidden partitioning — an Iceberg feature where partition values are derived from columns via transforms, so queries need not filter on a separate partition column
C.Hidden partitioning — the Iceberg catalog that adds Git-like branches, tags, and commits over tables so multi-table changes can be versioned and merged
D.Hidden partitioning — an Iceberg merge-on-read delete file marking removed rows by their file path and row position, efficient to apply but tied to specific data files
Answer + AI explanation with Pro
4. What is Partition evolution?
Junior
A.the Iceberg ability to change a table's partition scheme over time without rewriting existing data files
B.an Iceberg merge-on-read delete file marking removed rows by their file path and row position, efficient to apply but tied to specific data files
C.the Iceberg SQL statement that performs upserts by matching a source against a target and inserting, updating, or deleting in one atomic operation
D.the open-source REST Iceberg catalog originally open sourced by Snowflake, adding role-based access control and multi-engine governance
Answer + AI explanation with Pro
5. Which term means: "the Iceberg ability to change a table's partition scheme over time without rewriting existing data files"?
Junior
A.Partition evolution
B.Merge-on-read
C.compaction
D.REST catalog
Answer + AI explanation with Pro
6. Which statement is correct?
Junior
A.Partition evolution — the Iceberg row-level mode that records changes as delete and data files merged at query time, favoring fast writes over fast reads
B.Partition evolution — the Iceberg ability to change a table's partition scheme over time without rewriting existing data files
C.Partition evolution — an Iceberg update mode that rewrites whole data files on each change, giving fast reads but slower writes
D.Partition evolution — an Iceberg spec v3 binary structure that marks deleted rows compactly per data file, replacing positional delete files
Answer + AI explanation with Pro
7. What is Snapshot?
Junior
A.an Iceberg update mode that rewrites whole data files on each change, giving fast reads but slower writes
B.an Iceberg spec v3 binary structure that marks deleted rows compactly per data file, replacing positional delete files
C.an immutable, point-in-time version of an Iceberg table listing all data files, enabling time travel and rollback
D.an Iceberg table property declaring how rows should be ordered within data files so writers cluster data and readers prune more effectively
Answer + AI explanation with Pro
8. Which term means: "an immutable, point-in-time version of an Iceberg table listing all data files, enabling time travel and rollback"?
Junior
A.write distribution mode
B.sort order
C.Snapshot
D.REST catalog
Answer + AI explanation with Pro
9. Which statement is correct?
Junior
A.Snapshot — the Iceberg SQL statement that performs upserts by matching a source against a target and inserting, updating, or deleting in one atomic operation
B.Snapshot — an immutable, point-in-time version of an Iceberg table listing all data files, enabling time travel and rollback
C.Snapshot — the Iceberg property (none, hash, or range) controlling how Spark redistributes rows across tasks before writing to limit small files and skew
D.Snapshot — an Iceberg update mode that records deletes and updates separately and merges them at query time for faster writes
Answer + AI explanation with Pro
10. What is Manifest file?
Mid
A.the Iceberg row-level mode that records changes as delete and data files merged at query time, favoring fast writes over fast reads
B.an Iceberg metadata file that lists data files with their partition values and column statistics for pruning
C.an Iceberg update mode that records deletes and updates separately and merges them at query time for faster writes
D.an Iceberg merge-on-read delete file marking removed rows by their file path and row position, efficient to apply but tied to specific data files
Answer + AI explanation with Pro
11. Which term means: "an Iceberg metadata file that lists data files with their partition values and column statistics for pruning"?
Mid
A.merge-on-read
B.Apache Polaris
C.copy-on-write
D.Manifest file
Answer + AI explanation with Pro
12. Which statement is correct?
Mid
A.Manifest file — an Iceberg table property declaring how rows should be ordered within data files so writers cluster data and readers prune more effectively
B.Manifest file — an Iceberg merge-on-read delete file marking removed rows by their file path and row position, efficient to apply but tied to specific data files
C.Manifest file — an Iceberg metadata file that lists data files with their partition values and column statistics for pruning
D.Manifest file — an Iceberg update mode that records deletes and updates separately and merges them at query time for faster writes
Answer + AI explanation with Pro
13. What is Branch?
Mid
A.an open-source Iceberg REST catalog (1.0 in 2025) providing centralized governance and credential vending across engines
B.a named, independent line of Iceberg snapshots that supports isolated writes such as a write-audit-publish staging area
C.an Iceberg metadata file that lists data files with their partition values and column statistics for pruning
D.the Iceberg property (none, hash, or range) controlling how Spark redistributes rows across tasks before writing to limit small files and skew
Answer + AI explanation with Pro
14. Which term means: "a named, independent line of Iceberg snapshots that supports isolated writes such as a write-audit-publish staging area"?
Mid
A.Polaris catalog
B.Branch
C.write distribution mode
D.merge-on-read
Answer + AI explanation with Pro
15. Which statement is correct?
Mid
A.Branch — an Iceberg merge-on-read delete file marking removed rows by column values such as id equals five, flexible to write but costly to apply on read
B.Branch — an Iceberg feature where partition values are derived from columns via transforms, so queries need not filter on a separate partition column
C.Branch — a named, independent line of Iceberg snapshots that supports isolated writes such as a write-audit-publish staging area
D.Branch — an Iceberg spec v3 binary structure that marks deleted rows compactly per data file, replacing positional delete files
Answer + AI explanation with Pro
16. What is Tag?
Mid
A.an Iceberg catalog interface defined by a standard HTTP API so engines can share one catalog implementation across vendors
B.the Iceberg rewrite_data_files action that merges many small files into larger ones and can re-sort data to restore read performance
C.an Iceberg spec v3 binary structure that marks deleted rows compactly per data file, replacing positional delete files
D.a named reference to a specific Iceberg snapshot, often used to mark a release or retain a snapshot for a retention period
Answer + AI explanation with Pro
17. Which term means: "a named reference to a specific Iceberg snapshot, often used to mark a release or retain a snapshot for a retention period"?
Mid
A.rewrite manifests
B.Snapshot
C.Tag
D.metadata table
Answer + AI explanation with Pro
18. Which statement is correct?
Mid
A.Tag — the Iceberg SQL statement that performs upserts by matching a source against a target and inserting, updating, or deleting in one atomic operation
B.Tag — an Iceberg feature where partition values are derived from columns via transforms, so queries need not filter on a separate partition column
C.Tag — a named reference to a specific Iceberg snapshot, often used to mark a release or retain a snapshot for a retention period
D.Tag — the Iceberg catalog that adds Git-like branches, tags, and commits over tables so multi-table changes can be versioned and merged
Answer + AI explanation with Pro
19. What is Copy-on-write?
Mid
A.an Iceberg update mode that rewrites whole data files on each change, giving fast reads but slower writes
B.the Iceberg ability to query a table as of an earlier snapshot or timestamp for reproducibility and auditing
C.the Iceberg catalog implementation that tracks table pointers in a Hive Metastore, a common bridge for existing Hadoop deployments
D.the Iceberg ability to change a table's partition scheme over time without rewriting existing data files
Answer + AI explanation with Pro
20. Which term means: "an Iceberg update mode that rewrites whole data files on each change, giving fast reads but slower writes"?
Mid
A.Copy-on-write
B.write distribution mode
C.Apache Polaris
D.Nessie catalog
Answer + AI explanation with Pro
21. Which statement is correct?
Mid
A.Copy-on-write — an Iceberg update mode that rewrites whole data files on each change, giving fast reads but slower writes
B.Copy-on-write — an immutable, point-in-time version of an Iceberg table listing all data files, enabling time travel and rollback
C.Copy-on-write — an Iceberg virtual table (such as files, snapshots, history, or partitions) queried to inspect a table's structure and evolution
D.Copy-on-write — an Iceberg merge-on-read delete file marking removed rows by column values such as id equals five, flexible to write but costly to apply on read
Answer + AI explanation with Pro
22. What is Merge-on-read?
Mid
A.a named, independent line of Iceberg snapshots that supports isolated writes such as a write-audit-publish staging area
B.an Iceberg update mode that records deletes and updates separately and merges them at query time for faster writes
C.an Iceberg update mode that rewrites whole data files on each change, giving fast reads but slower writes
D.an Iceberg merge-on-read delete file marking removed rows by their file path and row position, efficient to apply but tied to specific data files
Answer + AI explanation with Pro
23. Which term means: "an Iceberg update mode that records deletes and updates separately and merges them at query time for faster writes"?
Mid
A.Merge-on-read
B.Deletion vector
C.Time travel
D.Manifest file
Answer + AI explanation with Pro
24. Which statement is correct?
Mid
A.Merge-on-read — an Iceberg spec v3 binary structure that marks deleted rows compactly per data file, replacing positional delete files
B.Merge-on-read — an immutable, point-in-time version of an Iceberg table listing all data files, enabling time travel and rollback
C.Merge-on-read — an Iceberg metadata file that lists data files with their partition values and column statistics for pruning
D.Merge-on-read — an Iceberg update mode that records deletes and updates separately and merges them at query time for faster writes
Answer + AI explanation with Pro
25. What is Deletion vector?
Senior
A.an Iceberg spec v3 binary structure that marks deleted rows compactly per data file, replacing positional delete files
B.an Iceberg update mode that rewrites whole data files on each change, giving fast reads but slower writes
C.an Iceberg spec v3 feature assigning each row a stable _row_id and sequence number to support change tracking
D.the Iceberg ability to query a table as of an earlier snapshot or timestamp for reproducibility and auditing
Answer + AI explanation with Pro
26. Which term means: "an Iceberg spec v3 binary structure that marks deleted rows compactly per data file, replacing positional delete files"?
Senior
A.Deletion vector
B.Hive catalog
C.MERGE INTO
D.rewrite manifests
Answer + AI explanation with Pro
27. Which statement is correct?
Senior
A.Deletion vector — the Iceberg rewrite_data_files action that merges many small files into larger ones and can re-sort data to restore read performance
B.Deletion vector — the Iceberg catalog implementation that stores table metadata pointers in the AWS Glue Data Catalog, the most adopted Iceberg catalog as of 2025
C.Deletion vector — the open-source REST Iceberg catalog originally open sourced by Snowflake, adding role-based access control and multi-engine governance
D.Deletion vector — an Iceberg spec v3 binary structure that marks deleted rows compactly per data file, replacing positional delete files
Answer + AI explanation with Pro
28. What is Row lineage?
Senior
A.the Iceberg row-level mode that records changes as delete and data files merged at query time, favoring fast writes over fast reads
B.an Iceberg merge-on-read delete file marking removed rows by column values such as id equals five, flexible to write but costly to apply on read
C.the Iceberg SQL statement that performs upserts by matching a source against a target and inserting, updating, or deleting in one atomic operation
D.an Iceberg spec v3 feature assigning each row a stable _row_id and sequence number to support change tracking
Answer + AI explanation with Pro
29. Which term means: "an Iceberg spec v3 feature assigning each row a stable _row_id and sequence number to support change tracking"?
Senior
A.Row lineage
B.Partition evolution
C.Hidden partitioning
D.metadata table
Answer + AI explanation with Pro
30. Which statement is correct?
Senior
A.Row lineage — the open-source REST Iceberg catalog originally open sourced by Snowflake, adding role-based access control and multi-engine governance
B.Row lineage — an Iceberg spec v3 feature assigning each row a stable _row_id and sequence number to support change tracking
C.Row lineage — the Iceberg catalog that adds Git-like branches, tags, and commits over tables so multi-table changes can be versioned and merged
D.Row lineage — an Iceberg merge-on-read delete file marking removed rows by column values such as id equals five, flexible to write but costly to apply on read
Answer + AI explanation with Pro
Showing 30 of 96 Iceberg questions — the full set, with answers, explanations and an AI tutor on every question, is inside.
Free to start
Start with a free readiness check
Sign up free for the 2-minute IT readiness check and a scored result. Answers, explanations and the AI tutor on every Iceberg question come with Pro.