Hive interview questions

63 Hive questions from the Big Data bank, written for Indian campus drives and tech interviews. Every question has a verified answer and an AI-tutor explanation on placd.

Free to start: the 2-minute IT readiness check — six questions and a result.

Take the free IT readiness check

or take a mock interview set up for this area

1. What is Hive managed table?

Junior
  1. A.an optional per-column probabilistic index in ORC files that lets Hive skip stripes which cannot contain a searched value
  2. B.a Hive engine mode that processes batches of rows in columnar form to improve CPU efficiency on ORC data
  3. C.a table where Hive owns the data, so DROP deletes the underlying files
  4. D.the Hive execution engine that runs queries as a directed acyclic graph of tasks, avoiding the multi-stage disk writes of classic MapReduce

Answer + AI explanation with Pro

2. Which term means: "a table where Hive owns the data, so DROP deletes the underlying files"?

Junior
  1. A.schema-on-read
  2. B.Vectorized execution
  3. C.Hive managed table
  4. D.Cost-based optimizer

Answer + AI explanation with Pro

3. Which statement is correct?

Junior
  1. A.Hive managed table — a Hive ACID operation that merges base and delta files into new base files, applying all updates and deletes
  2. B.Hive managed table — a table where Hive owns the data, so DROP deletes the underlying files
  3. C.Hive managed table — the Hive ACID layout where base files hold compacted data and delta files record subsequent inserts and deletes
  4. D.Hive managed table — the Hive optimizer (backed by Apache Calcite) that uses table and column statistics to reorder joins and choose efficient plans

Answer + AI explanation with Pro

4. What is Hive external table?

Junior
  1. A.a table where Hive owns only metadata, so DROP keeps the data
  2. B.a Hive join optimization that joins corresponding buckets of two bucketed tables in memory, avoiding a full shuffle
  3. C.the Hive service that stores table schemas, partitions, and storage locations in a relational database for query engines to consult
  4. D.the Hive background process that merges base and delta files of transactional tables, with minor compaction merging deltas and major compaction rewriting bases

Answer + AI explanation with Pro

5. Which term means: "a table where Hive owns only metadata, so DROP keeps the data"?

Junior
  1. A.Dynamic partition insert
  2. B.Major compaction
  3. C.Vectorized execution
  4. D.Hive external table

Answer + AI explanation with Pro

6. Which statement is correct?

Junior
  1. A.Hive external table — applying a schema at query time rather than at write time
  2. B.Hive external table — a Hive join optimization that joins corresponding buckets of two bucketed tables in memory, avoiding a full shuffle
  3. C.Hive external table — an optional per-column probabilistic index in ORC files that lets Hive skip stripes which cannot contain a searched value
  4. D.Hive external table — a table where Hive owns only metadata, so DROP keeps the data

Answer + AI explanation with Pro

7. What is Hive partitioning?

Mid
  1. A.the Hive background process that merges base and delta files of transactional tables, with minor compaction merging deltas and major compaction rewriting bases
  2. B.splitting data into directories by column value for partition pruning
  3. C.applying a schema at query time rather than at write time
  4. D.an optional per-column probabilistic index in ORC files that lets Hive skip stripes which cannot contain a searched value

Answer + AI explanation with Pro

8. Which term means: "splitting data into directories by column value for partition pruning"?

Mid
  1. A.LLAP
  2. B.Hive partitioning
  3. C.Hive managed table
  4. D.Dynamic partition insert

Answer + AI explanation with Pro

9. Which statement is correct?

Mid
  1. A.Hive partitioning — splitting data into directories by column value for partition pruning
  2. B.Hive partitioning — a Hive write mode that creates partitions automatically from column values in the data being inserted
  3. C.Hive partitioning — the Hive service that stores table schemas, partitions, and storage locations in a relational database for query engines to consult
  4. D.Hive partitioning — a Hive join optimization that joins corresponding buckets of two bucketed tables in memory, avoiding a full shuffle

Answer + AI explanation with Pro

10. What is Hive bucketing?

Mid
  1. A.the Hive query optimizer that uses table statistics through Apache Calcite to pick join orders and plans
  2. B.the Hive service that stores table schemas, partitions, and storage locations in a relational database for query engines to consult
  3. C.hashing a column into a fixed number of files within a partition
  4. D.a Hive ACID operation that merges base and delta files into new base files, applying all updates and deletes

Answer + AI explanation with Pro

11. Which term means: "hashing a column into a fixed number of files within a partition"?

Mid
  1. A.Metastore
  2. B.Partition pruning
  3. C.ORC bloom filter
  4. D.Hive bucketing

Answer + AI explanation with Pro

12. Which statement is correct?

Mid
  1. A.Hive bucketing — a table where Hive owns the data, so DROP deletes the underlying files
  2. B.Hive bucketing — hashing a column into a fixed number of files within a partition
  3. C.Hive bucketing — a Hive engine mode that processes batches of rows in columnar form to improve CPU efficiency on ORC data
  4. D.Hive bucketing — the Hive ACID layout where base files hold compacted data and delta files record subsequent inserts and deletes

Answer + AI explanation with Pro

13. What is schema-on-read?

Junior
  1. A.a Hive write mode that creates partitions automatically from column values in the data being inserted
  2. B.applying a schema at query time rather than at write time
  3. C.a Hive ACID operation that merges base and delta files into new base files, applying all updates and deletes
  4. D.the Hive optimization that reads only the partition directories matching a query's filter on partition columns

Answer + AI explanation with Pro

14. Which term means: "applying a schema at query time rather than at write time"?

Junior
  1. A.ACID tables
  2. B.Hive bucketing
  3. C.schema-on-read
  4. D.ORC bloom filter

Answer + AI explanation with Pro

15. Which statement is correct?

Junior
  1. A.schema-on-read — a Hive ACID operation that merges several delta files into a single delta file without rewriting base files
  2. B.schema-on-read — applying a schema at query time rather than at write time
  3. C.schema-on-read — the Hive feature processing batches of about 1024 rows per operation instead of one row at a time to cut CPU overhead
  4. D.schema-on-read — the Hive ACID layout where base files hold compacted data and delta files record subsequent inserts and deletes

Answer + AI explanation with Pro

16. What is Metastore?

Junior
  1. A.a Hive engine mode that processes batches of rows in columnar form to improve CPU efficiency on ORC data
  2. B.the Hive query optimizer that uses table statistics through Apache Calcite to pick join orders and plans
  3. C.a Hive join optimization that joins corresponding buckets of two bucketed tables in memory, avoiding a full shuffle
  4. D.the Hive service that stores table schemas, partitions, and storage locations in a relational database for query engines to consult

Answer + AI explanation with Pro

17. Which term means: "the Hive service that stores table schemas, partitions, and storage locations in a relational database for query engines to consult"?

Junior
  1. A.Cost-based optimizer
  2. B.LLAP
  3. C.Hive bucketing
  4. D.Metastore

Answer + AI explanation with Pro

18. Which statement is correct?

Junior
  1. A.Metastore — the Hive service that stores table schemas, partitions, and storage locations in a relational database for query engines to consult
  2. B.Metastore — a Hive join optimization that joins corresponding buckets of two bucketed tables in memory, avoiding a full shuffle
  3. C.Metastore — a Hive ACID operation that merges base and delta files into new base files, applying all updates and deletes
  4. D.Metastore — the Hive Live Long and Process service that keeps daemons and cached data resident in memory for low-latency interactive queries

Answer + AI explanation with Pro

19. What is ACID tables?

Junior
  1. A.the Hive Live Long and Process service that keeps daemons and cached data resident in memory for low-latency interactive queries
  2. B.applying a schema at query time rather than at write time
  3. C.Hive transactional tables, stored as ORC, that support row-level insert, update, and delete with snapshot isolation
  4. D.the Hive service that stores table schemas, partitions, and storage locations in a relational database for query engines to consult

Answer + AI explanation with Pro

20. Which term means: "Hive transactional tables, stored as ORC, that support row-level insert, update, and delete with snapshot isolation"?

Junior
  1. A.vectorized query execution
  2. B.Partition pruning
  3. C.Minor compaction
  4. D.ACID tables

Answer + AI explanation with Pro

21. Which statement is correct?

Junior
  1. A.ACID tables — a Hive ACID operation that merges several delta files into a single delta file without rewriting base files
  2. B.ACID tables — Hive transactional tables, stored as ORC, that support row-level insert, update, and delete with snapshot isolation
  3. C.ACID tables — the Hive feature processing batches of about 1024 rows per operation instead of one row at a time to cut CPU overhead
  4. D.ACID tables — a Hive ACID operation that merges base and delta files into new base files, applying all updates and deletes

Answer + AI explanation with Pro

22. What is LLAP?

Mid
  1. A.Hive transactional tables, stored as ORC, that support row-level insert, update, and delete with snapshot isolation
  2. B.the Hive Live Long and Process service that keeps daemons and cached data resident in memory for low-latency interactive queries
  3. C.a Hive write mode that creates partitions automatically from column values in the data being inserted
  4. D.the Hive execution engine that runs queries as a directed acyclic graph of tasks, avoiding the multi-stage disk writes of classic MapReduce

Answer + AI explanation with Pro

23. Which term means: "the Hive Live Long and Process service that keeps daemons and cached data resident in memory for low-latency interactive queries"?

Mid
  1. A.Hive partitioning
  2. B.Major compaction
  3. C.ACID tables
  4. D.LLAP

Answer + AI explanation with Pro

24. Which statement is correct?

Mid
  1. A.LLAP — hashing a column into a fixed number of files within a partition
  2. B.LLAP — splitting data into directories by column value for partition pruning
  3. C.LLAP — the Hive Live Long and Process service that keeps daemons and cached data resident in memory for low-latency interactive queries
  4. D.LLAP — a Hive engine mode that processes batches of rows in columnar form to improve CPU efficiency on ORC data

Answer + AI explanation with Pro

25. What is Delta and base files?

Mid
  1. A.the Hive ACID layout where base files hold compacted data and delta files record subsequent inserts and deletes
  2. B.a Hive write mode that creates partitions automatically from column values in the data being inserted
  3. C.splitting data into directories by column value for partition pruning
  4. D.the Hive execution engine that runs queries as a directed acyclic graph of tasks, avoiding the multi-stage disk writes of classic MapReduce

Answer + AI explanation with Pro

26. Which term means: "the Hive ACID layout where base files hold compacted data and delta files record subsequent inserts and deletes"?

Mid
  1. A.ACID compaction
  2. B.bucketed map join
  3. C.Hive cost-based optimizer
  4. D.Delta and base files

Answer + AI explanation with Pro

27. Which statement is correct?

Mid
  1. A.Delta and base files — the Hive background process that merges base and delta files of transactional tables, with minor compaction merging deltas and major compaction rewriting bases
  2. B.Delta and base files — the Hive optimization that reads only the partition directories matching a query's filter on partition columns
  3. C.Delta and base files — the Hive ACID layout where base files hold compacted data and delta files record subsequent inserts and deletes
  4. D.Delta and base files — Hive transactional tables, stored as ORC, that support row-level insert, update, and delete with snapshot isolation

Answer + AI explanation with Pro

28. What is Minor compaction?

Mid
  1. A.a Hive ACID operation that merges base and delta files into new base files, applying all updates and deletes
  2. B.a Hive ACID operation that merges several delta files into a single delta file without rewriting base files
  3. C.the Hive query optimizer that uses table statistics through Apache Calcite to pick join orders and plans
  4. D.applying a schema at query time rather than at write time

Answer + AI explanation with Pro

29. Which term means: "a Hive ACID operation that merges several delta files into a single delta file without rewriting base files"?

Mid
  1. A.Dynamic partition insert
  2. B.Minor compaction
  3. C.Vectorized execution
  4. D.bucketed map join

Answer + AI explanation with Pro

30. Which statement is correct?

Mid
  1. A.Minor compaction — the Hive ACID layout where base files hold compacted data and delta files record subsequent inserts and deletes
  2. B.Minor compaction — a Hive ACID operation that merges several delta files into a single delta file without rewriting base files
  3. C.Minor compaction — the Hive feature processing batches of about 1024 rows per operation instead of one row at a time to cut CPU overhead
  4. D.Minor compaction — splitting data into directories by column value for partition pruning

Answer + AI explanation with Pro

Showing 30 of 63 Hive questions — the full set, with answers, explanations and an AI tutor on every question, is inside.

Free to start

Start with a free readiness check

Sign up free for the 2-minute IT readiness check and a scored result. Answers, explanations and the AI tutor on every Hive question come with Pro.

Take the free IT readiness check

or take a mock interview set up for this area

24,000+ questions & coding problemsSoftware & IT16,274 questionsGovernment jobs26 examsAptitudenew questions every timeAI practice interviewwith feedback65 topics to practiseMechanical1,149 questionsGATE ME9 papersEngineering Mathematics381 questions2-minute checkfreeDSA Problems1,422Civil1,005 questionsGATE CE9 papersCS Fundamentals1,209 questionsYour scores6 skillsSystem Design25Electrical / EEE1,047 questionsGATE EE9 papersRun your codeC++ · Java · PythonLow-Level Design144Electronics & Comm.975 questionsGATE EC9 papersAI help on every questionFull-Stack6,282Chemical1,005 questionsGATE CH9 papersAI whiteboardsystem designWork abroadEurope · remote · transfersESE ME1 paperGATE practice papers2019–2026ESE CE1 paperDate alertsbefore the last dateESE EE1 paperBehavioural courseHR round practiceESE ET1 paperResume optimizerProSSC JE ME1 paperApplication trackerSSC JE CE1 paperCompany-wise prepSSC JE EE1 paperRole roadmapsRRB JE1 subjectPriced in ₹UPI · cardsISRO SC1 paperGATE CS9 papersIBPS SO IT1 paperUGC NET CS1 paperSSC CGL26 papersIBPS PO26 papersRRB NTPC26 papersSSC CHSL26 papersIBPS Clerk26 papersSBI Clerk26 papersRRB Group D26 papersSSC CPO26 papersSSC GD26 papers