48 Cloud Stack questions from the Data Engineering bank, written for Indian campus drives and tech interviews. Every question has a verified answer and an AI-tutor explanation on placd.
Free to start: the 2-minute IT readiness check — six questions and a result.
A.A serverless AWS ETL and data-integration service with a managed Spark runtime and a central Data Catalog for schema metadata.
B.Microsoft's unified SaaS analytics platform combining data engineering, warehousing, and BI on a single lake foundation called OneLake.
C.Google Cloud's object storage service used as the data-lake layer for BigQuery, Dataflow, and Spark workloads.
D.Azure Data Lake Storage Gen2, object storage with a hierarchical namespace optimized for big-data analytics workloads.
Answer + AI explanation with Pro
2. Which term means: "A serverless AWS ETL and data-integration service with a managed Spark runtime and a central Data Catalog for schema metadata."?
Junior
A.Synapse Analytics
B.Azure Data Factory
C.AWS Glue
D.S3
Answer + AI explanation with Pro
3. Which statement is correct?
Junior
A.AWS Glue — Azure's managed data-integration service for building ETL and ELT pipelines with code-free mapping data flows and many connectors.
B.AWS Glue — Azure Data Lake Storage Gen2, object storage with a hierarchical namespace optimized for big-data analytics workloads.
C.AWS Glue — A serverless AWS ETL and data-integration service with a managed Spark runtime and a central Data Catalog for schema metadata.
D.AWS Glue — Microsoft Fabric's single, tenant-wide logical data lake that stores all workspace data in open Delta format with shortcuts to external sources.
Answer + AI explanation with Pro
4. What is Athena?
Junior
A.Azure's managed data-integration service for building ETL and ELT pipelines with code-free mapping data flows and many connectors.
B.A layered lakehouse design organizing data into bronze (raw), silver (cleaned), and gold (curated) zones of increasing quality and aggregation.
C.An AWS serverless query service that runs SQL directly over data in S3 using Presto/Trino, billed per data scanned.
D.An open table format adding ACID transactions, schema evolution, hidden partitioning, and time travel to files in object storage.
Answer + AI explanation with Pro
5. Which term means: "An AWS serverless query service that runs SQL directly over data in S3 using Presto/Trino, billed per data scanned."?
Junior
A.medallion architecture
B.OneLake
C.Athena
D.Dataproc
Answer + AI explanation with Pro
6. Which statement is correct?
Junior
A.Athena — Google Cloud's managed Apache Airflow service for authoring and scheduling workflows.
B.Athena — Google Cloud's managed, autoscaling service for running Apache Beam batch and streaming pipelines.
C.Athena — An AWS serverless query service that runs SQL directly over data in S3 using Presto/Trino, billed per data scanned.
D.Athena — Azure's managed data-integration service for building ETL and ELT pipelines with code-free mapping data flows and many connectors.
Answer + AI explanation with Pro
7. What is Dataflow?
Junior
A.A layered lakehouse design organizing data into bronze (raw), silver (cleaned), and gold (curated) zones of increasing quality and aggregation.
B.Amazon's object storage service, the durable, scalable foundation for data lakes on AWS.
C.Google Cloud's managed, autoscaling service for running Apache Beam batch and streaming pipelines.
D.An in-process columnar analytics database (the SQLite of OLAP) ideal for fast local analysis of Parquet and CSV without a server.
Answer + AI explanation with Pro
8. Which term means: "Google Cloud's managed, autoscaling service for running Apache Beam batch and streaming pipelines."?
Junior
A.Cloud Storage
B.Synapse Analytics
C.catalog
D.Dataflow
Answer + AI explanation with Pro
9. Which statement is correct?
Junior
A.Dataflow — A layered lakehouse design organizing data into bronze (raw), silver (cleaned), and gold (curated) zones of increasing quality and aggregation.
B.Dataflow — An AWS serverless query service that runs SQL directly over data in S3 using Presto/Trino, billed per data scanned.
C.Dataflow — A serverless AWS ETL and data-integration service with a managed Spark runtime and a central Data Catalog for schema metadata.
D.Dataflow — Google Cloud's managed, autoscaling service for running Apache Beam batch and streaming pipelines.
Answer + AI explanation with Pro
10. What is Cloud Composer?
Junior
A.Google Cloud's managed Apache Airflow service for authoring and scheduling workflows.
B.Microsoft's unified SaaS analytics platform combining data engineering, warehousing, and BI on a single lake foundation called OneLake.
C.Google Cloud's object storage service used as the data-lake layer for BigQuery, Dataflow, and Spark workloads.
D.A layered lakehouse design organizing data into bronze (raw), silver (cleaned), and gold (curated) zones of increasing quality and aggregation.
Answer + AI explanation with Pro
11. Which term means: "Google Cloud's managed Apache Airflow service for authoring and scheduling workflows."?
Junior
A.Microsoft Fabric
B.Dataproc
C.Cloud Composer
D.catalog
Answer + AI explanation with Pro
12. Which statement is correct?
Junior
A.Cloud Composer — Microsoft Fabric's single, tenant-wide logical data lake that stores all workspace data in open Delta format with shortcuts to external sources.
B.Cloud Composer — Azure's integrated analytics service combining data warehousing and big-data Spark processing, now converging into Microsoft Fabric.
C.Cloud Composer — A service tracking table metadata and pointers (such as Iceberg REST catalog, Unity Catalog, or Glue) so multiple engines query the same tables consistently.
D.Cloud Composer — Google Cloud's managed Apache Airflow service for authoring and scheduling workflows.
Answer + AI explanation with Pro
13. What is Azure Data Factory?
Junior
A.Amazon's object storage service, the durable, scalable foundation for data lakes on AWS.
B.Azure's managed data-integration service for building ETL and ELT pipelines with code-free mapping data flows and many connectors.
C.An in-process columnar analytics database (the SQLite of OLAP) ideal for fast local analysis of Parquet and CSV without a server.
D.An AWS serverless query service that runs SQL directly over data in S3 using Presto/Trino, billed per data scanned.
Answer + AI explanation with Pro
14. Which term means: "Azure's managed data-integration service for building ETL and ELT pipelines with code-free mapping data flows and many connectors."?
Junior
A.Cloud Storage
B.Synapse Analytics
C.OneLake
D.Azure Data Factory
Answer + AI explanation with Pro
15. Which statement is correct?
Junior
A.Azure Data Factory — Google Cloud's object storage service used as the data-lake layer for BigQuery, Dataflow, and Spark workloads.
B.Azure Data Factory — Azure's managed data-integration service for building ETL and ELT pipelines with code-free mapping data flows and many connectors.
C.Azure Data Factory — A service tracking table metadata and pointers (such as Iceberg REST catalog, Unity Catalog, or Glue) so multiple engines query the same tables consistently.
D.Azure Data Factory — An in-process columnar analytics database (the SQLite of OLAP) ideal for fast local analysis of Parquet and CSV without a server.
Answer + AI explanation with Pro
16. What is Microsoft Fabric?
Mid
A.An AWS serverless query service that runs SQL directly over data in S3 using Presto/Trino, billed per data scanned.
B.Google Cloud's managed Spark and Hadoop service for running open-source big-data engines on ephemeral or autoscaling clusters.
C.Microsoft's unified SaaS analytics platform combining data engineering, warehousing, and BI on a single lake foundation called OneLake.
D.Google Cloud's object storage service used as the data-lake layer for BigQuery, Dataflow, and Spark workloads.
Answer + AI explanation with Pro
17. Which term means: "Microsoft's unified SaaS analytics platform combining data engineering, warehousing, and BI on a single lake foundation called OneLake."?
Mid
A.Microsoft Fabric
B.Apache Iceberg
C.S3
D.Dataproc
Answer + AI explanation with Pro
18. Which statement is correct?
Mid
A.Microsoft Fabric — Google Cloud's object storage service used as the data-lake layer for BigQuery, Dataflow, and Spark workloads.
B.Microsoft Fabric — Microsoft's unified SaaS analytics platform combining data engineering, warehousing, and BI on a single lake foundation called OneLake.
C.Microsoft Fabric — Azure Data Lake Storage Gen2, object storage with a hierarchical namespace optimized for big-data analytics workloads.
D.Microsoft Fabric — Google Cloud's managed Apache Airflow service for authoring and scheduling workflows.
Answer + AI explanation with Pro
19. What is OneLake?
Mid
A.Azure's integrated analytics service combining data warehousing and big-data Spark processing, now converging into Microsoft Fabric.
B.A layered lakehouse design organizing data into bronze (raw), silver (cleaned), and gold (curated) zones of increasing quality and aggregation.
C.Microsoft's unified SaaS analytics platform combining data engineering, warehousing, and BI on a single lake foundation called OneLake.
D.Microsoft Fabric's single, tenant-wide logical data lake that stores all workspace data in open Delta format with shortcuts to external sources.
Answer + AI explanation with Pro
20. Which term means: "Microsoft Fabric's single, tenant-wide logical data lake that stores all workspace data in open Delta format with shortcuts to external sources."?
Mid
A.ADLS Gen2
B.Apache Iceberg
C.OneLake
D.Synapse Analytics
Answer + AI explanation with Pro
21. Which statement is correct?
Mid
A.OneLake — Azure's managed data-integration service for building ETL and ELT pipelines with code-free mapping data flows and many connectors.
B.OneLake — An AWS serverless query service that runs SQL directly over data in S3 using Presto/Trino, billed per data scanned.
C.OneLake — A service tracking table metadata and pointers (such as Iceberg REST catalog, Unity Catalog, or Glue) so multiple engines query the same tables consistently.
D.OneLake — Microsoft Fabric's single, tenant-wide logical data lake that stores all workspace data in open Delta format with shortcuts to external sources.
Answer + AI explanation with Pro
22. What is Synapse Analytics?
Junior
A.A service tracking table metadata and pointers (such as Iceberg REST catalog, Unity Catalog, or Glue) so multiple engines query the same tables consistently.
B.Microsoft Fabric's single, tenant-wide logical data lake that stores all workspace data in open Delta format with shortcuts to external sources.
C.Google Cloud's managed Apache Airflow service for authoring and scheduling workflows.
D.Azure's integrated analytics service combining data warehousing and big-data Spark processing, now converging into Microsoft Fabric.
Answer + AI explanation with Pro
23. Which term means: "Azure's integrated analytics service combining data warehousing and big-data Spark processing, now converging into Microsoft Fabric."?
Junior
A.OneLake
B.Athena
C.Synapse Analytics
D.DuckDB
Answer + AI explanation with Pro
24. Which statement is correct?
Junior
A.Synapse Analytics — Google Cloud's managed Apache Airflow service for authoring and scheduling workflows.
B.Synapse Analytics — A layered lakehouse design organizing data into bronze (raw), silver (cleaned), and gold (curated) zones of increasing quality and aggregation.
C.Synapse Analytics — Microsoft's unified SaaS analytics platform combining data engineering, warehousing, and BI on a single lake foundation called OneLake.
D.Synapse Analytics — Azure's integrated analytics service combining data warehousing and big-data Spark processing, now converging into Microsoft Fabric.
Answer + AI explanation with Pro
25. What is ADLS Gen2?
Mid
A.Microsoft's unified SaaS analytics platform combining data engineering, warehousing, and BI on a single lake foundation called OneLake.
B.Azure Data Lake Storage Gen2, object storage with a hierarchical namespace optimized for big-data analytics workloads.
C.An open table format adding ACID transactions, schema evolution, hidden partitioning, and time travel to files in object storage.
D.Azure's managed data-integration service for building ETL and ELT pipelines with code-free mapping data flows and many connectors.
Answer + AI explanation with Pro
26. Which term means: "Azure Data Lake Storage Gen2, object storage with a hierarchical namespace optimized for big-data analytics workloads."?
Mid
A.OneLake
B.Apache Iceberg
C.Azure Data Factory
D.ADLS Gen2
Answer + AI explanation with Pro
27. Which statement is correct?
Mid
A.ADLS Gen2 — Microsoft's unified SaaS analytics platform combining data engineering, warehousing, and BI on a single lake foundation called OneLake.
B.ADLS Gen2 — Google Cloud's managed, autoscaling service for running Apache Beam batch and streaming pipelines.
C.ADLS Gen2 — An AWS serverless query service that runs SQL directly over data in S3 using Presto/Trino, billed per data scanned.
D.ADLS Gen2 — Azure Data Lake Storage Gen2, object storage with a hierarchical namespace optimized for big-data analytics workloads.
Answer + AI explanation with Pro
28. What is DuckDB?
Junior
A.An AWS serverless query service that runs SQL directly over data in S3 using Presto/Trino, billed per data scanned.
B.An in-process columnar analytics database (the SQLite of OLAP) ideal for fast local analysis of Parquet and CSV without a server.
C.Microsoft Fabric's single, tenant-wide logical data lake that stores all workspace data in open Delta format with shortcuts to external sources.
D.Google Cloud's managed Apache Airflow service for authoring and scheduling workflows.
Answer + AI explanation with Pro
29. Which term means: "An in-process columnar analytics database (the SQLite of OLAP) ideal for fast local analysis of Parquet and CSV without a server."?
Junior
A.DuckDB
B.Microsoft Fabric
C.Synapse Analytics
D.Dataflow
Answer + AI explanation with Pro
30. Which statement is correct?
Junior
A.DuckDB — An in-process columnar analytics database (the SQLite of OLAP) ideal for fast local analysis of Parquet and CSV without a server.
B.DuckDB — Azure Data Lake Storage Gen2, object storage with a hierarchical namespace optimized for big-data analytics workloads.
C.DuckDB — A service tracking table metadata and pointers (such as Iceberg REST catalog, Unity Catalog, or Glue) so multiple engines query the same tables consistently.
D.DuckDB — Azure's integrated analytics service combining data warehousing and big-data Spark processing, now converging into Microsoft Fabric.
Answer + AI explanation with Pro
Showing 30 of 48 Cloud Stack questions — the full set, with answers, explanations and an AI tutor on every question, is inside.
Free to start
Start with a free readiness check
Sign up free for the 2-minute IT readiness check and a scored result. Answers, explanations and the AI tutor on every Cloud Stack question come with Pro.