Cloud Platforms interview questions

87 Cloud Platforms questions from the Big Data bank, written for Indian campus drives and tech interviews. Every question has a verified answer and an AI-tutor explanation on placd.

Free to start: the 2-minute IT readiness check — six questions and a result.

Take the free IT readiness check

or take a mock interview set up for this area

1. What is Amazon EMR?

Junior
  1. A.the AWS managed service for running Spark, Hive, and other big-data frameworks on provisioned or serverless clusters
  2. B.the Azure unified SaaS analytics platform that brings data engineering, warehousing, and BI together on OneLake
  3. C.the BigQuery capability to train and run machine-learning models using SQL directly on warehouse data
  4. D.the single, tenant-wide data lake underlying Microsoft Fabric where all workloads store data in open Delta format

Answer + AI explanation with Pro

2. Which term means: "the AWS managed service for running Spark, Hive, and other big-data frameworks on provisioned or serverless clusters"?

Junior
  1. A.Amazon EMR
  2. B.Dataproc Serverless
  3. C.BigLake
  4. D.Amazon Athena

Answer + AI explanation with Pro

3. Which statement is correct?

Junior
  1. A.Amazon EMR — the AWS S3 bucket type providing managed Apache Iceberg storage with built-in maintenance like compaction and snapshot expiration
  2. B.Amazon EMR — a unit of BigQuery compute capacity that the engine allocates to execute the stages of a query
  3. C.Amazon EMR — the AWS managed service for running Spark, Hive, and other big-data frameworks on provisioned or serverless clusters
  4. D.Amazon EMR — the AWS option that runs Spark and Hive jobs on automatically provisioned capacity without managing clusters

Answer + AI explanation with Pro

4. What is AWS Glue?

Junior
  1. A.the Google Cloud storage engine unifying governance and fine-grained access over open table formats across BigQuery and external engines
  2. B.the AWS serverless ETL service with a managed Spark runtime and a central data catalog for schema discovery
  3. C.the in-memory acceleration layer for BigQuery that caches data to deliver sub-second responses for dashboards
  4. D.the AWS serverless query service that runs Trino-based SQL directly over data in S3 with no clusters to manage

Answer + AI explanation with Pro

5. Which term means: "the AWS serverless ETL service with a managed Spark runtime and a central data catalog for schema discovery"?

Junior
  1. A.Azure Synapse Analytics
  2. B.Athena
  3. C.AWS Glue
  4. D.S3 Tables

Answer + AI explanation with Pro

6. Which statement is correct?

Junior
  1. A.AWS Glue — the AWS serverless ETL service with a managed Spark runtime and a central data catalog for schema discovery
  2. B.AWS Glue — the GCP fully managed service for running Apache Beam batch and streaming pipelines with autoscaling
  3. C.AWS Glue — the GCP serverless data warehouse that separates storage from compute and queries petabytes with standard SQL
  4. D.AWS Glue — the Azure unified SaaS analytics platform that brings data engineering, warehousing, and BI together on OneLake

Answer + AI explanation with Pro

7. What is Amazon Athena?

Junior
  1. A.the AWS Glue component that scans data stores, infers schemas, and populates table definitions in the Glue Data Catalog
  2. B.the GCP managed service for running Spark and Hadoop clusters, including ephemeral and serverless options
  3. C.the AWS serverless query service that runs SQL directly over data in S3 using Presto and Trino engines
  4. D.the Google Cloud service that runs Spark batch and interactive workloads without provisioning or managing clusters

Answer + AI explanation with Pro

8. Which term means: "the AWS serverless query service that runs SQL directly over data in S3 using Presto and Trino engines"?

Junior
  1. A.Trino
  2. B.OneLake shortcut
  3. C.Google BigQuery
  4. D.Amazon Athena

Answer + AI explanation with Pro

9. Which statement is correct?

Junior
  1. A.Amazon Athena — the GCP managed service for running Spark and Hadoop clusters, including ephemeral and serverless options
  2. B.Amazon Athena — the AWS serverless query service that runs SQL directly over data in S3 using Presto and Trino engines
  3. C.Amazon Athena — the AWS managed service for running Spark, Hive, and other big-data frameworks on provisioned or serverless clusters
  4. D.Amazon Athena — the AWS option that runs Spark and Hive jobs on automatically provisioned capacity without managing clusters

Answer + AI explanation with Pro

10. What is Google BigQuery?

Junior
  1. A.the GCP serverless data warehouse that separates storage from compute and queries petabytes with standard SQL
  2. B.the GCP fully managed service for running Apache Beam batch and streaming pipelines with autoscaling
  3. C.the Google Cloud service that runs Spark batch and interactive workloads without provisioning or managing clusters
  4. D.a unit of BigQuery compute capacity that the engine allocates to execute the stages of a query

Answer + AI explanation with Pro

11. Which term means: "the GCP serverless data warehouse that separates storage from compute and queries petabytes with standard SQL"?

Junior
  1. A.Google BigQuery
  2. B.Fabric shortcut
  3. C.BigQuery BI Engine
  4. D.EMR Serverless

Answer + AI explanation with Pro

12. Which statement is correct?

Junior
  1. A.Google BigQuery — the GCP serverless data warehouse that separates storage from compute and queries petabytes with standard SQL
  2. B.Google BigQuery — the Google Cloud storage engine unifying governance and fine-grained access over open table formats across BigQuery and external engines
  3. C.Google BigQuery — the GCP fully managed service for running Apache Beam batch and streaming pipelines with autoscaling
  4. D.Google BigQuery — the AWS S3 bucket type providing managed Apache Iceberg storage with built-in maintenance like compaction and snapshot expiration

Answer + AI explanation with Pro

13. What is Google Dataproc?

Mid
  1. A.the Google Cloud service that runs Spark batch and interactive workloads without provisioning or managing clusters
  2. B.the in-memory acceleration layer for BigQuery that caches data to deliver sub-second responses for dashboards
  3. C.the AWS metadata store, compatible with the Hive metastore, that many AWS analytics services share for schema information
  4. D.the GCP managed service for running Spark and Hadoop clusters, including ephemeral and serverless options

Answer + AI explanation with Pro

14. Which term means: "the GCP managed service for running Spark and Hadoop clusters, including ephemeral and serverless options"?

Mid
  1. A.Glue Data Catalog
  2. B.BigQuery BI Engine
  3. C.Google Dataflow
  4. D.Google Dataproc

Answer + AI explanation with Pro

15. Which statement is correct?

Mid
  1. A.Google Dataproc — the AWS option that runs Spark and Hive jobs on automatically provisioned capacity without managing clusters
  2. B.Google Dataproc — the Google Cloud service that runs Spark batch and interactive workloads without provisioning or managing clusters
  3. C.Google Dataproc — the AWS serverless query service that runs SQL directly over data in S3 using Presto and Trino engines
  4. D.Google Dataproc — the GCP managed service for running Spark and Hadoop clusters, including ephemeral and serverless options

Answer + AI explanation with Pro

16. What is Google Dataflow?

Mid
  1. A.the GCP fully managed service for running Apache Beam batch and streaming pipelines with autoscaling
  2. B.the BigQuery capability to train and run machine-learning models using SQL directly on warehouse data
  3. C.the distributed SQL query engine that federates queries across data lakes and databases without moving the data
  4. D.the AWS serverless ETL service with a managed Spark runtime and a central data catalog for schema discovery

Answer + AI explanation with Pro

17. Which term means: "the GCP fully managed service for running Apache Beam batch and streaming pipelines with autoscaling"?

Mid
  1. A.Google Dataflow
  2. B.Glue crawler
  3. C.Google BigQuery
  4. D.AWS Glue

Answer + AI explanation with Pro

18. Which statement is correct?

Mid
  1. A.Google Dataflow — the Amazon Redshift feature that queries external data in S3 directly, joining it with tables stored in the warehouse
  2. B.Google Dataflow — a unit of BigQuery compute capacity that the engine allocates to execute the stages of a query
  3. C.Google Dataflow — the GCP fully managed service for running Apache Beam batch and streaming pipelines with autoscaling
  4. D.Google Dataflow — the single unified, tenant-wide data lake in Microsoft Fabric where all workloads store data in open Delta and Parquet formats

Answer + AI explanation with Pro

19. What is Azure Synapse Analytics?

Mid
  1. A.a unit of BigQuery compute capacity that the engine allocates to execute the stages of a query
  2. B.the Azure platform unifying data warehousing and big-data analytics over dedicated and serverless SQL and Spark pools
  3. C.the Google Cloud service that runs Spark batch and interactive workloads without provisioning or managing clusters
  4. D.a Microsoft Fabric reference that virtualizes external data into OneLake without copying it

Answer + AI explanation with Pro

20. Which term means: "the Azure platform unifying data warehousing and big-data analytics over dedicated and serverless SQL and Spark pools"?

Mid
  1. A.Amazon Athena
  2. B.BigQuery ML
  3. C.AWS Glue
  4. D.Azure Synapse Analytics

Answer + AI explanation with Pro

21. Which statement is correct?

Mid
  1. A.Azure Synapse Analytics — the Azure platform unifying data warehousing and big-data analytics over dedicated and serverless SQL and Spark pools
  2. B.Azure Synapse Analytics — a unit of BigQuery compute capacity that the engine allocates to execute the stages of a query
  3. C.Azure Synapse Analytics — the AWS metadata store, compatible with the Hive metastore, that many AWS analytics services share for schema information
  4. D.Azure Synapse Analytics — the distributed SQL query engine that federates queries across data lakes and databases without moving the data

Answer + AI explanation with Pro

22. What is Microsoft Fabric?

Mid
  1. A.the single, tenant-wide data lake underlying Microsoft Fabric where all workloads store data in open Delta format
  2. B.the AWS option that runs Spark and Hive jobs on automatically provisioned capacity without managing clusters
  3. C.the AWS serverless ETL service with a managed Spark runtime and a central data catalog for schema discovery
  4. D.the Azure unified SaaS analytics platform that brings data engineering, warehousing, and BI together on OneLake

Answer + AI explanation with Pro

23. Which term means: "the Azure unified SaaS analytics platform that brings data engineering, warehousing, and BI together on OneLake"?

Mid
  1. A.EMR Serverless
  2. B.Glue Data Catalog
  3. C.Trino
  4. D.Microsoft Fabric

Answer + AI explanation with Pro

24. Which statement is correct?

Mid
  1. A.Microsoft Fabric — the GCP fully managed service for running Apache Beam batch and streaming pipelines with autoscaling
  2. B.Microsoft Fabric — the Azure unified SaaS analytics platform that brings data engineering, warehousing, and BI together on OneLake
  3. C.Microsoft Fabric — the distributed SQL query engine that federates queries across data lakes and databases without moving the data
  4. D.Microsoft Fabric — the Azure platform unifying data warehousing and big-data analytics over dedicated and serverless SQL and Spark pools

Answer + AI explanation with Pro

25. What is OneLake?

Mid
  1. A.the Google Cloud service that runs Spark batch and interactive workloads without provisioning or managing clusters
  2. B.the AWS serverless ETL service with a managed Spark runtime and a central data catalog for schema discovery
  3. C.the single, tenant-wide data lake underlying Microsoft Fabric where all workloads store data in open Delta format
  4. D.the Google Cloud storage engine unifying governance and fine-grained access over open table formats across BigQuery and external engines

Answer + AI explanation with Pro

26. Which term means: "the single, tenant-wide data lake underlying Microsoft Fabric where all workloads store data in open Delta format"?

Mid
  1. A.EMR Serverless
  2. B.BigQuery BI Engine
  3. C.OneLake
  4. D.S3 Tables

Answer + AI explanation with Pro

27. Which statement is correct?

Mid
  1. A.OneLake — the AWS serverless query service that runs SQL directly over data in S3 using Presto and Trino engines
  2. B.OneLake — the AWS serverless ETL service with a managed Spark runtime and a central data catalog for schema discovery
  3. C.OneLake — the single, tenant-wide data lake underlying Microsoft Fabric where all workloads store data in open Delta format
  4. D.OneLake — the in-memory acceleration layer for BigQuery that caches data to deliver sub-second responses for dashboards

Answer + AI explanation with Pro

28. What is BigLake?

Senior
  1. A.the GCP storage engine that lets BigQuery and open engines query data lake tables, including Iceberg, with unified governance
  2. B.the AWS serverless query service that runs SQL directly over data in S3 using Presto and Trino engines
  3. C.a Microsoft Fabric reference that virtualizes external data into OneLake without copying it
  4. D.the in-memory acceleration layer for BigQuery that caches data to deliver sub-second responses for dashboards

Answer + AI explanation with Pro

29. Which term means: "the GCP storage engine that lets BigQuery and open engines query data lake tables, including Iceberg, with unified governance"?

Senior
  1. A.BigLake
  2. B.Glue crawler
  3. C.AWS Glue
  4. D.Google Dataproc

Answer + AI explanation with Pro

30. Which statement is correct?

Senior
  1. A.BigLake — the Google Cloud storage engine unifying governance and fine-grained access over open table formats across BigQuery and external engines
  2. B.BigLake — the GCP storage engine that lets BigQuery and open engines query data lake tables, including Iceberg, with unified governance
  3. C.BigLake — the AWS option that runs Spark and Hive jobs on automatically provisioned capacity without managing clusters
  4. D.BigLake — the AWS serverless query service that runs SQL directly over data in S3 using Presto and Trino engines

Answer + AI explanation with Pro

Showing 30 of 87 Cloud Platforms questions — the full set, with answers, explanations and an AI tutor on every question, is inside.

Free to start

Start with a free readiness check

Sign up free for the 2-minute IT readiness check and a scored result. Answers, explanations and the AI tutor on every Cloud Platforms question come with Pro.

Take the free IT readiness check

or take a mock interview set up for this area

24,000+ questions & coding problemsSoftware & IT16,274 questionsGovernment jobs26 examsAptitudenew questions every timeAI practice interviewwith feedback65 topics to practiseMechanical1,149 questionsGATE ME9 papersEngineering Mathematics381 questions2-minute checkfreeDSA Problems1,422Civil1,005 questionsGATE CE9 papersCS Fundamentals1,209 questionsYour scores6 skillsSystem Design25Electrical / EEE1,047 questionsGATE EE9 papersRun your codeC++ · Java · PythonLow-Level Design144Electronics & Comm.975 questionsGATE EC9 papersAI help on every questionFull-Stack6,282Chemical1,005 questionsGATE CH9 papersAI whiteboardsystem designWork abroadEurope · remote · transfersESE ME1 paperGATE practice papers2019–2026ESE CE1 paperDate alertsbefore the last dateESE EE1 paperBehavioural courseHR round practiceESE ET1 paperResume optimizerProSSC JE ME1 paperApplication trackerSSC JE CE1 paperCompany-wise prepSSC JE EE1 paperRole roadmapsRRB JE1 subjectPriced in ₹UPI · cardsISRO SC1 paperGATE CS9 papersIBPS SO IT1 paperUGC NET CS1 paperSSC CGL26 papersIBPS PO26 papersRRB NTPC26 papersSSC CHSL26 papersIBPS Clerk26 papersSBI Clerk26 papersRRB Group D26 papersSSC CPO26 papersSSC GD26 papers