DEA-C01 Sample Questions

DEA-C01 Sample Questions & Answers

Ingesting, transforming and orchestrating data pipelines carries the biggest share, ahead of picking the right data store and cataloging approach, keeping operations running smoothly day to day, and securing and encrypting data.

Launch the full DEA-C01 simulator →

Showing 8 of 17 free samples.

  1. Question 1Intermediate

    A data engineer maintains custom Python scripts that perform a data formatting process that many AWS Lambda functions use. When the data engineer needs to modify the Python scripts, the data engineer must manually update all the Lambda functions.The data engineer requires a less manual way to update the Lambda functions.Which solution will meet this requirement?

    Show answer & explanation

    Correct answer: B

  2. Question 2Intermediate

    A company created an extract, transform, and load (ETL) data pipeline in AWS Glue. A data engineer must crawl a table that is in Microsoft SQL Server. The data engineer needs to extract, transform, and load the output of the crawl to an Amazon S3 bucket. The data engineer also must orchestrate the data pipeline.Which AWS service or feature will meet these requirements MOST cost-effectively?

    Show answer & explanation

    Correct answer: B

  3. Question 3Intermediate

    A financial services company stores financial data in Amazon Redshift. A data engineer wants to run real-time queries on the financial data to support a web-based trading application. The data engineer wants to run the queries from within the trading application.Which solution will meet these requirements with the LEAST operational overhead?

    Show answer & explanation

    Correct answer: B

  4. Question 4Intermediate

    A company uses Amazon Athena for one-time queries against data that is in Amazon S3. The company has several use cases. The company must implement permission controls to separate query processes and access to query history among users, teams, and applications that are in the same AWS account.Which solution will meet these requirements?

    Show answer & explanation

    Correct answer: B

  5. Question 5Intermediate

    A data engineer needs to schedule a workflow that runs a set of AWS Glue jobs every day. The data engineer does not require the Glue jobs to run or finish at a specific time.Which solution will run the Glue jobs in the MOST cost-effective way?

    Show answer & explanation

    Correct answer: A

  6. Question 6Intermediate

    A data engineer needs to create an AWS Lambda function that converts the format of data from .csv to Apache Parquet. The Lambda function must run only if a user uploads a .csv file to an Amazon S3 bucket.Which solution will meet these requirements with the LEAST operational overhead?

    Show answer & explanation

    Correct answer: A

  7. Question 7Intermediate

    Data Ingestion and Transformation · Task Statement 1.1: Perform data ingestion - Managing fan-in and fan-out

    A media streaming company ingests clickstream data using Amazon Kinesis Data Streams. The data is consumed by a fleet of Amazon EC2 instances running a custom application. During peak hours, the application experiences read throttling, indicated by ReadProvisionedThroughputExceeded exceptions. The data engineering team needs to increase the read throughput for this specific consumer application without impacting other consumers or increasing the number of shards. Which feature should they implement?

    Show answer & explanation

    Correct answer: A

    Enhanced Fan-Out (EFO) allows a consumer to receive its own dedicated throughput of up to 2 MB/second per shard, independent of other consumers. This isolates the consumer from the shared 2 MB/second read limit of standard consumers, effectively resolving the throttling issue without needing to reshard the stream.

  8. Question 8Beginner

    Data Ingestion and Transformation · Task Statement 1.2: Transform and process data

    A data engineer is designing a data lake on Amazon S3. The data ingestion layer receives JSON files from various sources. The engineer wants to optimize the storage for analytics by converting these files to Apache Parquet format. The solution must be serverless, scale automatically, and allow for data transformation logic (like renaming columns) to be applied during the conversion. Which AWS service is BEST suited for this task?

    Show answer & explanation

    Correct answer: B

    AWS Glue is a fully managed, serverless ETL service. It can natively read JSON, apply complex transformations (like renaming, mapping, filtering) using PySpark or Scala, and write the output as Apache Parquet. It scales automatically based on the configured DPUs.

Ready for the real thing?

The full DEA-C01 simulator has every exam-style question, timed mode, and instant scoring.