DEA-C01 exam dumps

DEA-C01 practice question 201 of 550

AWS Certified Data Engineer - Associate. Associate level, Amazon Web Services. Free question with the correct answer and a full explanation.

DEA-C01 Question 201

Select 3

A data engineering team is setting up an AWS Glue Data Catalog for their organization. They want to catalog their data stored in Amazon S3 buckets and make it queryable using Amazon Athena. Which steps must the team take to create and populate the Data Catalog correctly?

  1. A

    Create an AWS Glue crawler, configure it with the S3 bucket location, and run the crawler to populate the Data Catalog.

  2. B

    Manually define tables and schemas in the AWS Glue Data Catalog without using a crawler.

  3. C

    Ensure an IAM role with proper permissions to the S3 bucket and Glue service is configured and attached to the crawler.

  4. D

    Use AWS Glue ETL jobs directly to create the Data Catalog without running a crawler.

  5. E

    Integrate the AWS Glue Data Catalog with Amazon Athena by enabling the Glue Data Catalog as the query metadata source in Athena settings.

Show answer and explanation

Correct answers: A, C, E

Explanation

To create and populate an AWS Glue Data Catalog for data stored in Amazon S3, you should use a Glue crawler to scan the S3 bucket and automatically define table schemas in the Data Catalog. Proper IAM permissions must be configured to allow the crawler access to the S3 bucket and Glue service. Additionally, you need to enable the Glue Data Catalog as the metadata source in Amazon Athena to make the cataloged data queryable. These steps ensure a seamless integration between S3, Glue, and Athena for data analysis.

  • A. Correct.

    Correct. AWS Glue crawlers are used to scan data in S3 and automatically populate the Data Catalog with table definitions and schema details.

  • B. Incorrect.

    Incorrect. While manual table creation is possible, it is not the recommended or automated way to catalog large datasets, especially for S3-based storage.

  • C. Correct.

    Correct. An IAM role with the necessary permissions is required for the crawler to access the S3 bucket and update the Glue Data Catalog.

  • D. Incorrect.

    Incorrect. AWS Glue ETL jobs are designed for data transformation and processing, not for directly creating or populating the Data Catalog.

  • E. Correct.

    Correct. To make the Glue Data Catalog queryable with Amazon Athena, you need to configure Athena to use the Glue Data Catalog as the metadata source.

Timed practice exam

Take a DEA-C01 practice test under exam conditions

65 questions in 130 minutes, drawn from this bank, with a score report and a per-question review when you finish.

Start timed exam