DEA-C01 exam dumps

DEA-C01 practice question 278 of 550

AWS Certified Data Engineer - Associate. Associate level, Amazon Web Services. Free question with the correct answer and a full explanation.

DEA-C01 Question 278

Select 2

A company needs to design a data processing pipeline to handle three types of data: structured sales records from a relational database, semi-structured customer feedback stored as JSON files, and unstructured images of product defects. Which combination of AWS services and approaches should they use to effectively model and store this data?

  1. A

    Use Amazon RDS for the sales records, Amazon S3 for the JSON files, and Amazon S3 with metadata tagging for the images

  2. B

    Use Amazon DynamoDB for the sales records, Amazon S3 for the JSON files, and Amazon Rekognition for the images

  3. C

    Use Amazon Redshift for the sales records, Amazon OpenSearch Service for the JSON files, and Amazon S3 for the images

  4. D

    Use Amazon RDS for the sales records, Amazon DynamoDB for the JSON files, and Amazon Rekognition for the images

  5. E

    Use Amazon Aurora for the sales records, Amazon S3 for the JSON files, and Amazon Rekognition for image analysis

Show answer and explanation

Correct answers: A, E

Explanation

Modeling and storing structured, semi-structured, and unstructured data requires selecting appropriate services. For structured data, relational databases like Amazon RDS or Aurora are the best fit. Semi-structured data such as JSON files can be effectively stored in Amazon S3 due to its scalability and flexibility. For unstructured data like images, Amazon S3 remains the primary storage solution, while services like Amazon Rekognition can be used for additional analysis. Using these services together ensures a well-architected and cost-effective data pipeline.

  • A. Correct.

    Correct: Amazon RDS is suitable for handling structured data like sales records from relational databases, Amazon S3 can handle semi-structured data like JSON files, and S3 with metadata tagging can effectively store and manage unstructured data like images.

  • B. Incorrect.

    Incorrect: While Amazon DynamoDB is a NoSQL database that can store semi-structured data, it is not ideal for structured sales records. Amazon Rekognition is used for image analysis, but the question focuses on data modeling and storage.

  • C. Incorrect.

    Incorrect: Amazon Redshift is a data warehouse optimized for analytics, not for transactional structured data. Amazon OpenSearch Service is used for search and analytics but is not ideal for storing semi-structured JSON data.

  • D. Incorrect.

    Incorrect: Amazon DynamoDB is not an ideal choice for storing semi-structured JSON files when S3 provides a more cost-effective and scalable solution.

  • E. Correct.

    Correct: Amazon Aurora (a relational database service) is suitable for structured sales records. Amazon S3 is ideal for storing semi-structured JSON files, and Amazon Rekognition can be used for analyzing unstructured image data, though S3 is still used for storing the images.

Timed practice exam

Take a DEA-C01 practice test under exam conditions

65 questions in 130 minutes, drawn from this bank, with a score report and a per-question review when you finish.

Start timed exam