nerdexam
Amazon

DEA-C01 · Question #35

A data engineer needs to maintain a central metadata repository that users access through Amazon EMR and Amazon Athena queries. The repository needs to provide the schema and properties of many tables

The correct answer is C. Use the AWS Glue Data Catalog.. https://aws.amazon.com/blogs/big-data/metadata-classification-lineage-and-discovery-using- apache-atlas-on-amazon-emr/

Data Store Management

Question

A data engineer needs to maintain a central metadata repository that users access through Amazon EMR and Amazon Athena queries. The repository needs to provide the schema and properties of many tables. Some of the metadata is stored in Apache Hive. The data engineer needs to import the metadata from Hive into the central metadata repository. Which solution will meet these requirements with the LEAST development effort?

Options

  • AUse Amazon EMR and Apache Ranger.
  • BUse a Hive metastore on an EMR cluster.
  • CUse the AWS Glue Data Catalog.
  • DUse a metastore on an Amazon RDS for MySQL DB instance.

How the community answered

(40 responses)
  • A
    8% (3)
  • B
    3% (1)
  • C
    85% (34)
  • D
    5% (2)

Explanation

https://aws.amazon.com/blogs/big-data/metadata-classification-lineage-and-discovery-using- apache-atlas-on-amazon-emr/

Topics

#AWS Glue Data Catalog#Metadata Management#Apache Hive#Amazon EMR

Community Discussion

No community discussion yet for this question.

Full DEA-C01 Practice