DEA-C01 · Question #35
A data engineer needs to maintain a central metadata repository that users access through Amazon EMR and Amazon Athena queries. The repository needs to provide the schema and properties of many tables
The correct answer is C. Use the AWS Glue Data Catalog.. https://aws.amazon.com/blogs/big-data/metadata-classification-lineage-and-discovery-using- apache-atlas-on-amazon-emr/
Question
A data engineer needs to maintain a central metadata repository that users access through Amazon EMR and Amazon Athena queries. The repository needs to provide the schema and properties of many tables. Some of the metadata is stored in Apache Hive. The data engineer needs to import the metadata from Hive into the central metadata repository. Which solution will meet these requirements with the LEAST development effort?
Options
- AUse Amazon EMR and Apache Ranger.
- BUse a Hive metastore on an EMR cluster.
- CUse the AWS Glue Data Catalog.
- DUse a metastore on an Amazon RDS for MySQL DB instance.
How the community answered
(40 responses)- A8% (3)
- B3% (1)
- C85% (34)
- D5% (2)
Explanation
https://aws.amazon.com/blogs/big-data/metadata-classification-lineage-and-discovery-using- apache-atlas-on-amazon-emr/
Topics
Community Discussion
No community discussion yet for this question.