A data engineer needs to maintain a central metadata repository that users access through Amazon EMR and Amazon Athena queries. The repository needs to provide the schema and properties of many tables. Some of the metadata is stored in Apache Hive. The data engineer needs to import the metadata from Hive into the central metadata repository.
Which solution will meet these requirements with the LEAST development effort?

Question

A data engineer needs to maintain a central metadata repository that users access through Amazon EMR and Amazon Athena queries. The repository needs to provide the schema and properties of many tables. Some of the metadata is stored in Apache Hive. The data engineer needs to import the metadata from Hive into the central metadata repository.

Which solution will meet these requirements with the LEAST development effort?

Brad Jarrett · Accepted Answer

Use the AWS Glue Data Catalog.

Brad Jarrett · Answer

Use Amazon EMR and Apache Ranger.

Brad Jarrett · Answer

Use a Hive metastore on an EMR cluster.

Brad Jarrett · Answer

Use a metastore on an Amazon RDS for MySQL DB instance.

Question list

List of questions

Question 1

(0)

Question 2

(0)

Question 3

(0)

Question 4

(0)

Question 5

(0)

Question 6

(0)

Question 7

(0)

Question 8

(0)

Question 9

(0)

Question 10

(0)

Related questions

Question 70 - DEA-C01 discussion

Suggested answer: C

0 comments