nerdexam
Microsoft

70-774 · Question #12

You are analyzing taxi trips in New York City. You leverage the Azure Data Factory to create data pipelines and to orchestrate data movement. You plan to develop a predictive model for 170 million…

The correct answer is C. Interactive Hive. https://azure.microsoft.com/en-gb/blog/general-availability-of-hdinsight-interactive-query-blazing- fast-data- warehouse-style-queries-on-hyper-scale-data-2/

Develop machine learning models

Question

You are analyzing taxi trips in New York City. You leverage the Azure Data Factory to create data pipelines and to orchestrate data movement. You plan to develop a predictive model for 170 million rows (37 GB) of raw data in Apache Hive by using Microsoft R Server to identify which factors contribute to the passenger tipping behavior. All of the platforms that are used for the analysis are the same. Each worker node has eight processor cores and 26 GB of memory. Which type of Azure HDInsight cluster should you use to produce results as quickly as possible?

Options

  • AHadoop
  • BHBase
  • CInteractive Hive
  • DSpark

How the community answered

(45 responses)
  • A
    4% (2)
  • B
    4% (2)
  • C
    82% (37)
  • D
    9% (4)

Explanation

https://azure.microsoft.com/en-gb/blog/general-availability-of-hdinsight-interactive-query-blazing- fast-data- warehouse-style-queries-on-hyper-scale-data-2/

Topics

#Interactive Hive#HDInsight cluster#big data processing#performance optimization

Community Discussion

No community discussion yet for this question.

Full 70-774 Practice