nerdexam
Cloudera

CCA-500 · Question #1

CCA-500 Question #1: Real Exam Question with Answer & Explanation

Sign in or unlock CCA-500 to reveal the answer and full explanation for question #1. The question stem and answer options stay visible for context.

Question

You need to analyze 60,000,000 images stored in JPEG format, each of which is approximately 25 KB. Because you Hadoop cluster isn't optimized for storing and processing many small files, you decide to do the following actions: 1. Group the individual images into a set of larger files 2. Use the set of larger files as input for a MapReduce job that processes them directly with python using Hadoop streaming. Which data serialization system gives the flexibility to do this?

Options

  • ACSV
  • BXML
  • CHTML
  • DAvro
  • ESequenceFiles
  • FJSON

Unlock CCA-500 to see the answer

You've previewed enough free CCA-500 questions. Unlock CCA-500 for full answers, explanations, the timed quiz mode, progress tracking, and the master PDF. Question stem and options stay visible so you can still see what's on the exam.

Full CCA-500 Practice