Apache-Hadoop-Developer試験無料問題集「Hortonworks Hadoop 2.0 Certification exam for Pig and Hive Developer 認定」

You are developing a MapReduce job for sales reporting. The mapper will process input keys representing the year (IntWritable) and input values representing product indentifies (Text).
Indentify what determines the data types used by the Mapper for a given job.

解説: (GoShiken メンバーにのみ表示されます)
You write MapReduce job to process 100 files in HDFS. Your MapReduce algorithm uses TextInputFormat: the mapper applies a regular expression over input values and emits key-values pairs with the key consisting of the matching text, and the value containing the filename and byte offset. Determine the difference between setting the number of reduces to one and settings the number of reducers to zero.

解説: (GoShiken メンバーにのみ表示されます)
Analyze each scenario below and indentify which best describes the behavior of the default partitioner?

解説: (GoShiken メンバーにのみ表示されます)
Which one of the following statements is true regarding a MapReduce job?

Consider the following two relations, A and B.

What is the output of the following Pig commands?
X = GROUP A BY S1; DUMP X;

What is a SequenceFile?

解説: (GoShiken メンバーにのみ表示されます)