Exit demo CCD-410 Cloudera Certified Developer for Apache Hadoop (CCDH) PDF format · free preview

Cloudera CCD-410 - Questions & Answers

Free preview · every answer includes a full explanation

Product page: https://prepkeys.com/ccd-410.html

Question 1
Single choice

When is the earliest point at which the reduce method of a given Reducer can be called?

A.

As soon as at least one mapper has finished processing its input split.

B.

As soon as a mapper has emitted at least one record.

C.

Not until all mappers have finished processing all records.

D.

It depends on the InputFormat used for the job.

Question 2
Single choice

Which describes how a client reads a file from HDFS?

A.

The client queries the NameNode for the block location(s). The NameNode returns the block location (s) to the client. The client reads the data directory off the DataNode(s).

B.

The client queries all DataNodes in parallel. The DataNode that contains the requested data responds directly to the client. The client reads the data directly off the DataNode.

C.

The client contacts the NameNode for the block location(s). The NameNode then queries the DataNodes for block locations. The DataNodes respond to the NameNode, and the NameNode redirects the client to the DataNode that holds the requested data block(s). The client then reads the data directly off the DataNode.

D.

The client contacts the NameNode for the block location(s). The NameNode contacts the DataNode that holds the requested data block. Data is transferred from the DataNode to the NameNode, and then from the NameNode to the client.

Question 3
Single choice

You are developing a combiner that takes as input Text keys, IntWritable values, and emits Text keys, IntWritable values.

Which interface should your class implement?

A.

Combiner <Text, IntWritable, Text, IntWritable>

B.

Mapper <Text, IntWritable, Text, IntWritable>

C.

Reducer <Text, Text, IntWritable, IntWritable>

D.

Reducer <Text, IntWritable, Text, IntWritable>

E.

Combiner <Text, Text, IntWritable, IntWritable>

Question 4
Single choice

Indentify the utility that allows you to create and run MapReduce jobs with any executable or script as the mapper and/or the reducer?

A.

Oozie

B.

Sqoop

C.

Flume

D.

Hadoop Streaming

E.

mapred

Question 5
Single choice

How are keys and values presented and passed to the reducers during a standard sort and shuffle phase of MapReduce?

A.

Keys are presented to reducer in sorted order; values for a given key are not sorted.

B.

Keys are presented to reducer in sorted order; values for a given key are sorted in ascending order.

C.

Keys are presented to a reducer in random order; values for a given key are not sorted.

D.

Keys are presented to a reducer in random order; values for a given key are sorted in ascending order.

Question 6
Single choice

Assuming default settings, which best describes the order of data provided to a reducer's reduce method:

A.

The keys given to a reducer aren't in a predictable order, but the values associated with those keys always are.

B.

Both the keys and values passed to a reducer always appear in sorted order.

C.

Neither keys nor values are in any predictable order.

D.

The keys given to a reducer are in sorted order but the values associated with each key are in no predictable order

Question 7
Single choice

You wrote a map function that throws a runtime exception when it encounters a control character in input data. The input supplied to your mapper contains twelve such characters totals, spread across five file splits. The first four file splits each have two control characters and the last split has four control characters.

Indentify the number of failed task attempts you can expect when you run the job with mapred.max.map.attempts set to 4:

A.

You will have forty-eight failed task attempts

B.

You will have seventeen failed task attempts

C.

You will have five failed task attempts

D.

You will have twelve failed task attempts

E.

You will have twenty failed task attempts

Question 8
Single choice

You want to populate an associative array in order to perform a map-side join. You've decided to put this information in a text file, place that file into the DistributedCache and read it in your Mapper before any records are processed.

Indentify which method in the Mapper you should use to implement code for reading the file and populating the associative array?

A.

combine

B.

map

C.

init

D.

configure

Question 9
Single choice

You've written a MapReduce job that will process 500 million input records and generated 500 million key- value pairs. The data is not uniformly distributed. Your MapReduce job will create a significant amount of intermediate data that it needs to transfer between mappers and reduces which is a potential bottleneck.
A custom implementation of which interface is most likely to reduce the amount of intermediate data transferred across the network?

A.

Partitioner

B.

OutputFormat

C.

WritableComparable

D.

Writable

E.

InputFormat

F.

Combiner

Question 10
Single choice

Can you use MapReduce to perform a relational join on two large tables sharing a key? Assume that the two tables are formatted as comma-separated files in HDFS.

A.

Yes.

B.

Yes, but only if one of the tables fits into memory

C.

Yes, so long as both tables fit into memory.

D.

No, MapReduce cannot perform relational operations.

E.

No, but it can be done with either Pig or Hive.

Showing 10 of 60 questions · Unlock the full set