Your client application submits a MapReduce job to your Hadoop cluster. Identify the Hadoop daemon on which the Hadoop framework will look for an available slot schedule a MapReduce operation.
A. NameNode
B. DataNode
C. Secondary NameNode
D. TaskTracker
E. JobTracker
正解:E
解説: (Pass4Test メンバーにのみ表示されます)
質問 2:
Which describes how a client reads a file from HDFS?
A. The client contacts the NameNode for the block location(s). The NameNode contacts the DataNode that holds the requested data block. Data is transferred from the DataNode to the NameNode, and then from the NameNode to the client.
B. The client queries the NameNode for the block location(s). The NameNode returns the block location(s) to the client. The client reads the data directory off the DataNode(s).
C. The client queries all DataNodes in parallel. The DataNode that contains the requested data responds directly to the client. The client reads the data directly off the DataNode.
D. The client contacts the NameNode for the block location(s). The NameNode then queries the DataNodes for block locations. The DataNodes respond to the NameNode, and the NameNode redirects the client to the DataNode that holds the requested data block(s). The client then reads the data directly off the DataNode.
正解:B
解説: (Pass4Test メンバーにのみ表示されます)
質問 3:
Assuming default settings, which best describes the order of data provided to a reducer's reduce method:
A. The keys given to a reducer aren't in a predictable order, but the values associated with those keys always are.
B. The keys given to a reducer are in sorted order but the values associated with each key are in no predictable order
C. Both the keys and values passed to a reducer always appear in sorted order.
D. Neither keys nor values are in any predictable order.
正解:B
解説: (Pass4Test メンバーにのみ表示されます)
質問 4:
You want to perform analysis on a large collection of images. You want to store this data in HDFS and process it with MapReduce but you also want to give your data analysts and data scientists the ability to process the data directly from HDFS with an interpreted high-level programming language like Python. Which format should you use to store this data in HDFS?
A. CSV
B. XML
C. HTML
D. JSON
E. Avro
F. SequenceFiles
正解:E
解説: (Pass4Test メンバーにのみ表示されます)
質問 5:
You need to perform statistical analysis in your MapReduce job and would like to call methods in the Apache Commons Math library, which is distributed as a 1.3 megabyte Java archive (JAR) file. Which is the best way to make this library available to your MapReducer job at runtime?
A. Have your system administrator copy the JAR to all nodes in the cluster and set its location in the HADOOP_CLASSPATH environment variable before you submit your job.
B. Package your code and the Apache Commands Math library into a zip file named JobJar.zip
C. When submitting the job on the command line, specify the -libjars option followed by the JAR file path.
D. Have your system administrator place the JAR file on a Web server accessible to all cluster nodes and then set the HTTP_JAR_URL environment variable to its location.
正解:C
解説: (Pass4Test メンバーにのみ表示されます)
質問 6:
In a MapReduce job with 500 map tasks, how many map task attempts will there be?
A. At most 500.
B. Exactly 500.
C. Between 500 and 1000.
D. It depends on the number of reduces in the job.
E. At least 500.
正解:E
解説: (Pass4Test メンバーにのみ表示されます)
787 お客様のコメント





冈*绫 -
CCD-410問題集一つで万全の試験対策が出来て素敵な問題集になっている。受験直前までの仕上げ学習をガッチリサポート!