フリー問題

Databricks Certified Associate Developer for Apache Spark のフリー問題 11 / 20 問目

問題文

A job needs a small lookup dictionary available inside a transformation on every executor, and also needs a running count of malformed rows visible on the driver at the end. Which pair of shared variables fits in Apache Spark?

選択肢

  1. A broadcast variable for both, with a second broadcast created after each partition.
  2. A broadcast variable for the lookup, and an accumulator for the count.
  3. A cached DataFrame for the lookup and a global Python variable for the count.
  4. An accumulator for the lookup and a broadcast variable for the count, because accumulators are the only shared variables that can be read inside a transformation and broadcast variables are the only ones the driver can inspect after the job finishes.

解答・解説を確認するには

正解と解説の確認、回答の記録には無料登録が必要です。登録すると演習モードでフリー問題に回答し、正誤と解説をその場で確認できます。