Databricks Certified Associate Developer for Apache Spark 3.0 - Associate-Developer-Apache-Spark무료 덤프문제 풀어보기

Which of the following code blocks returns all unique values across all values in columns value and productId in DataFrame transactionsDf in a one-column DataFrame?

정답: E
설명: (Fast2test 회원만 볼 수 있음)
Which of the following code blocks saves DataFrame transactionsDf in location /FileStore/transactions.csv as a CSV file and throws an error if a file already exists in the location?

정답: C
설명: (Fast2test 회원만 볼 수 있음)
Which of the following code blocks returns a DataFrame that has all columns of DataFrame transactionsDf and an additional column predErrorSquared which is the squared value of column predError in DataFrame transactionsDf?

정답: D
설명: (Fast2test 회원만 볼 수 있음)
Which of the following statements about RDDs is incorrect?

정답: E
설명: (Fast2test 회원만 볼 수 있음)
The code block shown below should return a new 2-column DataFrame that shows one attribute from column attributes per row next to the associated itemName, for all suppliers in column supplier whose name includes Sports. Choose the answer that correctly fills the blanks in the code block to accomplish this.
Sample of DataFrame itemsDf:
1.+------+----------------------------------+-----------------------------+-------------------+
2.|itemId|itemName |attributes |supplier |
3.+------+----------------------------------+-----------------------------+-------------------+
4.|1 |Thick Coat for Walking in the Snow|[blue, winter, cozy] |Sports Company Inc.|
5.|2 |Elegant Outdoors Summer Dress |[red, summer, fresh, cooling]|YetiX |
6.|3 |Outdoors Backpack |[green, summer, travel] |Sports Company Inc.|
7.+------+----------------------------------+-----------------------------+-------------------+ Code block:
itemsDf.__1__(__2__).select(__3__, __4__)

정답: D
설명: (Fast2test 회원만 볼 수 있음)
Which of the following describes characteristics of the Dataset API?

정답: E
설명: (Fast2test 회원만 볼 수 있음)
Which is the highest level in Spark's execution hierarchy?

정답: A
The code block displayed below contains an error. The code block should use Python method find_most_freq_letter to find the letter present most in column itemName of DataFrame itemsDf and return it in a new column most_frequent_letter. Find the error.
Code block:
1. find_most_freq_letter_udf = udf(find_most_freq_letter)
2. itemsDf.withColumn("most_frequent_letter", find_most_freq_letter("itemName"))

정답: A
설명: (Fast2test 회원만 볼 수 있음)
Which of the following code blocks returns a DataFrame that matches the multi-column DataFrame itemsDf, except that integer column itemId has been converted into a string column?

정답: D
설명: (Fast2test 회원만 볼 수 있음)
Which of the following code blocks creates a new one-column, two-row DataFrame dfDates with column date of type timestamp?

정답: C
설명: (Fast2test 회원만 볼 수 있음)
Which of the following code blocks sorts DataFrame transactionsDf both by column storeId in ascending and by column productId in descending order, in this priority?

정답: E
설명: (Fast2test 회원만 볼 수 있음)
The code block shown below should return a copy of DataFrame transactionsDf without columns value and productId and with an additional column associateId that has the value 5. Choose the answer that correctly fills the blanks in the code block to accomplish this.
transactionsDf.__1__(__2__, __3__).__4__(__5__, 'value')

정답: E
설명: (Fast2test 회원만 볼 수 있음)
Which of the following is a problem with using accumulators?

정답: B
설명: (Fast2test 회원만 볼 수 있음)
The code block displayed below contains an error. The code block should return a new DataFrame that only contains rows from DataFrame transactionsDf in which the value in column predError is at least 5. Find the error.
Code block:
transactionsDf.where("col(predError) >= 5")

정답: B
설명: (Fast2test 회원만 볼 수 있음)
Which of the following describes a valid concern about partitioning?

정답: D
설명: (Fast2test 회원만 볼 수 있음)

우리와 연락하기

문의할 점이 있으시면 메일을 보내오세요. 12시간이내에 답장드리도록 하고 있습니다.

근무시간: ( UTC+9 ) 9:00-24:00
월요일~토요일

서포트: 바로 연락하기 

English Deutsch 繁体中文 日本語