Limited-Time Offer: Enjoy 50% Savings! Ends in 00h 00m 00s Coupon code: 50OFF
Skip to content

Free Alibaba ACA Big Data Certification Exam ACA-BigData1 Exam Questions

Page: 1 / 8 Total 78 questions

Want more questions? Get Premium Access.

Question 1

DataWorks provides powerful scheduling capabilities including time-based or dependency-based

task trigger mechanisms to perform tens of millions of tasks accurately and punctually each day based

on DAG relationships. It supports multiple scheduling frequency configurations like: (Number of correct

answers: 4)

Score 2

Correct Answer: A. By Minute; B. By Hour; C. By Day; D. By Week

Question 2

Data Migration Unit (DMU) is used to measure the amount of resources consumed by data integration, including CPU, memory, and network. One DMU represents the minimum amount of

resources used for a data synchronization task.

Score 1

Correct Answer: A. True

Question 3

Assume that Task 1 is configured to run at 02:00 each day. In this case, the scheduling system

automatically generates a snapshot at the time predefined by the periodic node task at 23:30 each day.

That is, the instance of Task 1 will run at 02:00 the next day. If the system detects the upstream task is

complete, the system automatically runs the Task 1 instance at 02:00 the next day.

Score 1

Correct Answer: A. True

Question 4

A distributed file system like GFS and Hadoop are design to have much larger block(or chunk) size

like 64MB or 128MB, which of the following descriptions are correct? (Number of correct answers: 4)

Score 2

Correct Answer: A. It reduces clients' need to interact with the master because reads and writes on the same block( or chunck) require only one initial request to the master for block location information; B. Since on a large block(or chunk), a client is more likely to perform many operations on a given block, it can reduce network overhead by keeping a persistent TCP connection to the metadata server over an extended period of time; C. It reduces the size of the metadata stored on the master; D. The servers storing those blocks may become hot spots if many clients are accessing the same small files

Question 5

When we use the MaxCompute tunnel command to upload the log.txt file to the t_log table, the t_log is a partition table and the partitioning column is (p1 string, p2 string). Which of the following commands is correct?

Correct Answer: A. tunnel upload log.txt t_log/p1='b1'', p2='b2'

Question 6

A dataset includes the following items (time, region, sales amount). If you want to present the

information above in a chart, ______ is applicable.

Score 2

Correct Answer: A. Bubble Chart

Question 7

In DataWorks, a task should be instantiated first before a scheduled task is running every time, that is, generating a corresponding instance which is executed for running the scheduled task. The status is different in each phase of the scheduling process, including ________. (Number of correct answers: 3)

Correct Answer: A. Not running; B. Running; C. Running Successfully

Question 8

DataService Studio in DataWorks aims to build a data service bus to help enterprises centrally

manage private and public APIs. DataService Studio allows you to quickly create APIs based on data

tables and register existing APIs with the DataService Studio platform for centralized management and

release. Which of the following descriptions about DataService Studio in DataWorks is INCORRECT?

Score 2

Correct Answer: C. To meet the personalized query requirements of advanced users, DataService Studio provides the custom Python script mode to allow you compile the API query by yourself. It also supports multi-table association, complex query conditions, and aggregate functions.

Question 9

In MaxCompute, if error occurs in Tunnel transmission due to network or Tunnel service, the user

can resume the last update operation through the command

tunnel resume;.

Score 1

Correct Answer: A. True

Question 10

DataWorks can be used to develop and configure data sync tasks. Which of the following statements

are correct? (Number of correct answers: 3)

Score 2

Correct Answer: A. The data source configuration in the project management is required to add data source; B. Some of the columns in source tables can be extracted to create a mapping relationship between fields, and constants or variables can't be added; D. Clean-up rules can be set to clear or preserve existing data before data write