Free Alibaba ACA Big Data Certification Exam ACA-BigData1 Exam Questions
Page: 1 / 8Total 78 questions
Want more questions? Get Premium Access.
Question 1
DataWorks provides powerful scheduling capabilities including time-based or dependency-based
task trigger mechanisms to perform tens of millions of tasks accurately and punctually each day based
on DAG relationships. It supports multiple scheduling frequency configurations like: (Number of correct
answers: 4)
Score 2
Correct Answer:A. By Minute; B. By Hour; C. By Day; D. By Week
Question 2
Data Migration Unit (DMU) is used to measure the amount of resources consumed by data integration, including CPU, memory, and network. One DMU represents the minimum amount of
resources used for a data synchronization task.
Score 1
Correct Answer:A. True
Question 3
Assume that Task 1 is configured to run at 02:00 each day. In this case, the scheduling system
automatically generates a snapshot at the time predefined by the periodic node task at 23:30 each day.
That is, the instance of Task 1 will run at 02:00 the next day. If the system detects the upstream task is
complete, the system automatically runs the Task 1 instance at 02:00 the next day.
Score 1
Correct Answer:A. True
Question 4
A distributed file system like GFS and Hadoop are design to have much larger block(or chunk) size
like 64MB or 128MB, which of the following descriptions are correct? (Number of correct answers: 4)
Score 2
Correct Answer:A. It reduces clients' need to interact with the master because reads and writes on the same block( or
chunck) require only one initial request to the master for block location information; B. Since on a large block(or chunk), a client is more likely to perform many operations on a given block, it
can reduce network overhead by keeping a persistent TCP connection to the metadata server over an
extended period of time; C. It reduces the size of the metadata stored on the master; D. The servers storing those blocks may become hot spots if many clients are accessing the same small
files
Question 5
When we use the MaxCompute tunnel command to upload the log.txt file to the t_log table, the t_log is a partition table and the partitioning column is (p1 string, p2 string). Which of the following commands is correct?
A dataset includes the following items (time, region, sales amount). If you want to present the
information above in a chart, ______ is applicable.
Score 2
Correct Answer:A. Bubble Chart
Question 7
In DataWorks, a task should be instantiated first before a scheduled task is running every time, that is, generating a corresponding instance which is executed for running the scheduled task. The status is different in each phase of the scheduling process, including ________. (Number of correct answers: 3)
Correct Answer:A. Not running; B. Running; C. Running Successfully
Question 8
DataService Studio in DataWorks aims to build a data service bus to help enterprises centrally
manage private and public APIs. DataService Studio allows you to quickly create APIs based on data
tables and register existing APIs with the DataService Studio platform for centralized management and
release. Which of the following descriptions about DataService Studio in DataWorks is INCORRECT?
Score 2
Correct Answer:C. To meet the personalized query requirements of advanced users, DataService Studio provides the
custom Python script mode to allow you compile the API query by yourself. It also supports multi-table
association, complex query conditions, and aggregate functions.
Question 9
In MaxCompute, if error occurs in Tunnel transmission due to network or Tunnel service, the user
can resume the last update operation through the command
tunnel resume;.
Score 1
Correct Answer:A. True
Question 10
DataWorks can be used to develop and configure data sync tasks. Which of the following statements
are correct? (Number of correct answers: 3)
Score 2
Correct Answer:A. The data source configuration in the project management is required to add data source; B. Some of the columns in source tables can be extracted to create a mapping relationship between
fields, and constants or variables can't be added; D. Clean-up rules can be set to clear or preserve existing data before data write