โ† All ACA Flashcard Decks

Big Data and Analytics Services Flashcards

7 cards from real ACA practice questions. Tap to flip, then mark Knew It or Still Learning โ€” missed cards come back until you master them.

Read the first 7 Big Data and Analytics Services flashcards as text
  1. What does the term 'partitioned table' mean in the context of MaxCompute?

    Answer: A table divided into logical segments by partition key values to improve query performance and reduce scan costs

    In MaxCompute, a partitioned table divides data into logical partitions by one or more column values (e.g., date), allowing queries to scan only relevant partitions and reducing compute costs.

  2. Which Alibaba Cloud service provides a fully managed platform for building, training, and deploying machine learning models?

    Answer: Machine Learning Platform for AI (PAI)

    Platform for AI (PAI) is Alibaba Cloud's end-to-end machine learning platform covering data preparation, model training, evaluation, and online/batch inference deployment.

  3. In Alibaba Cloud's big data services, what is DataV primarily used for?

    Answer: Building large-screen, real-time data visualization dashboards for public displays

    DataV is Alibaba Cloud's data visualization service designed to create large-format, real-time visualization dashboards often used for command centers, exhibitions, and situation rooms.

  4. Which access control mechanism does MaxCompute use to grant fine-grained permissions on tables, projects, and resources?

    Answer: Role-based access control (RBAC) combined with policy-based ACLs

    MaxCompute uses role-based access control (RBAC) and fine-grained ACL policies to control who can access projects, tables, functions, and resources within a project.

  5. What is the primary purpose of the 'External Table' feature in MaxCompute?

    Answer: Querying data stored in external systems like OSS directly from MaxCompute without importing it

    MaxCompute External Tables allow you to query data stored in external sources such as OSS (Object Storage Service) directly using SQL without physically importing the data into MaxCompute.

  6. When using E-MapReduce (EMR) on Alibaba Cloud, where is it recommended to store large datasets to reduce costs and decouple storage from compute?

    Answer: On Alibaba Cloud OSS (Object Storage Service)

    Best practice with EMR is to store large datasets on OSS, allowing you to scale compute clusters independently, shut them down when idle, and avoid paying for persistent local storage.

  7. Which Alibaba Cloud big data service provides a Serverless Spark computing environment that allows running Spark jobs without managing clusters?

    Answer: MaxCompute Spark

    MaxCompute Spark is Alibaba Cloud's serverless Spark environment that lets you submit Spark jobs on MaxCompute infrastructure without provisioning or managing clusters.