โ† All AWS Flashcard Decks

Data Analytics & Machine Learning Flashcards

7 cards from real AWS practice questions. Tap to flip, then mark Knew It or Still Learning โ€” missed cards come back until you master them.

Read the first 7 Data Analytics & Machine Learning flashcards as text
  1. Which AWS service provides a fully managed, petabyte-scale data warehouse for running complex analytic queries?

    Answer: Amazon Redshift

    Amazon Redshift is AWS's fully managed, petabyte-scale data warehouse optimized for analytic queries.

  2. A team needs to run ad-hoc SQL queries directly against data stored in Amazon S3 without managing servers. Which service fits best?

    Answer: Amazon Athena

    Amazon Athena is a serverless query service that lets you run standard SQL directly on data in S3.

  3. Which AWS service is a fully managed extract, transform, and load (ETL) service that catalogs data and prepares it for analytics?

    Answer: AWS Glue

    AWS Glue is a serverless ETL service that includes a Data Catalog for discovering and preparing data.

  4. Which service provides business intelligence dashboards and interactive visualizations with SPICE in-memory acceleration?

    Answer: Amazon QuickSight

    Amazon QuickSight is AWS's serverless BI service that uses the SPICE engine for fast in-memory analytics.

  5. A company wants to ingest and process streaming clickstream data in real time. Which service is most appropriate?

    Answer: Amazon Kinesis Data Streams

    Amazon Kinesis Data Streams is built for real-time ingestion and processing of streaming data like clickstreams.

  6. Which fully managed service is designed to build, train, and deploy machine learning models at scale?

    Answer: Amazon SageMaker

    Amazon SageMaker is the flagship managed platform for building, training, and deploying ML models.

  7. Which AWS service runs managed Apache Spark and Hadoop clusters for big data processing?

    Answer: Amazon EMR

    Amazon EMR provides managed clusters for running big data frameworks such as Apache Spark and Hadoop.