AWS-Certified-Machine-Learning-Specialty Exam Question 116

A Machine Learning Specialist must build out a process to query a dataset on Amazon S3 using Amazon Athena. The dataset contains more than 800,000 records stored as plaintext CSV files. Each record contains
200 columns and is approximately 1.5 MB in size. Most queries will span 5 to 10 columns only.
How should the Machine Learning Specialist transform the dataset to minimize query runtime?
  • AWS-Certified-Machine-Learning-Specialty Exam Question 117

    A Machine Learning Specialist is packaging a custom ResNet model into a Docker container so the company can leverage Amazon SageMaker for training. The Specialist is using Amazon EC2 P3 instances to train the model and needs to properly configure the Docker container to leverage the NVIDIA GPUs.
    What does the Specialist need to do?
  • AWS-Certified-Machine-Learning-Specialty Exam Question 118

    A Data Science team within a large company uses Amazon SageMaker notebooks to access data stored in Amazon S3 buckets. The IT Security team is concerned that internet-enabled notebook instances create a security vulnerability where malicious code running on the instances could compromise data privacy. The company mandates that all instances stay within a secured VPC with no internet access, and data communication traffic must stay within the AWS network.
    How should the Data Science team configure the notebook instance placement to meet these requirements?
  • AWS-Certified-Machine-Learning-Specialty Exam Question 119

    A Data Scientist uses logistic regression to build a fraud detection model. While the model accuracy is 99%, 90% of the fraud cases are not detected by the model.
    What action will definitively help the model detect more than 10% of fraud cases?
  • AWS-Certified-Machine-Learning-Specialty Exam Question 120

    A company is using Amazon Polly to translate plaintext documents to speech for automated company announcements However company acronyms are being mispronounced in the current documents How should a Machine Learning Specialist address this issue for future documents'?