QUESTION: 25

A company wants to classify user behavior as either fraudulent or normal. Based on internal research, a Machine Learning Specialist would like to build a binary classifier based on two features: age of account and transaction month. The class distribution for these features is illustrated in the figure provided.

Based on this information, which model would have the HIGHEST recall with respect to the fraudulent class?

Decision tree
Linear support vector machine (SVM)
Naive Bayesian classifier
Single Perceptron with sigmoidal activation function

Answer(s): C

Reveal Solution Next Question

QUESTION: 26

A Machine Learning Specialist kicks off a hyperparameter tuning job for a tree-based ensemble model using Amazon SageMaker with Area Under the ROC Curve (AUC) as the objective metric. This workflow will eventually be deployed in a pipeline that retrains and tunes hyperparameters each night to model click-through on data that goes stale every 24 hours.

With the goal of decreasing the amount of time it takes to train these models, and ultimately to decrease costs, the Specialist wants to reconfigure the input hyperparameter range(s).

Which visualization will accomplish this?

A histogram showing whether the most important input feature is Gaussian.
A scatter plot with points colored by target variable that uses t-Distributed Stochastic Neighbor Embedding (t-SNE) to visualize the large number of input variables in an easier-to-read dimension.
A scatter plot showing the performance of the objective metric over each training iteration.
A scatter plot showing the correlation between maximum tree depth and the objective metric.

Answer(s): D

Reveal Solution Next Question

QUESTION: 27

A Machine Learning Specialist is creating a new natural language processing application that processes a dataset comprised of 1 million sentences. The aim is to then run Word2Vec to generate embeddings of the sentences and enable different types of predictions.

Here is an example from the dataset:

"The quck BROWN FOX jumps over the lazy dog.”

Which of the following are the operations the Specialist needs to perform to correctly sanitize and prepare the data in a repeatable manner? (Choose three.)

Perform part-of-speech tagging and keep the action verb and the nouns only.
Normalize all words by making the sentence lowercase.
Remove stop words using an English stopword dictionary.
Correct the typography on "quck" to "quick.”
One-hot encode all words in the sentence.
Tokenize the sentence into words.

Answer(s): B,C,F

Reveal Solution Next Question

QUESTION: 28

A company is using Amazon Polly to translate plaintext documents to speech for automated company announcements. However, company acronyms are being mispronounced in the current documents.
How should a Machine Learning Specialist address this issue for future documents?

Convert current documents to SSML with pronunciation tags.
Create an appropriate pronunciation lexicon.
Output speech marks to guide in pronunciation.
Use Amazon Lex to preprocess the text files for pronunciation

Answer(s): B

Reveal Solution Next Question

Free MLS-C01 Exam Braindumps (page: 20)

QUESTION: 25

QUESTION: 26

QUESTION: 27

QUESTION: 28

MLS-C01 Exam Discussions & Posts