WorksheetsData Mining (Quiz 1)
Total questions: 20
Worksheet time: 10mins
Identify the CORRECT data mining tasks based on the scenario given:
Case 1: To identify items that bought together based on
customer transaction’s ID
Case 2: To find the similar symptoms among COVID 19 patient
Case 3: To predict whether the target customer prefer to buy a
new smartphone.
1:Clustering 2:Association Analysis 3:Classification
1:Association Analysis 2:Clustering 3:Classification
1:Association Analysis 2:Classification 3:Clustering
1:Classification 2:Clustering 3:Classification
"The oil and gas company want to categorize the petroleum data into distinct groups in order to find the best rank of petrol."
Find the CORRECT category based on the statement above.
Supervised learning
Unsupervised learning
Continuous learning
Categorical learning
Find the BEST term for flower in the example to predict the flower category based on the flower characteristic.
Outcome
Attribute
Data
Knowledge
Which of the following is NOT a problem in mining huge amount of data?
Cost of the learning set – related to sample size for training and cost to achieve
good accuracy.
Time and speed – time taken to achieve certain level of accuracy or for
evaluation.
Interoperability – the ability of model to operate in different platform
Integration - the ability to merge different sources of data type.
Which of the following is NOT a data mining task?
Analyzing the relationship of Twitter users towards a particular brands.
Forecasting the air traffic levels based on existing routes.
Segmenting the customers of a company according to their age.
Classifying the propensity of Covid-19 patient to have prolong Covid or not.
What are the tasks of Data Mining?
Association and correctional analysis classification
Prediction and characterization
Cluster analysis and evolution analysis
Market basket analysis
All of the above
Which of the following refers to the problem of finding abstracted patterns (or structures) in the unlabeled data?
Supervised learning
Unsupervised learning
Hybrid learning
Reinforcement learning
The supervised algorithms for data mining are
Linear Regression and K-means Clustering
Decision Tree and K-nearest neighbor
Polynomial Regression and Hierarchical clustering
Apriori and Random Forest
What are the tasks of Data Mining?
Prediction and categorization
Cluster analysis and summarization
Association and Market basket analysis
Anomaly and Deviation analysis
All the options
Searching patterns from the uncategorized data refers to ____________.
Supervised learning
Unsupervised learning
Hybrid learning
Reinforcement learning
Which of the following option is the correct combination of the pre-processing tasks in the knowledge discovery process?
Data transformation, Data cleaning, Data selection, Modeling
Data cleaning, Data selection, Classification, Data transformation
Data selection, Data cleaning, Data transformation, Data integration
Data integration, Data transformation, Data cleaning, Clustering
In predicting the cases of COVID-19, the final number of total patients can be considered as the __________.
Features
Observation
Attribute
Outcome
The efficiency of data mining algorithms is based on execution time, meanwhile, the scalability of data mining algorithms refers to the increment of attributes or instances. The statement is best described as ____________ of the algorithms.
robustness
performance
diverse data types
high dimensionality
Which of the following statement is FALSE?
Data mining can best be described as business intelligence (BI) technology that has various techniques to extract comprehensible, hidden and useful information from a population of data.
Data mining is an Artificial Intelligence powered tool that can discover useful information
from human that can then be used to improve actions.
The output of a data mining task can be in the form of patterns, trends or rules that are implicit in the data.
Data mining can discover the anomalies that might be significant in a particular business.
“Children over age of 2 whose Body Mass Index (BMI) is less than the 5th percentile are considered underweight”. This statement refers to _____________.
data
information
knowledge
wisdom
The following are the problems in mining huge amount of data EXCEPT
Cost of the learning set – related to sample size for training and cost to achieve good
accuracy.
Time and speed–time taken to achieve a certain level of accuracy or for evaluation.
Ethics - Proliferation of security and privacy concerns by organizations.
Integration - the ability to merge different sources of data type.
Which one of the following correctly refers to the task of the classification?
Classification is a learning process without training and testing.
Determining a set of instances into categories.
Partitioning of a set of objects into several classes.
Segmenting the objects into several subgroups.
Which of the following is NOT a data mining task?
Analyzing the relationship of Facebook users towards a particular product.
Forecasting the air traffic levels based on existing routes.
Dividing the customers of a company according to their location.
Classifying the propensity of the COVID-19 patient to have prolong COVID-19 or not.
Which of the following is NOT a type of Text Data Mining?
Keyword – based association analysis which finding a set of words that are frequently appeared in the set of documents related to COVID-19
Market Basket Analysis – is one of the key techniques used by large retailers to uncover associations between items. It works by looking for combinations of items that occur together frequently in transactions.
Similarity Detection – finding a set of similar text and words in the documents.
Clustering – extracting valuable information from medical literature by grouping similar examples based on set of keywords.
What are the tasks of Data Mining?
Association and correctional analysis classification
Prediction and characterization
Cluster analysis and evolution analysis
Market basket analysis
