Revenue generated from a particular unit
The department in which the unit was purchased
The individual items purchased
10. Mya is investigating the factors that impact soda consumption. She examines a host of variables that help explain the
amount consumed. Which type of data mining methodology is she most likely to use?
11. The predicted value from a logistic regression will be:
12. Bridget has partitioned data into two subsets. The original file contains 300,000 observations. The subset she is
currently working with has 60,000 observations. Which subset is she most likely to be using?
13. Which methodology is used to group products that customers purchase together?
14. The testing set in data partitioning is:
The first subset of data, which usually contains 70% of the records
The second subset of data, which usually contains less than 70% of the records
The initial dataset from which subsets are created
The first subset of data, which usually contains 30% of the records
15. The higher the “score” for a particular member in logistic regression, the:
higher the likelihood that member is in category 1
lower the likelihood that member is in category 1
higher the likelihood that member is in category 0
higher the likelihood that member is not in a category