K nearest neighbor | Operations Management homework help


The purpose of this assignment is to perform k-Nearest Neighbor classification, interpret the results, and analyze whether or not the information generated can be used to address a specific business problem.

For this assignment, you will use the “Adult Incomes” data set from the Topic Materials.

ABC Survey Company collects data via surveys that it then sells to marketing departments. Marketing departments typically do not like missing data. Since survey takers typically do not like to answer questions regarding their salary, the one question usually missing from the survey results is, “Is your annual salary $50,000 or more?”

You are the analyst who has been tasked with finding a way to impute (i.e., fill-in) the answer to the question, “Is your annual salary $50,000 or more?” This information can best be imputed based upon how individuals answer other survey questions related to their marital status, educational level, occupation, and familial relationship status. If this important question can be accurately imputed, then the worth of the survey data provided by ABC Survey Company increases dramatically.

Question 1: Using only “Marital_Status,” “Education,” “Occupation,” and “Relationship” variables, find the number of neighbors (k) that minimizes the error rate. Use a range of k between 3 and 10. Include the “k Selection Error Log” output when submitting the answer.

Question 2: Using the same variables and the k selected in Question 1, rerun the nearest neighbor model using the feature selection option in the IBM SPSS Modeler. What is the set of variables that minimize the error rate? Include the “Predictor Selection Error Log” output when submitting the answer.

Question 3: Using the value of k and the set of variables that minimizes the error rate, rerun the k-Nearest Neighbor model. What is the classification table? Include the pivot table output when submitting the answer.

Question 4: Consider the following individual: Marital_Status=Never-married, Education=Masters, Occupation=Sales, and Relationship=Not-in-family. Based on the k-Nearest Neighbor model from Question 3, how would this individual be classified? Provide the predicted income level (“>50K” or “<=50K”) and explain the process that you used to determine the income level. Include the table illustrating the data when submitting the answer.

Question 5: Describe the model building process you used to determine whether or not a particular survey taker earned an annual salary of $50,000 or more. Include discussion of the accuracy of the k-Nearest Neighbor model and how it can be used in practice to impute the answer to the question, “Is your annual salary $50,000 or more?”

General Requirements:

Submit the answers to Questions 1-5 including the specified screenshots and software outputs, in a Word document.

APA format is not required, but solid academic writing is expected.

Order a unique copy of this paper
(550 words)

Approximate price: $22

Basic features
  • Free title page and bibliography
  • Unlimited revisions
  • Plagiarism-free guarantee
  • Money-back guarantee
  • 24/7 support
On-demand options
  • Writer’s samples
  • Part-by-part delivery
  • Overnight delivery
  • Copies of used sources
  • Expert Proofreading
Paper format
  • 275 words per page
  • 12 pt Arial/Times New Roman
  • Double line spacing
  • Any citation style (APA, MLA, Chicago/Turabian, Harvard)

Our guarantees

We value our customers and so we ensure that what we do is 100% original..
With us you are guaranteed of quality work done by our qualified experts.Your information and everything that you do with us is kept completely confidential.

Money-back guarantee

You have to be 100% sure of the quality of your product to give a money-back guarantee. This describes us perfectly. Make sure that this guarantee is totally transparent.

Read more

Zero-plagiarism guarantee

The Product ordered is guaranteed to be original. Orders are checked by the most advanced anti-plagiarism software in the market to assure that the Product is 100% original. The Company has a zero tolerance policy for plagiarism.

Read more

Free-revision policy

The Free Revision policy is a courtesy service that the Company provides to help ensure Customer’s total satisfaction with the completed Order. To receive free revision the Company requires that the Customer provide the request within fourteen (14) days from the first completion date and within a period of thirty (30) days for dissertations.

Read more

Privacy policy

The Company is committed to protect the privacy of the Customer and it will never resell or share any of Customer’s personal information, including credit card data, with any third party. All the online transactions are processed through the secure and reliable online payment systems.

Read more

Fair-cooperation guarantee

By placing an order with us, you agree to the service we provide. We will endear to do all that it takes to deliver a comprehensive paper as per your requirements. We also count on your cooperation to ensure that we deliver on this mandate.

Read more

Calculate the price of your order

550 words
We'll send you the first draft for approval by September 11, 2018 at 10:52 AM
Total price:
The price is based on these factors:
Academic level
Number of pages