Question: use this document/link https://drive.google.com/open?id=138puSOxDk-Zu1sn5B_4LVJ__q6hwz41k Data Preprocessing. Create dummies for day of week, carrier, departure airport, and arrival airport. This will give you 17 dummies. Bin

use this document/link https://drive.google.com/open?id=138puSOxDk-Zu1sn5B_4LVJ__q6hwz41k

Data Preprocessing. Create dummies for day of week, carrier, departure airport, and arrival airport. This will give you 17 dummies. Bin the scheduled departure time into eight bins (in XLMiner use Transform Bin Continuous Data and select equal width). After binning CRS DEP TIME into the 8 bins, this new variable should be broken down into dummies (because the effect will not be linear, due to the morning and afternoon rush hours). This will avoid treating the departure time as a continuous predictor, since it is reasonable that delays are related to rush-hour times. Partition the data into training and validation sets. (show all steps)

Fit a classi?cation tree to the ?ight delay variable using all the relevant predictors. Do not include DEP_TIME (actual departure time) in the model because it is unknown at the time of prediction (unless we are generating our predictions of delays after the plane takes off, which is unlikely). In the third step of the Classi?cation Tree menu, choose Maximum # levels to be displayed = 6. Use the best-pruned tree, setting the minimum number of observations in the ?nal nodes to 1. Express the resulting tree as a set of rules.

If you needed to ?y between DCA and EWR on a Monday at 7 AM, would you be able to use this tree? What other information would you need? Is it available in practice? What information is redundant?

Fit another tree, this time excluding the Weather predictor. (Why?) Select the option of seeing both the full tree and the best-pruned tree. You will ?nd that the best-pruned tree contains a single terminal node.

How is this tree used for classi?cation? (What is the rule for classifying?)

To what is this rule equivalent?

Examine the full tree. What are the top three predictors according to this tree?

Why, technically, does the pruned tree result in a tree with a single node?

What is the disadvantage of using the top levels of the full tree as opposed to the best-pruned tree?

Compare this general result to that from logistic regression in the example in Chapter 10. What are possible reasons for the classi?cation trees failure to ?nd a good predictive model?

Step by Step Solution

There are 3 Steps involved in it

1 Expert Approved Answer

Step: 1 Unlock blur-text-image

Question Has Been Solved by an Expert!

Get step-by-step solutions from verified subject matter experts

Step: 2 Unlock

Step: 3 Unlock

Students Have Also Explored These Related Databases Questions!

Data set: https://file.io/UwfKfn This is to be done in XLMiner --- Please ask for data set Create dummies for day of week, carrier, departure airport, and arrival airport. This will give you 17...

Managing Scope Changes Case Study Scope changes on a project can occur regardless of how well the project is planned or executed. Scope changes can be the result of something that was omitted during...

Predicting Delayed Flights: The file FlightDelays.xls contains information on all commercial flights departing the Washington DC area and arriving at New York during January 2004. For each flight...

Please help in this assaiment, view AS Read aloud | Draw v Highlig In this assignment, you will design a visualization for a dataset. You are free to use any graphics or charting tool you please...

Scenario You and your group work for a multi-national consulting firm. Your team is a group of capital budget consultants and are seeking new engagements. You intend on responding to a Request for...

FACT SHEET FOR YOUR CONFERENCE ASSOCIATION NAME: Statistics Notes Name of Conference: List the name you are going to give your conference, this will be used in all of your marketing (HTM 2025 Annual...

Predicting Delayed Flights. The file FlightDelays.csv contains information on all commercial flights departing the Washington, DC area and arriving at New York during January 2004. For each flight,...

can you site what page numbers you found the answers on please and thank you CASE 12 EMIRATES AIRLINE IN 2017" With three decades. Emirates Airline went from a small introduction of Boeing 777 long...

Document1 - Microsoft Word Home Page layout Retorences Maling Review View Times New Roman - 12 AA FF 2. 41 Intert ou Copy Paste Format Painter Clipboard AaBbceDd AaBbcend AaB AaBbcc AaBbc Aabbccd....

HI GOOD AFTERNOON, IS IT POSSIBLE THAT I MAY GET SOME ASSISTANCE WITH THIS ASSIGNMENT. SEE ATTACHMENT INSTRUCTIONS ARE AS FOLLOWS: ONLY THE RELEVANT HISTORY IS IMPORTANT, IN TERMS OF A PROBLEM...

Refer to Mixon Companys balance sheets in Exercise 13. Express the balance sheets in commonsize percents. Round to the nearest one-tenth of a percent. In Exercise13 2006 2005 2004 Cash Accounts...

Draw the graph of the function z = f(x, y) = x + y in R. Hint: For any contant r > 0, set z = r which gives x + y = r. What shape does this equation correspond to? What happens when r become...

Accruals, accounts payable, and notes payable are listed on the balance sheet as A - accumulated liabilities B | | current liabilities C non - current liabilities D | | accrued liabilities Select an...

Ming Chen started a business and had the following transactions in June. Owner invested $66,000 cash in the company along with $15,000 of equipment in exchange for its common stock. The company paid...

1. To identify the programs strengths and weaknesses. This includes determining if the program is meeting the learning objectives, if the quality of the learning environment is satisfactory, and if...

4. Cost justification for training is based on numerical indicators. (Here the company has a strong orientation toward evaluation.)

2. To assess whether the content, organization, and administration of the program including the schedule, accommodations, trainers, and materialscontribute to learning and the use of training content...