{"id":272,"date":"2018-10-31T04:34:42","date_gmt":"2018-10-31T04:34:42","guid":{"rendered":"http:\/\/mgmtp15.epgpbooks.inflibnet.ac.in\/?post_type=chapter&#038;p=272"},"modified":"2018-10-31T04:59:27","modified_gmt":"2018-10-31T04:59:27","slug":"test-of-goodness-of-fit-and-independence-chi-square-test-as-a-test-of-independence","status":"publish","type":"chapter","link":"https:\/\/ebooks.inflibnet.ac.in\/mgmtp15\/chapter\/test-of-goodness-of-fit-and-independence-chi-square-test-as-a-test-of-independence\/","title":{"rendered":"Test of Goodness of Fit and Independence: Chi-Square-test-as a test of independence"},"content":{"raw":"<div>\r\n\r\n&nbsp;\r\n\r\n<strong>Test of Goodness of Fit and Independence: Chi-Square-test-as a test of independence<\/strong>\r\n\r\n&nbsp;\r\n\r\n<strong>Learning objective<\/strong>\r\n<ul>\r\n \t<li>After reading this module the students will be able to<\/li>\r\n \t<li>Understand the concept of non-parametric tests Apply Chi-square as Test for independence.<\/li>\r\n \t<li>Gain knowledge about the procedure of conducting Chi-square test.<\/li>\r\n<\/ul>\r\n&nbsp;\r\n\r\n<strong>Introduction<\/strong>\r\n\r\n&nbsp;\r\n<p style=\"text-align: justify\">The given set of data can be analyzed with the help of various tools available on the basis of the following parameters;<\/p>\r\n&nbsp;\r\n\r\nSize of the Sample\r\n\r\nSize of the Population\r\n\r\nScale used for measurement of data And dependency of measurement\r\n\r\n&nbsp;\r\n<p style=\"text-align: justify\">The tests may be classified in to two category mainly Parametric and Non-Parametric. Three test i.e. t, z and F are used to estimate and test the population parameters and prerequisite of application of these test are Interval and ratio Scale to be used<\/p>\r\n&nbsp;\r\n\r\nHypothesis testing for specific parameters\r\n\r\n&nbsp;\r\n\r\nAssumption of normality and Standard deviation is known or not should be clear\r\n\r\n&nbsp;\r\n<p style=\"text-align: justify\">The absence of these conditions leads to the application of Non-Parametric Tests or distribution free tests. These tests are applied in following conditions;<\/p>\r\n&nbsp;\r\n<p style=\"text-align: justify\">Do not require specific population distribution and data can be nominal or ordinal Does not takes in to consideration of population parameters<\/p>\r\n&nbsp;\r\n\r\nDoes not require normally distributed population\r\n\r\n&nbsp;\r\n<p style=\"text-align: justify\">These test are very easy to apply and can use nominal or ordinal data as well for calculation. These tests provide broad based conclusion with approximate solution and does not necessarily require normally distributed population. The \u03c7-square test is one of the non-parametric tests used to test hypothesis.<\/p>\r\n&nbsp;\r\n\r\n\u03c72\u00a0 <strong>test for Independence<\/strong>\r\n<p style=\"text-align: justify\">The test is applicable in the situation when there are two categorical variables from a single population. Its\u2019 purpose is to find out if there is a significant association between the two variables or not. For example, in an election survey, voters may be categorized on the basis of gender (i.e. male or female) and on the basis of party inclination ( i.e Democrat, Republican, or Independent). Chi-square test for independence is conducted to determine whether gender is related to party inclination or not.<\/p>\r\n\r\n<\/div>\r\n&nbsp;\r\n<div>\r\n\r\nThis test is suitable under the following conditions:\r\n\r\n&nbsp;\r\n\r\n\u00a7\u00a0 The sample is selected through simple random sampling.\r\n\r\n&nbsp;\r\n\r\n\u00a7\u00a0 The variables are of categorical nature.\r\n\r\n&nbsp;\r\n\r\n\u00a7\u00a0 The expected frequency count for each cell of the contingency table should not be less than 5.\r\n\r\n&nbsp;\r\n\r\nProcedure\r\n\r\n&nbsp;\r\n<p style=\"text-align: justify\">The procedure to test the association between two independent variables where the sample data is presented in the form of contingency table with n rows and m columns is summarized as \u2013<\/p>\r\n&nbsp;\r\n\r\n1. State the null and alternative hypotheses\r\n\r\n&nbsp;\r\n\r\nH0: No relationship or association exists between variables.\r\n\r\n&nbsp;\r\n\r\nHa:\u00a0 A relationship or association exists between variables i.e., they are related.\r\n\r\n&nbsp;\r\n<p style=\"text-align: justify\">2.\u00a0 Select a random sample and record the observed frequencies (O) in each cell of the contingency table and calculate the row, column and grand total.<\/p>\r\n&nbsp;\r\n\r\n3<strong>.<\/strong> Calculate the expected frequencies (E) for each cell:\r\n\r\n&nbsp;\r\n\r\nE= Row total*column total\/Grand total\r\n\r\n&nbsp;\r\n\r\n4.\u00a0 Compute the value of test statistic, \u03c72= \u03a3 [(O - E)2 \/ E ],\r\n\r\n&nbsp;\r\n\r\nwhere O is the observed frequency count and E is the expected frequency count.\r\n\r\n&nbsp;\r\n\r\n5. Calculate the degrees of freedom\r\n\r\n&nbsp;\r\n\r\ndf = (c - 1) * (r - 1)\r\n\r\n&nbsp;\r\n<p style=\"text-align: justify\">where c is the number of levels for one categorical variable, and r is the number of levels for the other categorical variable.<\/p>\r\n&nbsp;\r\n\r\n6.Use the level of significance \u03b1 and df to find the table value of \u03c72 at\u00a0 \u03b1.\r\n\r\n&nbsp;\r\n<p style=\"text-align: justify\">7.\u00a0 Compare the calculated and table value .If calculated value of chi-square is less than the table value, accept the null hypotheis otherwise reject it<\/p>\r\n\r\n<\/div>\r\n<div>\r\n\r\n&nbsp;\r\n\r\n<strong>Example<\/strong>\r\n\r\n&nbsp;\r\n<p style=\"text-align: justify\">A simple random sample of 1000 prospective voters was taken. They were categorized on the basis of gender ( namely M\/F) and on the basis of party liking (Republican, Democrat, or Independent). The contingency table given below shows the result<\/p>\r\n&nbsp;\r\n\r\nDo the M's party liking differ significantly from the F's preferences? Use a 0.05 level of significance.\r\n\r\n&nbsp;\r\n\r\n<strong>Solution<\/strong>\r\n\r\n&nbsp;\r\n<p style=\"text-align: justify\">As discussed above following procedure is followed,<\/p>\r\n&nbsp;\r\n<p style=\"text-align: justify\">The first step is to state the null hypothesis and an alternative hypothesis. H0: Gender and party likings are independent.<\/p>\r\n&nbsp;\r\n\r\nHa: Gender and party likings are not independent.\r\n\r\n&nbsp;\r\n\r\nFor this analysis, the significance level is 0.05, chi-square test for independence will be used.\r\n\r\n&nbsp;\r\n\r\nDegrees of freedom, the expected frequency counts, and the chi-square test statistic are calculated. df = (c-\r\n\r\n1) * (r - 1) = (2 - 1) * (3 - 1) = 2\r\n\r\n&nbsp;\r\n\r\nE1,1 = (800 * 900) \/ 2000 = 720000\/2000 = 360 E1,2 = (800 * 900) \/ 2000 = 360\r\n\r\n&nbsp;\r\n\r\nE1,3 = (800 * 200) \/ 2000 = 80 E2,1 = (1200 * 900) \/ 2000 = 540\r\n\r\n&nbsp;\r\n\r\n<span style=\"text-align: initial;font-size: 1em\">E2,2 = (1200 *900) \/ 2000 = 540<\/span>\r\n\r\n&nbsp;\r\n\r\n<span style=\"text-align: initial;font-size: 1em\">E2,3 = (1200 * 200) \/ 2000 = 120<\/span>\r\n\r\n&nbsp;\r\n\r\n<em style=\"text-align: initial;font-size: 1em\">x<\/em><span style=\"text-align: initial;font-size: 1em\">2 = \u03a3 [ (On ,m \u2013 En, m)2 \/ En ,m ]<\/span>\r\n\r\n&nbsp;\r\n\r\n<em style=\"text-align: initial;font-size: 1em\">x<\/em><span style=\"text-align: initial;font-size: 1em\">2 = (400 - 360)2\/360 + (300 - 360)2\/360 + (100 - 80)2\/80<\/span>\r\n\r\n&nbsp;\r\n\r\n<span style=\"text-align: initial;font-size: 1em\">+\u00a0 (500 - 540)2\/540 + (600 - 540)2\/540 + (100 - 120)2\/120 <\/span><em style=\"text-align: initial;font-size: 1em\">x<\/em><span style=\"text-align: initial;font-size: 1em\">2 = 4.44 + 10.00 + 5.0 + 2.96 + 6.66 + 3.34 = 32.4<\/span>\r\n\r\n<\/div>\r\n<div>\r\n\r\n&nbsp;\r\n<p style=\"text-align: justify\">This calculated value of chi-square statistic having 2 degrees of freedom is more than the table value (refer Chi-square table ,hence null hypothesis is not accepted. Thus, we conclude that there is a relationship between gender and voting preference.<\/p>\r\n&nbsp;\r\n\r\n<strong>Self-Check Questions:<\/strong>\r\n\r\n&nbsp;\r\n<p style=\"text-align: justify\"><strong>Question 1<\/strong>: Two hundred randomly selected adults were asked whether TV shows as a whole are primarily entertaining, educational or boring. The respondents were categorized by gender. Their responses are given in the following table-<\/p>\r\n\r\n<table class=\"aligncenter\" style=\"width: 60%\" border=\"1\">\r\n<tbody>\r\n<tr>\r\n<td style=\"width: 64.0625px\"><strong>Gender<\/strong><\/td>\r\n<td style=\"width: 102.063px\"><strong>Opinion Entertaining<\/strong><\/td>\r\n<td style=\"width: 96.0625px\"><strong>Educational<\/strong><\/td>\r\n<td style=\"width: 109.063px\"><strong>Waste of time<\/strong><\/td>\r\n<td style=\"width: 50.0625px\"><strong>Total<\/strong><\/td>\r\n<\/tr>\r\n<tr>\r\n<td style=\"width: 64.0625px\"><strong>Female<\/strong><\/td>\r\n<td style=\"width: 102.063px\">52<\/td>\r\n<td style=\"width: 96.0625px\">28<\/td>\r\n<td style=\"width: 109.063px\">30<\/td>\r\n<td style=\"width: 50.0625px\">110<\/td>\r\n<\/tr>\r\n<tr>\r\n<td style=\"width: 64.0625px\"><strong>Male<\/strong><\/td>\r\n<td style=\"width: 102.063px\">28<\/td>\r\n<td style=\"width: 96.0625px\">12<\/td>\r\n<td style=\"width: 109.063px\">50<\/td>\r\n<td style=\"width: 50.0625px\">90<\/td>\r\n<\/tr>\r\n<tr>\r\n<td style=\"width: 64.0625px\"><strong>Total<\/strong><\/td>\r\n<td style=\"width: 102.063px\">80<\/td>\r\n<td style=\"width: 96.0625px\">40<\/td>\r\n<td style=\"width: 109.063px\">80<\/td>\r\n<td style=\"width: 50.0625px\">200<\/td>\r\n<\/tr>\r\n<\/tbody>\r\n<\/table>\r\n&nbsp;\r\n<p style=\"text-align: justify\">Is this evidence convincing that there is a relationship between gender and opinion in the population of interest?<\/p>\r\n&nbsp;\r\n<p style=\"text-align: justify\"><strong>Solution \u2013 <\/strong>Let us take the null hypothesis that the opinion of adults is independent of adults is independent of gender.<\/p>\r\n&nbsp;\r\n<p style=\"text-align: justify\">Since, contingency table is of size 2x3,the degrees of freedom would be (2-1)(3-1) = 2.This implies that we need to calculate only to calculate only two expected frequencies and the other four can automatically be determined as shown below:<\/p>\r\n\r\n<\/div>\r\n<div>\r\n\r\n&nbsp;\r\n\r\nE11=Row 1 total x Column 1 total\r\n\r\n&nbsp;\r\n\r\n<span style=\"text-align: initial;font-size: 1em\">E13=110-(44+22)=44<\/span>\r\n\r\n&nbsp;\r\n\r\n<span style=\"text-align: initial;font-size: 1em\">E21=80-E11=40-22==18<\/span>\r\n\r\n&nbsp;\r\n\r\n<span style=\"text-align: initial;font-size: 1em\">E22=40-E12=40-22=18<\/span>\r\n\r\n&nbsp;\r\n\r\n<span style=\"text-align: initial;font-size: 1em\">E23=80-E13=80-44=36<\/span>\r\n\r\n&nbsp;\r\n\r\n<span style=\"text-align: initial;font-size: 1em\">The contingency table of expected frequencies is as follows:<\/span>\r\n\r\n<\/div>\r\n<div>\r\n<table class=\"aligncenter\" style=\"width: 60%\" border=\"1\">\r\n<tbody>\r\n<tr>\r\n<td><strong>Gender<\/strong><\/td>\r\n<td><strong>Entertaining<\/strong><\/td>\r\n<td><strong>Opinion Educational<\/strong><\/td>\r\n<td><strong>Waste of time<\/strong><\/td>\r\n<td><strong>Total<\/strong><\/td>\r\n<\/tr>\r\n<tr>\r\n<td><strong>Female<\/strong><\/td>\r\n<td>44<\/td>\r\n<td>22<\/td>\r\n<td>44<\/td>\r\n<td>110<\/td>\r\n<\/tr>\r\n<tr>\r\n<td><strong>Male<\/strong><\/td>\r\n<td>36<\/td>\r\n<td>18<\/td>\r\n<td>36<\/td>\r\n<td>90<\/td>\r\n<\/tr>\r\n<tr>\r\n<td><strong>Total<\/strong><\/td>\r\n<td>80<\/td>\r\n<td>40<\/td>\r\n<td>80<\/td>\r\n<td>200<\/td>\r\n<\/tr>\r\n<\/tbody>\r\n<\/table>\r\n&nbsp;\r\n\r\nArranging the observed and expected frequencies as follows to calculate the value of <em>x<\/em>2-test statistic:\r\n<table class=\"aligncenter\" style=\"width: 60%\" border=\"1\">\r\n<tbody>\r\n<tr>\r\n<td><strong>Observed(O)<\/strong><\/td>\r\n<td><strong>Expected(E)<\/strong><\/td>\r\n<td><strong>O-E<\/strong><\/td>\r\n<td><strong>(O-E)2<\/strong><\/td>\r\n<td><strong>(O-E)2\/E<\/strong><\/td>\r\n<\/tr>\r\n<tr>\r\n<td>52<\/td>\r\n<td>44<\/td>\r\n<td>8<\/td>\r\n<td>64<\/td>\r\n<td>1.454<\/td>\r\n<\/tr>\r\n<tr>\r\n<td>28<\/td>\r\n<td>22<\/td>\r\n<td>6<\/td>\r\n<td>36<\/td>\r\n<td>1.636<\/td>\r\n<\/tr>\r\n<tr>\r\n<td>30<\/td>\r\n<td>44<\/td>\r\n<td>14<\/td>\r\n<td>196<\/td>\r\n<td>4.455<\/td>\r\n<\/tr>\r\n<tr>\r\n<td>28<\/td>\r\n<td>36<\/td>\r\n<td>-8<\/td>\r\n<td>64<\/td>\r\n<td>1.777<\/td>\r\n<\/tr>\r\n<tr>\r\n<td>12<\/td>\r\n<td>18<\/td>\r\n<td>-6<\/td>\r\n<td>36<\/td>\r\n<td>2<\/td>\r\n<\/tr>\r\n<tr>\r\n<td>50<\/td>\r\n<td>36<\/td>\r\n<td>14<\/td>\r\n<td>196<\/td>\r\n<td>5.444<\/td>\r\n<\/tr>\r\n<tr>\r\n<td><\/td>\r\n<td><\/td>\r\n<td><\/td>\r\n<td><\/td>\r\n<td>16.766<\/td>\r\n<\/tr>\r\n<\/tbody>\r\n<\/table>\r\n<\/div>\r\n<div>\r\n\r\n&nbsp;\r\n<p style=\"text-align: justify\">Since, calculated value of <em>x<\/em>2=16.766 is more than its critical value, <em>x<\/em>2=5.99 at \u03b1=0.05 and <em>df<\/em> = 2 ,the null hypothesis is rejected. Hence, we conclude that the opinion of adults is not independent of gender.<\/p>\r\n&nbsp;\r\n\r\n<strong style=\"text-align: initial;font-size: 1em\">Question 2: <\/strong><span style=\"text-align: initial;font-size: 1em\">A sample analysis of examination results of 500 students was made. It was found that 220 students had failed ,170 had secured a third division 90 were placed in second division and 20 got a first division. Are these figures commensurate with the general examination result which is the ratio of 4:3:2:1 for the various categories respectively?<\/span>\r\n\r\n<\/div>\r\n<div><\/div>\r\n<div>\r\n\r\n&nbsp;\r\n<p style=\"text-align: justify\"><strong>Solution<\/strong>- Let us take the null hypothesis that the observed results are commensurate with the general examination result which is the ratio 4:3:2:1.<\/p>\r\n&nbsp;\r\n\r\nThe expected number of students who have failed, obtained a third division second division and first division, respectively, are\r\n\r\n&nbsp;\r\n\r\nE1=500*4\/10=200, E2=500*3\/10=150;E3=500*2\/10=100 AND E4=500*1\/10=50\r\n\r\n&nbsp;\r\n\r\nThe contingency table of expected and observed frequencies is as follows:\r\n<table class=\"aligncenter\" style=\"width: 60%\" border=\"1\">\r\n<tbody>\r\n<tr>\r\n<td><strong>Category<\/strong><\/td>\r\n<td><strong>O<\/strong><\/td>\r\n<td><strong>E<\/strong><\/td>\r\n<td><strong>(O-E)<\/strong><strong>2<\/strong><\/td>\r\n<td><strong>\u03a7<\/strong><strong>2<\/strong><strong>=(O-E)<\/strong><strong>2<\/strong><strong>\/E<\/strong><\/td>\r\n<\/tr>\r\n<tr>\r\n<td>Failed<\/td>\r\n<td>220<\/td>\r\n<td>200<\/td>\r\n<td>400<\/td>\r\n<td>2<\/td>\r\n<\/tr>\r\n<tr>\r\n<td>3rd division<\/td>\r\n<td>170<\/td>\r\n<td>150<\/td>\r\n<td>400<\/td>\r\n<td>2.667<\/td>\r\n<\/tr>\r\n<tr>\r\n<td>2nd division<\/td>\r\n<td>90<\/td>\r\n<td>100<\/td>\r\n<td>100<\/td>\r\n<td>1<\/td>\r\n<\/tr>\r\n<tr>\r\n<td>1st division<\/td>\r\n<td>20<\/td>\r\n<td>50<\/td>\r\n<td>900<\/td>\r\n<td>18<\/td>\r\n<\/tr>\r\n<tr>\r\n<td><\/td>\r\n<td><\/td>\r\n<td><\/td>\r\n<td><\/td>\r\n<td>23.667<\/td>\r\n<\/tr>\r\n<\/tbody>\r\n<\/table>\r\n&nbsp;\r\n<p style=\"text-align: justify\">Since calculated value of <em>x<\/em>2 = 23.667 is more than its table value , <em>x<\/em>2=7.81 at \u03b1 = 0.05 level of significance and df= n \u2013 1 = 4 -1 =3 the hypothesis is rejected.<\/p>\r\n&nbsp;\r\n<p style=\"text-align: justify\"><strong>Question 3<\/strong>: Based on information on 1000 randomly selected fields about the tenancy status of the cultivation of these fields and use of fertilizers ,collected in an AGRO ECONOMY survey, the following classification was noted:<\/p>\r\n\r\n<table class=\"aligncenter\" style=\"width: 60%\" border=\"1\">\r\n<tbody>\r\n<tr>\r\n<td><\/td>\r\n<td><strong>Owned<\/strong><\/td>\r\n<td><strong>Rented<\/strong><\/td>\r\n<td><strong>Total<\/strong><\/td>\r\n<\/tr>\r\n<tr>\r\n<td><strong>Using fertilizers<\/strong><\/td>\r\n<td>416<\/td>\r\n<td>184<\/td>\r\n<td>600<\/td>\r\n<\/tr>\r\n<tr>\r\n<td><strong>Not using fertilizers<\/strong><\/td>\r\n<td>64<\/td>\r\n<td>336<\/td>\r\n<td>400<\/td>\r\n<\/tr>\r\n<tr>\r\n<td><strong>Total<\/strong><\/td>\r\n<td>480<\/td>\r\n<td>5220<\/td>\r\n<td>1000<\/td>\r\n<\/tr>\r\n<\/tbody>\r\n<\/table>\r\n<\/div>\r\n<div><span style=\"font-size: 1em;text-align: initial\">Would you conclude that owner cultivators are more towards the use of fertilizers at 5%level of significance? Carry out a chi-square test as per testing procedure.<\/span><\/div>\r\n&nbsp;\r\n<p style=\"text-align: justify\"><strong>Solution<\/strong>: Let us take the hypothesis that ownership of fields and the use of fertilizers are independent attributes. Since, contingency table is of size 2*2 the degree of freedom would be (2-1)(2-1)=1. This implies that we need to calculate only one expected frequency and others can be automatically determined as follows:<\/p>\r\n&nbsp;\r\n\r\nE11=600*480\/1000=288\r\n\r\n&nbsp;\r\n\r\nE12 =600-288=312\r\n\r\n&nbsp;\r\n\r\nE21=480-288=192\r\n\r\n&nbsp;\r\n\r\nE22=208\r\n\r\n&nbsp;\r\n\r\nThe contingency table of expected frequencies is as follows:\r\n<table class=\"aligncenter\" style=\"width: 60%\" border=\"1\">\r\n<tbody>\r\n<tr>\r\n<td><strong>Observed<\/strong><\/td>\r\n<td><strong>Expected<\/strong><\/td>\r\n<td><strong>(O-E)<\/strong><strong>2<\/strong><\/td>\r\n<td><strong><em>x<\/em><\/strong><strong>2<\/strong><strong>=(O-E)<\/strong><strong>2<\/strong><strong>\/E<\/strong><\/td>\r\n<\/tr>\r\n<tr>\r\n<td>416<\/td>\r\n<td>288<\/td>\r\n<td>16,384<\/td>\r\n<td>56.889<\/td>\r\n<\/tr>\r\n<tr>\r\n<td>64<\/td>\r\n<td>192<\/td>\r\n<td>16,384<\/td>\r\n<td>85.333<\/td>\r\n<\/tr>\r\n<tr>\r\n<td>184<\/td>\r\n<td>312<\/td>\r\n<td>16,384<\/td>\r\n<td>52.513<\/td>\r\n<\/tr>\r\n<tr>\r\n<td>336<\/td>\r\n<td>208<\/td>\r\n<td>16384<\/td>\r\n<td>78.769<\/td>\r\n<\/tr>\r\n<tr>\r\n<td><\/td>\r\n<td><\/td>\r\n<td><\/td>\r\n<td>273.534<\/td>\r\n<\/tr>\r\n<\/tbody>\r\n<\/table>\r\n&nbsp;\r\n\r\nThe calculated value of <em>x<\/em>2=273.534 at \u03b1=0.05 level of significance and df= (n-1) (r-1) =(2-1) (2-1) = 1 is much more than its table value,\u03c72=3.84. The null hypothesis H0 is rejected. Hence, it can be conducted\r\n<p style=\"text-align: justify\">that owners\u2019 cultivators are more inclined towards the use of fertilizers.<\/p>\r\n&nbsp;\r\n\r\n<strong>Summary<\/strong>\r\n\r\n&nbsp;\r\n<p style=\"text-align: justify\">The tests may be classified in to two category mainly Parametric and Non-Parametric. Three test i.e. t test, z test and f test are used to estimate and test the population parameters and prerequisite of application of these test are-interval and ratio Scale to be used, hypothesis testing for specific parameters, assumption of normality and Standard deviation is known or not should be clear.<\/p>\r\n&nbsp;\r\n<p style=\"text-align: justify\">The absence of these conditions leads to the application of Non-Parametric Tests or distribution free tests. These test are very easy to apply and can use nominal or ordinal data as well for calculation. These tests provide broad based conclusion with approximate solution and does not necessarily require normally distributed population. The Chi-Square test is one of the non-parametric tests used to test hypothesis. Chi-Square Test for Independence test is applied when you have two categorical variables from a single population. It is used to determine whether there is a significant association between the two variables.<\/p>\r\n&nbsp;\r\n<p style=\"text-align: center\"><strong>Learn More:<\/strong><\/p>\r\n\r\n<ol>\r\n \t<li>Sharma, J K (2014), Business Statistics, S Chand &amp; Company, N Delhi.<\/li>\r\n \t<li>Bajpai, N (2010) Business Statistics, Pearson, N Delhi.<\/li>\r\n \t<li>Trevor Hastie, Robert Tibshirani, Jerome Friedman (2009), The Elements of Statistical Learning: Data Mining, Inference, and Prediction, 2nd Edition, Springer.<\/li>\r\n \t<li>Darrell Huff (2010), How to Lie with Statistics,\u00a0 W. W. Norton, California.<\/li>\r\n \t<li>K.R. Gupta (2012), Practical Statistics, Atlantic Publishers &amp; Distributors (P) Ltd., N. Delhi.<\/li>\r\n<\/ol>","rendered":"<div>\n<p>&nbsp;<\/p>\n<p><strong>Test of Goodness of Fit and Independence: Chi-Square-test-as a test of independence<\/strong><\/p>\n<p>&nbsp;<\/p>\n<p><strong>Learning objective<\/strong><\/p>\n<ul>\n<li>After reading this module the students will be able to<\/li>\n<li>Understand the concept of non-parametric tests Apply Chi-square as Test for independence.<\/li>\n<li>Gain knowledge about the procedure of conducting Chi-square test.<\/li>\n<\/ul>\n<p>&nbsp;<\/p>\n<p><strong>Introduction<\/strong><\/p>\n<p>&nbsp;<\/p>\n<p style=\"text-align: justify\">The given set of data can be analyzed with the help of various tools available on the basis of the following parameters;<\/p>\n<p>&nbsp;<\/p>\n<p>Size of the Sample<\/p>\n<p>Size of the Population<\/p>\n<p>Scale used for measurement of data And dependency of measurement<\/p>\n<p>&nbsp;<\/p>\n<p style=\"text-align: justify\">The tests may be classified in to two category mainly Parametric and Non-Parametric. Three test i.e. t, z and F are used to estimate and test the population parameters and prerequisite of application of these test are Interval and ratio Scale to be used<\/p>\n<p>&nbsp;<\/p>\n<p>Hypothesis testing for specific parameters<\/p>\n<p>&nbsp;<\/p>\n<p>Assumption of normality and Standard deviation is known or not should be clear<\/p>\n<p>&nbsp;<\/p>\n<p style=\"text-align: justify\">The absence of these conditions leads to the application of Non-Parametric Tests or distribution free tests. These tests are applied in following conditions;<\/p>\n<p>&nbsp;<\/p>\n<p style=\"text-align: justify\">Do not require specific population distribution and data can be nominal or ordinal Does not takes in to consideration of population parameters<\/p>\n<p>&nbsp;<\/p>\n<p>Does not require normally distributed population<\/p>\n<p>&nbsp;<\/p>\n<p style=\"text-align: justify\">These test are very easy to apply and can use nominal or ordinal data as well for calculation. These tests provide broad based conclusion with approximate solution and does not necessarily require normally distributed population. The \u03c7-square test is one of the non-parametric tests used to test hypothesis.<\/p>\n<p>&nbsp;<\/p>\n<p>\u03c72\u00a0 <strong>test for Independence<\/strong><\/p>\n<p style=\"text-align: justify\">The test is applicable in the situation when there are two categorical variables from a single population. Its\u2019 purpose is to find out if there is a significant association between the two variables or not. For example, in an election survey, voters may be categorized on the basis of gender (i.e. male or female) and on the basis of party inclination ( i.e Democrat, Republican, or Independent). Chi-square test for independence is conducted to determine whether gender is related to party inclination or not.<\/p>\n<\/div>\n<p>&nbsp;<\/p>\n<div>\n<p>This test is suitable under the following conditions:<\/p>\n<p>&nbsp;<\/p>\n<p>\u00a7\u00a0 The sample is selected through simple random sampling.<\/p>\n<p>&nbsp;<\/p>\n<p>\u00a7\u00a0 The variables are of categorical nature.<\/p>\n<p>&nbsp;<\/p>\n<p>\u00a7\u00a0 The expected frequency count for each cell of the contingency table should not be less than 5.<\/p>\n<p>&nbsp;<\/p>\n<p>Procedure<\/p>\n<p>&nbsp;<\/p>\n<p style=\"text-align: justify\">The procedure to test the association between two independent variables where the sample data is presented in the form of contingency table with n rows and m columns is summarized as \u2013<\/p>\n<p>&nbsp;<\/p>\n<p>1. State the null and alternative hypotheses<\/p>\n<p>&nbsp;<\/p>\n<p>H0: No relationship or association exists between variables.<\/p>\n<p>&nbsp;<\/p>\n<p>Ha:\u00a0 A relationship or association exists between variables i.e., they are related.<\/p>\n<p>&nbsp;<\/p>\n<p style=\"text-align: justify\">2.\u00a0 Select a random sample and record the observed frequencies (O) in each cell of the contingency table and calculate the row, column and grand total.<\/p>\n<p>&nbsp;<\/p>\n<p>3<strong>.<\/strong> Calculate the expected frequencies (E) for each cell:<\/p>\n<p>&nbsp;<\/p>\n<p>E= Row total*column total\/Grand total<\/p>\n<p>&nbsp;<\/p>\n<p>4.\u00a0 Compute the value of test statistic, \u03c72= \u03a3 [(O &#8211; E)2 \/ E ],<\/p>\n<p>&nbsp;<\/p>\n<p>where O is the observed frequency count and E is the expected frequency count.<\/p>\n<p>&nbsp;<\/p>\n<p>5. Calculate the degrees of freedom<\/p>\n<p>&nbsp;<\/p>\n<p>df = (c &#8211; 1) * (r &#8211; 1)<\/p>\n<p>&nbsp;<\/p>\n<p style=\"text-align: justify\">where c is the number of levels for one categorical variable, and r is the number of levels for the other categorical variable.<\/p>\n<p>&nbsp;<\/p>\n<p>6.Use the level of significance \u03b1 and df to find the table value of \u03c72 at\u00a0 \u03b1.<\/p>\n<p>&nbsp;<\/p>\n<p style=\"text-align: justify\">7.\u00a0 Compare the calculated and table value .If calculated value of chi-square is less than the table value, accept the null hypotheis otherwise reject it<\/p>\n<\/div>\n<div>\n<p>&nbsp;<\/p>\n<p><strong>Example<\/strong><\/p>\n<p>&nbsp;<\/p>\n<p style=\"text-align: justify\">A simple random sample of 1000 prospective voters was taken. They were categorized on the basis of gender ( namely M\/F) and on the basis of party liking (Republican, Democrat, or Independent). The contingency table given below shows the result<\/p>\n<p>&nbsp;<\/p>\n<p>Do the M&#8217;s party liking differ significantly from the F&#8217;s preferences? Use a 0.05 level of significance.<\/p>\n<p>&nbsp;<\/p>\n<p><strong>Solution<\/strong><\/p>\n<p>&nbsp;<\/p>\n<p style=\"text-align: justify\">As discussed above following procedure is followed,<\/p>\n<p>&nbsp;<\/p>\n<p style=\"text-align: justify\">The first step is to state the null hypothesis and an alternative hypothesis. H0: Gender and party likings are independent.<\/p>\n<p>&nbsp;<\/p>\n<p>Ha: Gender and party likings are not independent.<\/p>\n<p>&nbsp;<\/p>\n<p>For this analysis, the significance level is 0.05, chi-square test for independence will be used.<\/p>\n<p>&nbsp;<\/p>\n<p>Degrees of freedom, the expected frequency counts, and the chi-square test statistic are calculated. df = (c-<\/p>\n<p>1) * (r &#8211; 1) = (2 &#8211; 1) * (3 &#8211; 1) = 2<\/p>\n<p>&nbsp;<\/p>\n<p>E1,1 = (800 * 900) \/ 2000 = 720000\/2000 = 360 E1,2 = (800 * 900) \/ 2000 = 360<\/p>\n<p>&nbsp;<\/p>\n<p>E1,3 = (800 * 200) \/ 2000 = 80 E2,1 = (1200 * 900) \/ 2000 = 540<\/p>\n<p>&nbsp;<\/p>\n<p><span style=\"text-align: initial;font-size: 1em\">E2,2 = (1200 *900) \/ 2000 = 540<\/span><\/p>\n<p>&nbsp;<\/p>\n<p><span style=\"text-align: initial;font-size: 1em\">E2,3 = (1200 * 200) \/ 2000 = 120<\/span><\/p>\n<p>&nbsp;<\/p>\n<p><em style=\"text-align: initial;font-size: 1em\">x<\/em><span style=\"text-align: initial;font-size: 1em\">2 = \u03a3 [ (On ,m \u2013 En, m)2 \/ En ,m ]<\/span><\/p>\n<p>&nbsp;<\/p>\n<p><em style=\"text-align: initial;font-size: 1em\">x<\/em><span style=\"text-align: initial;font-size: 1em\">2 = (400 &#8211; 360)2\/360 + (300 &#8211; 360)2\/360 + (100 &#8211; 80)2\/80<\/span><\/p>\n<p>&nbsp;<\/p>\n<p><span style=\"text-align: initial;font-size: 1em\">+\u00a0 (500 &#8211; 540)2\/540 + (600 &#8211; 540)2\/540 + (100 &#8211; 120)2\/120 <\/span><em style=\"text-align: initial;font-size: 1em\">x<\/em><span style=\"text-align: initial;font-size: 1em\">2 = 4.44 + 10.00 + 5.0 + 2.96 + 6.66 + 3.34 = 32.4<\/span><\/p>\n<\/div>\n<div>\n<p>&nbsp;<\/p>\n<p style=\"text-align: justify\">This calculated value of chi-square statistic having 2 degrees of freedom is more than the table value (refer Chi-square table ,hence null hypothesis is not accepted. Thus, we conclude that there is a relationship between gender and voting preference.<\/p>\n<p>&nbsp;<\/p>\n<p><strong>Self-Check Questions:<\/strong><\/p>\n<p>&nbsp;<\/p>\n<p style=\"text-align: justify\"><strong>Question 1<\/strong>: Two hundred randomly selected adults were asked whether TV shows as a whole are primarily entertaining, educational or boring. The respondents were categorized by gender. Their responses are given in the following table-<\/p>\n<table class=\"aligncenter\" style=\"width: 60%\">\n<tbody>\n<tr>\n<td style=\"width: 64.0625px\"><strong>Gender<\/strong><\/td>\n<td style=\"width: 102.063px\"><strong>Opinion Entertaining<\/strong><\/td>\n<td style=\"width: 96.0625px\"><strong>Educational<\/strong><\/td>\n<td style=\"width: 109.063px\"><strong>Waste of time<\/strong><\/td>\n<td style=\"width: 50.0625px\"><strong>Total<\/strong><\/td>\n<\/tr>\n<tr>\n<td style=\"width: 64.0625px\"><strong>Female<\/strong><\/td>\n<td style=\"width: 102.063px\">52<\/td>\n<td style=\"width: 96.0625px\">28<\/td>\n<td style=\"width: 109.063px\">30<\/td>\n<td style=\"width: 50.0625px\">110<\/td>\n<\/tr>\n<tr>\n<td style=\"width: 64.0625px\"><strong>Male<\/strong><\/td>\n<td style=\"width: 102.063px\">28<\/td>\n<td style=\"width: 96.0625px\">12<\/td>\n<td style=\"width: 109.063px\">50<\/td>\n<td style=\"width: 50.0625px\">90<\/td>\n<\/tr>\n<tr>\n<td style=\"width: 64.0625px\"><strong>Total<\/strong><\/td>\n<td style=\"width: 102.063px\">80<\/td>\n<td style=\"width: 96.0625px\">40<\/td>\n<td style=\"width: 109.063px\">80<\/td>\n<td style=\"width: 50.0625px\">200<\/td>\n<\/tr>\n<\/tbody>\n<\/table>\n<p>&nbsp;<\/p>\n<p style=\"text-align: justify\">Is this evidence convincing that there is a relationship between gender and opinion in the population of interest?<\/p>\n<p>&nbsp;<\/p>\n<p style=\"text-align: justify\"><strong>Solution \u2013 <\/strong>Let us take the null hypothesis that the opinion of adults is independent of adults is independent of gender.<\/p>\n<p>&nbsp;<\/p>\n<p style=\"text-align: justify\">Since, contingency table is of size 2&#215;3,the degrees of freedom would be (2-1)(3-1) = 2.This implies that we need to calculate only to calculate only two expected frequencies and the other four can automatically be determined as shown below:<\/p>\n<\/div>\n<div>\n<p>&nbsp;<\/p>\n<p>E11=Row 1 total x Column 1 total<\/p>\n<p>&nbsp;<\/p>\n<p><span style=\"text-align: initial;font-size: 1em\">E13=110-(44+22)=44<\/span><\/p>\n<p>&nbsp;<\/p>\n<p><span style=\"text-align: initial;font-size: 1em\">E21=80-E11=40-22==18<\/span><\/p>\n<p>&nbsp;<\/p>\n<p><span style=\"text-align: initial;font-size: 1em\">E22=40-E12=40-22=18<\/span><\/p>\n<p>&nbsp;<\/p>\n<p><span style=\"text-align: initial;font-size: 1em\">E23=80-E13=80-44=36<\/span><\/p>\n<p>&nbsp;<\/p>\n<p><span style=\"text-align: initial;font-size: 1em\">The contingency table of expected frequencies is as follows:<\/span><\/p>\n<\/div>\n<div>\n<table class=\"aligncenter\" style=\"width: 60%\">\n<tbody>\n<tr>\n<td><strong>Gender<\/strong><\/td>\n<td><strong>Entertaining<\/strong><\/td>\n<td><strong>Opinion Educational<\/strong><\/td>\n<td><strong>Waste of time<\/strong><\/td>\n<td><strong>Total<\/strong><\/td>\n<\/tr>\n<tr>\n<td><strong>Female<\/strong><\/td>\n<td>44<\/td>\n<td>22<\/td>\n<td>44<\/td>\n<td>110<\/td>\n<\/tr>\n<tr>\n<td><strong>Male<\/strong><\/td>\n<td>36<\/td>\n<td>18<\/td>\n<td>36<\/td>\n<td>90<\/td>\n<\/tr>\n<tr>\n<td><strong>Total<\/strong><\/td>\n<td>80<\/td>\n<td>40<\/td>\n<td>80<\/td>\n<td>200<\/td>\n<\/tr>\n<\/tbody>\n<\/table>\n<p>&nbsp;<\/p>\n<p>Arranging the observed and expected frequencies as follows to calculate the value of <em>x<\/em>2-test statistic:<\/p>\n<table class=\"aligncenter\" style=\"width: 60%\">\n<tbody>\n<tr>\n<td><strong>Observed(O)<\/strong><\/td>\n<td><strong>Expected(E)<\/strong><\/td>\n<td><strong>O-E<\/strong><\/td>\n<td><strong>(O-E)2<\/strong><\/td>\n<td><strong>(O-E)2\/E<\/strong><\/td>\n<\/tr>\n<tr>\n<td>52<\/td>\n<td>44<\/td>\n<td>8<\/td>\n<td>64<\/td>\n<td>1.454<\/td>\n<\/tr>\n<tr>\n<td>28<\/td>\n<td>22<\/td>\n<td>6<\/td>\n<td>36<\/td>\n<td>1.636<\/td>\n<\/tr>\n<tr>\n<td>30<\/td>\n<td>44<\/td>\n<td>14<\/td>\n<td>196<\/td>\n<td>4.455<\/td>\n<\/tr>\n<tr>\n<td>28<\/td>\n<td>36<\/td>\n<td>-8<\/td>\n<td>64<\/td>\n<td>1.777<\/td>\n<\/tr>\n<tr>\n<td>12<\/td>\n<td>18<\/td>\n<td>-6<\/td>\n<td>36<\/td>\n<td>2<\/td>\n<\/tr>\n<tr>\n<td>50<\/td>\n<td>36<\/td>\n<td>14<\/td>\n<td>196<\/td>\n<td>5.444<\/td>\n<\/tr>\n<tr>\n<td><\/td>\n<td><\/td>\n<td><\/td>\n<td><\/td>\n<td>16.766<\/td>\n<\/tr>\n<\/tbody>\n<\/table>\n<\/div>\n<div>\n<p>&nbsp;<\/p>\n<p style=\"text-align: justify\">Since, calculated value of <em>x<\/em>2=16.766 is more than its critical value, <em>x<\/em>2=5.99 at \u03b1=0.05 and <em>df<\/em> = 2 ,the null hypothesis is rejected. Hence, we conclude that the opinion of adults is not independent of gender.<\/p>\n<p>&nbsp;<\/p>\n<p><strong style=\"text-align: initial;font-size: 1em\">Question 2: <\/strong><span style=\"text-align: initial;font-size: 1em\">A sample analysis of examination results of 500 students was made. It was found that 220 students had failed ,170 had secured a third division 90 were placed in second division and 20 got a first division. Are these figures commensurate with the general examination result which is the ratio of 4:3:2:1 for the various categories respectively?<\/span><\/p>\n<\/div>\n<div><\/div>\n<div>\n<p>&nbsp;<\/p>\n<p style=\"text-align: justify\"><strong>Solution<\/strong>&#8211; Let us take the null hypothesis that the observed results are commensurate with the general examination result which is the ratio 4:3:2:1.<\/p>\n<p>&nbsp;<\/p>\n<p>The expected number of students who have failed, obtained a third division second division and first division, respectively, are<\/p>\n<p>&nbsp;<\/p>\n<p>E1=500*4\/10=200, E2=500*3\/10=150;E3=500*2\/10=100 AND E4=500*1\/10=50<\/p>\n<p>&nbsp;<\/p>\n<p>The contingency table of expected and observed frequencies is as follows:<\/p>\n<table class=\"aligncenter\" style=\"width: 60%\">\n<tbody>\n<tr>\n<td><strong>Category<\/strong><\/td>\n<td><strong>O<\/strong><\/td>\n<td><strong>E<\/strong><\/td>\n<td><strong>(O-E)<\/strong><strong>2<\/strong><\/td>\n<td><strong>\u03a7<\/strong><strong>2<\/strong><strong>=(O-E)<\/strong><strong>2<\/strong><strong>\/E<\/strong><\/td>\n<\/tr>\n<tr>\n<td>Failed<\/td>\n<td>220<\/td>\n<td>200<\/td>\n<td>400<\/td>\n<td>2<\/td>\n<\/tr>\n<tr>\n<td>3rd division<\/td>\n<td>170<\/td>\n<td>150<\/td>\n<td>400<\/td>\n<td>2.667<\/td>\n<\/tr>\n<tr>\n<td>2nd division<\/td>\n<td>90<\/td>\n<td>100<\/td>\n<td>100<\/td>\n<td>1<\/td>\n<\/tr>\n<tr>\n<td>1st division<\/td>\n<td>20<\/td>\n<td>50<\/td>\n<td>900<\/td>\n<td>18<\/td>\n<\/tr>\n<tr>\n<td><\/td>\n<td><\/td>\n<td><\/td>\n<td><\/td>\n<td>23.667<\/td>\n<\/tr>\n<\/tbody>\n<\/table>\n<p>&nbsp;<\/p>\n<p style=\"text-align: justify\">Since calculated value of <em>x<\/em>2 = 23.667 is more than its table value , <em>x<\/em>2=7.81 at \u03b1 = 0.05 level of significance and df= n \u2013 1 = 4 -1 =3 the hypothesis is rejected.<\/p>\n<p>&nbsp;<\/p>\n<p style=\"text-align: justify\"><strong>Question 3<\/strong>: Based on information on 1000 randomly selected fields about the tenancy status of the cultivation of these fields and use of fertilizers ,collected in an AGRO ECONOMY survey, the following classification was noted:<\/p>\n<table class=\"aligncenter\" style=\"width: 60%\">\n<tbody>\n<tr>\n<td><\/td>\n<td><strong>Owned<\/strong><\/td>\n<td><strong>Rented<\/strong><\/td>\n<td><strong>Total<\/strong><\/td>\n<\/tr>\n<tr>\n<td><strong>Using fertilizers<\/strong><\/td>\n<td>416<\/td>\n<td>184<\/td>\n<td>600<\/td>\n<\/tr>\n<tr>\n<td><strong>Not using fertilizers<\/strong><\/td>\n<td>64<\/td>\n<td>336<\/td>\n<td>400<\/td>\n<\/tr>\n<tr>\n<td><strong>Total<\/strong><\/td>\n<td>480<\/td>\n<td>5220<\/td>\n<td>1000<\/td>\n<\/tr>\n<\/tbody>\n<\/table>\n<\/div>\n<div><span style=\"font-size: 1em;text-align: initial\">Would you conclude that owner cultivators are more towards the use of fertilizers at 5%level of significance? Carry out a chi-square test as per testing procedure.<\/span><\/div>\n<p>&nbsp;<\/p>\n<p style=\"text-align: justify\"><strong>Solution<\/strong>: Let us take the hypothesis that ownership of fields and the use of fertilizers are independent attributes. Since, contingency table is of size 2*2 the degree of freedom would be (2-1)(2-1)=1. This implies that we need to calculate only one expected frequency and others can be automatically determined as follows:<\/p>\n<p>&nbsp;<\/p>\n<p>E11=600*480\/1000=288<\/p>\n<p>&nbsp;<\/p>\n<p>E12 =600-288=312<\/p>\n<p>&nbsp;<\/p>\n<p>E21=480-288=192<\/p>\n<p>&nbsp;<\/p>\n<p>E22=208<\/p>\n<p>&nbsp;<\/p>\n<p>The contingency table of expected frequencies is as follows:<\/p>\n<table class=\"aligncenter\" style=\"width: 60%\">\n<tbody>\n<tr>\n<td><strong>Observed<\/strong><\/td>\n<td><strong>Expected<\/strong><\/td>\n<td><strong>(O-E)<\/strong><strong>2<\/strong><\/td>\n<td><strong><em>x<\/em><\/strong><strong>2<\/strong><strong>=(O-E)<\/strong><strong>2<\/strong><strong>\/E<\/strong><\/td>\n<\/tr>\n<tr>\n<td>416<\/td>\n<td>288<\/td>\n<td>16,384<\/td>\n<td>56.889<\/td>\n<\/tr>\n<tr>\n<td>64<\/td>\n<td>192<\/td>\n<td>16,384<\/td>\n<td>85.333<\/td>\n<\/tr>\n<tr>\n<td>184<\/td>\n<td>312<\/td>\n<td>16,384<\/td>\n<td>52.513<\/td>\n<\/tr>\n<tr>\n<td>336<\/td>\n<td>208<\/td>\n<td>16384<\/td>\n<td>78.769<\/td>\n<\/tr>\n<tr>\n<td><\/td>\n<td><\/td>\n<td><\/td>\n<td>273.534<\/td>\n<\/tr>\n<\/tbody>\n<\/table>\n<p>&nbsp;<\/p>\n<p>The calculated value of <em>x<\/em>2=273.534 at \u03b1=0.05 level of significance and df= (n-1) (r-1) =(2-1) (2-1) = 1 is much more than its table value,\u03c72=3.84. The null hypothesis H0 is rejected. Hence, it can be conducted<\/p>\n<p style=\"text-align: justify\">that owners\u2019 cultivators are more inclined towards the use of fertilizers.<\/p>\n<p>&nbsp;<\/p>\n<p><strong>Summary<\/strong><\/p>\n<p>&nbsp;<\/p>\n<p style=\"text-align: justify\">The tests may be classified in to two category mainly Parametric and Non-Parametric. Three test i.e. t test, z test and f test are used to estimate and test the population parameters and prerequisite of application of these test are-interval and ratio Scale to be used, hypothesis testing for specific parameters, assumption of normality and Standard deviation is known or not should be clear.<\/p>\n<p>&nbsp;<\/p>\n<p style=\"text-align: justify\">The absence of these conditions leads to the application of Non-Parametric Tests or distribution free tests. These test are very easy to apply and can use nominal or ordinal data as well for calculation. These tests provide broad based conclusion with approximate solution and does not necessarily require normally distributed population. The Chi-Square test is one of the non-parametric tests used to test hypothesis. Chi-Square Test for Independence test is applied when you have two categorical variables from a single population. It is used to determine whether there is a significant association between the two variables.<\/p>\n<p>&nbsp;<\/p>\n<p style=\"text-align: center\"><strong>Learn More:<\/strong><\/p>\n<ol>\n<li>Sharma, J K (2014), Business Statistics, S Chand &amp; Company, N Delhi.<\/li>\n<li>Bajpai, N (2010) Business Statistics, Pearson, N Delhi.<\/li>\n<li>Trevor Hastie, Robert Tibshirani, Jerome Friedman (2009), The Elements of Statistical Learning: Data Mining, Inference, and Prediction, 2nd Edition, Springer.<\/li>\n<li>Darrell Huff (2010), How to Lie with Statistics,\u00a0 W. W. Norton, California.<\/li>\n<li>K.R. Gupta (2012), Practical Statistics, Atlantic Publishers &amp; Distributors (P) Ltd., N. Delhi.<\/li>\n<\/ol>\n","protected":false},"author":3,"menu_order":25,"template":"","meta":{"pb_show_title":"on","pb_short_title":"","pb_subtitle":"","pb_authors":["dr-deependra-sharma"],"pb_section_license":""},"chapter-type":[],"contributor":[59],"license":[],"class_list":["post-272","chapter","type-chapter","status-publish","hentry","contributor-dr-deependra-sharma"],"part":3,"_links":{"self":[{"href":"https:\/\/ebooks.inflibnet.ac.in\/mgmtp15\/wp-json\/pressbooks\/v2\/chapters\/272","targetHints":{"allow":["GET"]}}],"collection":[{"href":"https:\/\/ebooks.inflibnet.ac.in\/mgmtp15\/wp-json\/pressbooks\/v2\/chapters"}],"about":[{"href":"https:\/\/ebooks.inflibnet.ac.in\/mgmtp15\/wp-json\/wp\/v2\/types\/chapter"}],"author":[{"embeddable":true,"href":"https:\/\/ebooks.inflibnet.ac.in\/mgmtp15\/wp-json\/wp\/v2\/users\/3"}],"version-history":[{"count":5,"href":"https:\/\/ebooks.inflibnet.ac.in\/mgmtp15\/wp-json\/pressbooks\/v2\/chapters\/272\/revisions"}],"predecessor-version":[{"id":280,"href":"https:\/\/ebooks.inflibnet.ac.in\/mgmtp15\/wp-json\/pressbooks\/v2\/chapters\/272\/revisions\/280"}],"part":[{"href":"https:\/\/ebooks.inflibnet.ac.in\/mgmtp15\/wp-json\/pressbooks\/v2\/parts\/3"}],"metadata":[{"href":"https:\/\/ebooks.inflibnet.ac.in\/mgmtp15\/wp-json\/pressbooks\/v2\/chapters\/272\/metadata\/"}],"wp:attachment":[{"href":"https:\/\/ebooks.inflibnet.ac.in\/mgmtp15\/wp-json\/wp\/v2\/media?parent=272"}],"wp:term":[{"taxonomy":"chapter-type","embeddable":true,"href":"https:\/\/ebooks.inflibnet.ac.in\/mgmtp15\/wp-json\/pressbooks\/v2\/chapter-type?post=272"},{"taxonomy":"contributor","embeddable":true,"href":"https:\/\/ebooks.inflibnet.ac.in\/mgmtp15\/wp-json\/wp\/v2\/contributor?post=272"},{"taxonomy":"license","embeddable":true,"href":"https:\/\/ebooks.inflibnet.ac.in\/mgmtp15\/wp-json\/wp\/v2\/license?post=272"}],"curies":[{"name":"wp","href":"https:\/\/api.w.org\/{rel}","templated":true}]}}