Question

In: Statistics and Probability

Topic: Discussion: Mean, Median and Mode To complete the Discussion activity, please do the following: Answer...

Topic: Discussion: Mean, Median and Mode

To complete the Discussion activity, please do the following:

Answer each question fully. Use a minimum of 250 words for a complete discussion post. While it is not required in this discussion, feel free to bring in outside resources to support your answers. Outside resources include materials outside of the textbook, such as a website.

An average is an attempt to summarize a collection of data into just one number.

  • Explain how the mean, median, and mode all represent averages in this context.
  • Why is the mean a balance point? Why is the median a midway point? Why is the mode the most common data point?
  • List three areas of daily life in which you think the mean, median, or mode would be the best choice to describe an “average and explain why.

Solutions

Expert Solution

Introduction

A measure of central tendency is a single value that attempts to describe a set of data by identifying the central position within that set of data. As such, measures of central tendency are sometimes called measures of central location. They are also classed as summary statistics. The mean (often called the average) is most likely the measure of central tendency that you are most familiar with, but there are others, such as the median and the mode.

The mean, median and mode are all valid measures of central tendency, but under different conditions, some measures of central tendency become more appropriate to use than others. In the following sections, we will look at the mean, mode and median, and learn how to calculate them and under what conditions they are most appropriate to be used.

Mean (Arithmetic)

The mean (or average) is the most popular and well known measure of central tendency. It can be used with both discrete and continuous data, although its use is most often with continuous data (see our Types of Variable guide for data types). The mean is equal to the sum of all the values in the data set divided by the number of values in the data set. So, if we have n values in a data set and they have values x1,x2, …,xn, the sample mean, usually denoted by x¯ (pronounced "x bar"), is:

x¯=x1+x2+⋯+xn/n

This formula is usually written in a slightly different manner using the Greek capitol letter, ∑, pronounced "sigma", which means "sum of...":

x¯=∑xi

You may have noticed that the above formula refers to the sample mean. So, why have we called it a sample mean? This is because, in statistics, samples and populations have very different meanings and these differences are very important, even if, in the case of the mean, they are calculated in the same way. To acknowledge that we are calculating the population mean and not the sample mean, we use the Greek lower case letter "mu", denoted as μ:

μ=∑x/n

The mean is essentially a model of your data set. It is the value that is most common. You will notice, however, that the mean is not often one of the actual values that you have observed in your data set. However, one of its important properties is that it minimises error in the prediction of any one value in your data set. That is, it is the value that produces the lowest amount of error from all other values in the data set.

An important property of the mean is that it includes every value in your data set as part of the calculation. In addition, the mean is the only measure of central tendency where the sum of the deviations of each value from the mean is always zero.

When not to use the mean

The mean has one main disadvantage: it is particularly susceptible to the influence of outliers. These are values that are unusual compared to the rest of the data set by being especially small or large in numerical value. For example, consider the wages of staff at a factory below:

Staff 1 2 3 4 5 6 7 8 9 10
Salary 15k 18k 16k 14k 15k 15k 12k 17k 90k 95k

The mean salary for these ten staff is $30.7k. However, inspecting the raw data suggests that this mean value might not be the best way to accurately reflect the typical salary of a worker, as most workers have salaries in the $12k to 18k range. The mean is being skewed by the two large salaries. Therefore, in this situation, we would like to have a better measure of central tendency. As we will find out later, taking the median would be a better measure of central tendency in this situation.

Another time when we usually prefer the median over the mean (or mode) is when our data is skewed (i.e., the frequency distribution for our data is skewed). If we consider the normal distribution - as this is the most frequently assessed in statistics - when the data is perfectly normal, the mean, median and mode are identical. Moreover, they all represent the most typical value in the data set. However, as the data becomes skewed the mean loses its ability to provide the best central location for the data because the skewed data is dragging it away from the typical value. However, the median best retains this position and is not as strongly influenced by the skewed values. This is explained in more detail in the skewed distribution section later in this guide.

Median

The median is the middle score for a set of data that has been arranged in order of magnitude. The median is less affected by outliers and skewed data. In order to calculate the median, suppose we have the data below:

65 55 89 56 35 14 56 55 87 45 92

We first need to rearrange that data into order of magnitude (smallest first):

14 35 45 55 55 56 56 65 87 89 92

Our median mark is the middle mark - in this case, 56 (highlighted in bold). It is the middle mark because there are 5 scores before it and 5 scores after it. This works fine when you have an odd number of scores, but what happens when you have an even number of scores? What if you had only 10 scores? Well, you simply have to take the middle two scores and average the result. So, if we look at the example below:

65 55 89 56 35 14 56 55 87 45

We again rearrange that data into order of magnitude (smallest first):

14 35 45 55 55 56 56 65 87 89

Only now we have to take the 5th and 6th score in our data set and average them to get a median of 55.5.

Mode

The mode is the most frequent score in our data set. On a histogram it represents the highest bar in a bar chart or histogram.

We are now stuck as to which mode best describes the central tendency of the data. This is particularly problematic when we have continuous data because we are more likely not to have any one value that is more frequent than the other. For example, consider measuring 30 peoples' weight (to the nearest 0.1 kg). How likely is it that we will find two or more people with exactly the same weight (e.g., 67.4 kg)? The answer, is probably very unlikely - many people might be close, but with such a small sample (30 people) and a large range of possible weights, you are unlikely to find two people with exactly the same weight; that is, to the nearest 0.1 kg. This is why the mode is very rarely used with continuous data.


Related Solutions

Mean , Median and Mode
The mean of 8,11,6,14,m and 13 is 11 a) Find the value of m. b) Find the median and mode of given data set.
how do i interpret mean, mode and median?
how do i interpret mean, mode and median?
What is the best measure of center (mean, median, or mode) for each topic? And why?...
What is the best measure of center (mean, median, or mode) for each topic? And why? Country Infant mortality Health expenditure Obesity rate Average income Life expectancy Diabetes rate Leading cause of death
What are the mean, median, and mode of a set of data, and how do they...
What are the mean, median, and mode of a set of data, and how do they differ from each other? What are the different type measures of dispersion? Provide examples of each from your experience.
What is the average, or mean, for the following sets? Median? Mode? Answer Average 7.43 10.00...
What is the average, or mean, for the following sets? Median? Mode? Answer Average 7.43 10.00 13.88 10.40 Median 6 10 14.5 10 Mode 5 12 19 7 Calculate the standard deviation and variance for a, b, c and d. explain the meaning of your results.   {5,4,12,2,1,6,5,10,20,5,10,6,11,7} {19,11,13,5,12,1,4,6,18,10,12,9,3,7,20} {10,16,5,15,19,6,8,20,19,14,12,17,11,13,19,18} {19,6,9,17,7,10,11,15,2,14,16,7,12,3,8} Using the information from question 1 and 2 calculate the z-score for a, b, c and d. explain the meaning of your results.
Find the mean, median, and mode of the following set of data. (Enter solutions for mode...
Find the mean, median, and mode of the following set of data. (Enter solutions for mode from smallest to largest. If there are any unused answer boxes, enter NONE in the last boxes.) (a) 6 6 7 8 9 12 Mean Median Mode Mode (b) 6 6 7 8 9 108 Mean Median Mode Mode
find the mean, median, mode, range, and Quartiles of the following 13,13,10,8,7,6,4,5,14,12,4,2,9,8,8,12
find the mean, median, mode, range, and Quartiles of the following 13,13,10,8,7,6,4,5,14,12,4,2,9,8,8,12
We are going to calculate the mean, median, and mode for two sets of data. Please...
We are going to calculate the mean, median, and mode for two sets of data. Please show your answer to one decimal place if necessary. Here is the first data set. 26 53 85 63 59 85 74 75 88 83 20 what is the mean (¯xx¯) of this data set? What is the median of this data set? What is the mode of this data set? Here is the second data set. 85 47 43 48 42 30 71...
We are going to calculate the mean, median, and mode for two sets of data. Please...
We are going to calculate the mean, median, and mode for two sets of data. Please show your answer to one decimal place if necessary. Here is the first data set. 22 76 25 85 28 29 72 22 95 50 23 what is the mean (¯xx¯) of this data set? What is the median of this data set? What is the mode of this data set? Here is the second data set. 30 22 41 38 47 39 79...
How do you understand and interpret mean, median, mode, standard deviation, and variance?
How do you understand and interpret mean, median, mode, standard deviation, and variance?
ADVERTISEMENT
ADVERTISEMENT
ADVERTISEMENT