Statistics
NCERT Class 11 Mathematics Chapter 13: Statistics (Pages 257–288)
Statistics at a Glance
CBSE
Class 11
Mathematics
Mathematics
13
257–288
7 study resources
Statistics is a chapter in the CBSE Class 11 Mathematics syllabus from Mathematics. This chapter hub brings together revision notes, practice questions, worksheets, flashcards, formula sheet to help students learn, practice, and revise Statistics effectively.
Scroll down to find Statistics notes, practice questions, worksheets, and revision resources — all in one place. Use the sidebar to jump to any section, or browse the full page below.
NCERT Class 11 Mathematics Chapter 13: Statistics (Pages 257–288)
CBSE
Class 11
Mathematics
Mathematics
13
257–288
7 study resources
Download the Statistics revision guide with key points, summaries, and quick revision notes for CBSE Class 11 Mathematics.
Key Points
Statistics deals with data analysis.
Statistics is about collecting, analyzing, and interpreting data to make informed decisions.
Measures of central tendency: Mean, Median, Mode.
These values summarize data, indicating where data points cluster. Mean is the average, median is the middle value, and mode is the most frequent.
Understanding variability in data.
Beyond averages, understanding how data is spread out (dispersion) is key for complete data interpretation.
Range: A basic measure of dispersion.
Range is calculated as the difference between the maximum and minimum values in a dataset (Range = Max - Min).
Mean Deviation (M.D.).
M.D. quantifies the average absolute deviation from a central tendency, calculated as M.D. = (Sum of |x - a|) / n.
Steps for Mean Deviation.
1) Calculate central measure (mean/median). 2) Find deviations. 3) Calculate absolute values. 4) Average the deviations.
Standard Deviation (S.D.).
A robust measure of dispersion indicating how much values deviate from the mean, calculated using the formula σ = √(Σ(x - x̄)²/n).
Variance is the square of standard deviation.
Variance (σ²) quantifies the degree of dispersion in a dataset and helps in understanding data variability.
Empirical Rule for Normal Distribution.
In a normal distribution, about 68% of data falls within 1 S.D. from the mean, 95% within 2 S.D., and 99.7% within 3 S.D.
Grouped data representation.
When data is organized in classes, calculations like Mean and S.D. can be performed using class midpoints.
Finding median in grouped data.
Identify the median class where N/2 lies in the cumulative frequency table, then apply the median formula.
Step-deviation method for ease.
In complex data, one can shift the mean and work with step deviations to simplify calculations.
Difference between Mean Deviation and Standard Deviation.
M.D. considers absolute deviations, while S.D. considers squared deviations, reflecting larger variances more effectively.
Importance of the total frequency (N).
In statistical calculations, knowing the total number of observations is crucial for accurate mean and variance computations.
Limitations of Mean Deviation.
Mean deviation might not reflect dispersion well in data with high variability or outliers.
Quartile deviation measures dispersion.
Quartile Deviation focuses on the middle 50% of the data, indicating variability within this central range.
Identify and utilize cumulative frequencies.
Cumulative frequencies assist in determining medians and understanding data distribution over ranges.
Practical applications of statistics.
Statistics apply to diverse fields like health, economics, and social sciences for informed decision-making.
Historical context of statistics.
Statistics has evolved through significant contributions from various scholars, aiding in data analysis throughout history.
Tools and techniques in statistical analysis.
Students must familiarize themselves with various statistical tools, like calculators and software, for efficient data analysis.
Review important statistical vocabulary.
Key terms such as 'population', 'sample', 'parameter', and 'statistic' are fundamental for understanding statistics.
Practice important questions and exam-style problems from Statistics. These questions cover key topics from the CBSE Class 11 Mathematics syllabus.
How to practice: Start with the questions below to test your understanding of Statistics. Use the revision guide to review concepts you find difficult, then come back and retry the questions for better retention.
What is the primary purpose of statistics?
Which of the following is NOT a measure of central tendency?
If the number of observations is even, how is the median computed?
What is meant by 'measure of dispersion'?
Which of the following statements is true regarding the two batsmen's scores in the given context?
What can be inferred if a set of data has a small range?
Which method would you use to display the frequency of data points?
How do we calculate the mean of a data set?
In the context of data analysis, why is it essential to understand variability?
What statistical measure indicates the most frequently occurring value in a dataset?
If a data set has an outlier, how does it affect the mean?
Which of the following accurately describes the median when data is sorted?
What does a high value of standard deviation indicate?
Which term describes the difference between the highest and lowest values in a dataset?
What will happen to the mean if we add a new number that is much larger than the existing mean?
What is the range of the data set: {4, 8, 15, 16, 23, 42}?
Find the range of the following scores: 22, 28, 19, 35, 30.
If the lowest score in a set is 10 and the highest score is 50, what is the range?
A student scored 55, 67, 78, and 82 in four tests. What is the range of their scores?
In a survey, the ages of participants ranged from 22 to 45. What can be concluded about the range?
If the range of a dataset is 0, what does it imply about the data?
Given the data set {5, 7, 9, 20, 22}, what is the range?
The marks of five students are 45, 74, 68, 86, and 52. What is the range of their marks?
Which statement about the range is false?
What is the range of the following data set: {13, 20, 34, 18, 27}?
In a football game, the scores of team A are {2, 3, 5, 8, 7} and team B are {1, 4, 5, 6, 3}. What is the range difference between the two teams?
The temperature over a week was recorded as 72°F, 75°F, 78°F, 71°F, 77°F. What is the range of temperatures?
If the top scorer in a team scored 100 runs and the lowest scorer scored 10 runs, what is the team's range in runs?
The range of a set of heights is calculated as the difference between the tallest and the shortest person. Which of the following statements is true?
What will be the range if all values in a dataset are increased by a constant value?
What is the formula for calculating range in a dataset?
If the scores of a batsman are 45, 50, 55, 60, and 65, what is the range of these scores?
Which of the following is a measure of dispersion that considers all data points?
If the mean deviation from the mean of a dataset is 12, what does this imply?
In the context of data analysis, why is standard deviation preferred over range?
If the mean of a dataset is 50 and the deviations from the mean are -2, 0, 3, and 5, what is the mean deviation?
What type of data does the standard deviation measure most effectively?
If two datasets have the same mean but different standard deviations, what can you infer?
What is the effect of outliers on the mean and standard deviation?
Which formula represents the standard deviation for a sample of n observations?
What happens to the variance if each value in a dataset is multiplied by 2?
How is mean deviation calculated if the median is used as a measure of central tendency?
In a dataset with a high standard deviation, what can be expected about the distance of values from the mean?
What is the primary limitation of using range as a measure of dispersion?
What is the mean deviation about the mean for the data set 4, 8, 6, 5, 3?
If the numbers are 7, 8, 9 and 10, what is the mean deviation about the median?
The data set is 14, 18, 20, 22, and 24. What is the mean deviation from the mean?
What is the mean deviation of the data set 2, 4, 6, 8, and 10 about the mean?
For the data set {3, 5, 7, 9, 11}, what is the mean deviation about the mean?
Calculate the mean deviation for the data set {10, 20, 30, 40, 50} about the median.
Which of the following describes the mean deviation?
If a dataset has a mean of 50 and the mean deviation is 0, which statement must be true?
How do you calculate the mean deviation for grouped data?
Consider the data set {5, 15, 25, 25, 35}. How do you find the mean deviation about the median?
Why might the mean deviation be preferred over standard deviation?
In a uniform distribution, what can you say about the mean deviation of the data?
In a set of data, if the mean deviation is known to be low, what does this imply?
Given the following frequencies: {2, 3, 4, 6} and corresponding midpoints {1, 2, 3, 4}, what is the mean deviation?
If a dataset consists of values with 0 variance, what is the mean deviation?
What is the formula for calculating variance for a sample?
If all values in a data set are increased by 5, how does the variance change?
For given data points, which statement regarding standard deviation is true?
In a dataset of {4, 8, 6, 5, 3}, what is the standard deviation?
Variance of a dataset is affected primarily by which of the following?
When measuring the spread of a frequency distribution, which formula is used for variance?
Which condition will lead to a lower standard deviation?
If the variance of a population is 9, what is the standard deviation?
In calculating standard deviation, which mistake is common?
Which of these datasets will have the highest standard deviation?
If all data points of a dataset are multiplied by a constant, how does this affect the variance?
Download and practice Statistics worksheets to improve problem-solving accuracy and speed for CBSE Class 11 Mathematics exams.
This worksheet covers essential long-answer questions to help you build confidence in Statistics from Mathematics for Class 11 (Mathematics).
Questions
Define statistics and explain its importance in decision making with examples.
Statistics is a branch of mathematics dealing with collecting, analyzing, interpreting, presenting, and organizing data. It is crucial in decision making because it allows individuals and organizations to make informed conclusions based on data rather than assumptions. For example, in business, statistics help in market analysis to understand consumer preferences. In healthcare, statistical methods can analyze the effectiveness of a new treatment. These analyses can drive strategic decisions that affect numerous stakeholders. Consequently, acquiring statistical skills enhances a person's ability to interpret data efficiently and make sound decisions based on empirical evidence.
Explain the measures of central tendency and how they are calculated in a dataset.
Measures of central tendency include mean, median, and mode. The mean is calculated by summing all data points and dividing by the number of observations. The formula is: Mean (x̄) = Σx / n. The median is the middle value when the data points are arranged in ascending order. In case of an even number of observations, the median is the average of the two middle numbers. The mode is the value that appears most frequently in a data set. For example, consider the dataset: {1, 2, 2, 3, 4}. The mean is (1+2+2+3+4)/5 = 2. The median is 2, and the mode is also 2. This illustrates how central tendency provides a summary of the data's typical value.
Discuss how variance and standard deviation are used to measure dispersion in a dataset.
Variance measures how far data points are from the mean. It is calculated by finding the average of the squared differences from the mean. The formula is: Variance (σ²) = Σ(xi - x̄)² / n, where xi represents each data point and n is the number of points. Standard deviation is the square root of variance and provides a measure of dispersion in the same unit as the data. It reflects the spread of data points; a small standard deviation indicates that points are close to the mean, while a large one shows they are spread out. For instance, if we have a dataset of exam scores, high standard deviation indicates varied performance levels among students, which may necessitate different teaching strategies.
What is the range of a dataset, and how do you calculate it? Give an example.
The range is a simple measure of dispersion that indicates the extent of variation within a dataset. It is calculated by subtracting the smallest value from the largest value in the dataset. The formula for range is: Range = Maximum value - Minimum value. For example, if we have the dataset: {3, 7, 9, 5, 12}, the maximum value is 12, and the minimum value is 3. Thus, the range is 12 - 3 = 9. This implies that the scores have a spread of 9 units, giving a quick sense of the overall variability in the dataset.
Define and differentiate between mean deviation and standard deviation.
Mean deviation is the average of the absolute deviations of each data point from the mean or median, while standard deviation measures the square root of the variance, providing insight into data spread relative to the mean. Mean deviation is calculated using the formula: M.D. = Σ|xi - x̄| / n, where |xi - x̄| are absolute deviations. In contrast, standard deviation is calculated using: σ = √(Σ(xi - x̄)² / n). While mean deviation gives us an idea of dispersion without indicating direction (positive or negative), standard deviation takes into account the squared values of deviations, making it sensitive to outliers. Thus, standard deviation can be more informative but also more complex to understand in practical contexts.
Explain how the median is calculated for grouped data and illustrate with a sample dataset.
To calculate the median for grouped data, we first determine the cumulative frequency and identify the class interval containing the median. The median is calculated using the formula: Median = l + (N/2 - CF) / f × h, where l is the lower limit of the median class, CF is the cumulative frequency of the class preceding the median class, f is the frequency of the median class, and h is the class width. For example, consider the grouped data: Class intervals: [10-20], [20-30], [30-40], with frequencies 5, 15, 10 respectively. The total frequency N = 30. Thus, N/2 = 15. The cumulative frequencies are 5, 20, 30. The median falls in the second class ([20-30]) since the cumulative frequency just exceeds 15. Applying the formula gives us the final median value.
Describe how outliers can affect the mean, median, and mode of a dataset.
Outliers are extreme values that differ significantly from other observations; they can skew distributions and greatly influence statistical measures. The mean is particularly sensitive to outliers, as it incorporates all data points. For example, in the dataset {1, 2, 2, 3, 100}, the mean is 21.6, which does not represent the center of the majority of data. The median, however, is less affected; in the same dataset, it remains 2, showing a better central tendency. The mode, being the most frequent value, is also resilient to outliers unless the outlier affects frequency. Thus, outliers can distort the mean, while the median and mode may provide a more accurate representation of central tendency in skewed datasets.
What is a frequency distribution, and how does it relate to measures of central tendency?
A frequency distribution is a summary of how often each value occurs in a dataset, typically organized into classes or intervals for ease of analysis. It provides a visual representation of data, helping to identify patterns, trends, or outliers. Measures of central tendency, such as mean, median, and mode, summarize frequency distributions by indicating where most values lie. For instance, in a frequency distribution of exam scores, the mean score highlights the average student performance, while the median indicates the score at which half the students performed better or worse. This relationship allows statisticians to draw insights from data distributions effectively and make predictions or informed decisions based on empirical evidence.
Provide the formula for calculating quartiles and explain their significance in data analysis.
Quartiles divide a dataset into four equal parts, providing insight into data dispersion and distribution shape. The first quartile (Q1) is the median of the lower half of the dataset, the second quartile (Q2) is the median, and the third quartile (Q3) is the median of the upper half. Calculation involves arranging data in ascending order: Q1 = (n + 1) * 1/4 th observation, Q2 = (n + 1) * 1/2 th observation, Q3 = (n + 1) * 3/4 th observation. Quartiles are crucial for understanding the spread and identifying outliers in data analysis, as they help illustrate how data points are distributed relative to the median. For example, in income data analysis, knowing the quartiles can indicate income disparity and aid in policymaking.
This worksheet challenges you with deeper, multi-concept long-answer questions from Statistics to prepare for higher-weightage questions in Class 11.
Questions
Batsman A and B have scored 30, 91, 0, 64, 42, 80, 30, 5, 117, 71 and 53, 46, 48, 50, 53, 53, 58, 60, 57, 52 runs respectively. Calculate the mean, median, range, and standard deviation for both batsmen. Discuss how these statistics reflect the performance and consistency of each batsman.
Mean: Both batsmen have a mean of 53. Median: Both have a median of 53. Range of A = 117; Range of B = 14. Standard deviation calculations show A is more variable in performance. A graph can illustrate the distribution.
Given the data set: 4, 7, 8, 9, 10, 12, 13, 17. Calculate the mean deviation about the mean and the median. Compare the results. What does this indicate about the data set's dispersion?
Mean = 9.125, Median = 10.5. M.D. about mean = 2.75, about median = 3.25. The which indicates a stronger clustering of data points around the mean compared to median.
For a continuous frequency distribution of students' scores: Class 0-10 (f=6), 10-20 (f=7), 20-30 (f=15), 30-40 (f=12), calculate the mean, variance, and standard deviation. Discuss how these measures assess the central tendency and dispersion.
Mean: calculate mid-points and totals. Variance and standard deviation follow from the mean calculated. Discuss the implications regarding clustering of scores.
Consider the following measures: Range, Mean deviation, Variance, and Standard deviation. Define each and compare their applicability when analyzing data sets. Provide an example for different scenarios illustrating their practical application.
Define each measure. Comparisons: Range is a simple spread measure. Mean deviation reflects average distances from a center. Variance and standard deviation provide insights into variability. Example: Use varied data sets to highlight each measure's contribution.
A student took scores of subjects with data: 70, 80, 60, 75, 90. Find the standard deviation using both direct and shortcut methods. Discuss which method was easier and reflect on why that might be your preference.
Direct method gives a standard deviation of 10.00; shortcut method confirms it is easier for larger datasets. Compare calculations.
Explore the effect of outliers on mean and median in a given dataset (e.g., 10, 20, 30, 40, 100). Calculate and analyze the shifts in these measures and suggest better representation options for skewed data.
Calculate mean (40) and median (30). Outlier (100) skews mean more than median. Discuss IQRs, box plots for better insights in future situations.
Use the following data on a discrete frequency distribution: x values 1, 2, 3, 4, 5 corresponding to frequencies 2, 3, 5, 2, 1 respectively. Calculate mean and variance. Discuss how these determine the dataset's shape.
Mean = 2.83, Variance = 1.35. Discuss normality and skewness implications from these values.
Critique the limitations of mean deviation compared to standard deviation when analyzing data variability. Provide examples in which mean deviation might fail or succeed.
Mean deviation often overlooks sign directionality; standard deviation compensates by squaring values. Provide scenarios that show advantages/disadvantages.
Plot the data of class intervals: 0-10 (5), 10-20 (10), 20-30 (15). Calculate the mean for this group and apply a visual graph representation. Discuss the benefits of graphical data interpretation.
Mean = 15. Visual graphs illustrate total trends; gauges central tendencies shape. Discuss perceptiveness gained.
Discuss how standard deviation and variance are used to analyze data spread in different fields (e.g., finance, education). Provide relevant examples that apply these measures practically.
Standard deviation tracks volatility in finance; variance in education highlights student performance disparities. Illustrate with field-appropriate scenarios.
The final worksheet presents challenging long-answer questions that test your depth of understanding and exam-readiness for Statistics in Class 11.
Questions
Evaluate the impact of using only the mean as a measure of central tendency when analyzing a dataset with extreme outliers. Compare this with the median and discuss the implications for data interpretation.
The mean can be skewed by outliers, while the median remains unaffected, offering a more reliable indicator of central tendency in such cases. Discuss how each measure represents the dataset differently.
Discuss how the calculation of standard deviation can vary between grouped and ungrouped data. Why is it important to understand these differences in practical applications?
The formula for standard deviation varies based on data type, influencing results. Comparison of calculations shows importance in settings like research, finance, or quality control where accuracy is crucial.
Create a dataset of your choice, calculate its mean, median, mode, and standard deviation. Analyze how these statistics reveal different facets of your data.
The dataset will present distinct trends through various statistics, highlighting how they complement each other. Discuss anomalies or typical behaviors observed.
Explain how the concept of variability can influence decision-making in fields such as healthcare or education. Provide specific examples to illustrate your point.
Variability indicates consistency in a dataset, impacting decisions like resource allocation in healthcare. An example might include varying recovery times across patient demographics affecting treatment plans.
Analyze the role of quartiles and interquartile range (IQR) in understanding data distribution. Why might these measures be preferred in certain analyses over the range or mean deviation?
Quartiles and IQR provide insights into data dispersion without being affected by extreme values, unlike range. Discuss when these are important in statistics for robust data representation.
If an experiment is repeated and observations are drawn from a non-normal distribution, how would this affect the calculation and interpretation of standard deviation?
In non-normal distributions, standard deviation might not adequately represent variability; alternative measures like median absolute deviation might be necessary. Analyze how this affects confidence in statistical conclusions.
Discuss the potential pitfalls of interpreting the standard deviation in stratified datasets. How can stratification impact the understanding of data variability?
Stratification may lead to misrepresentations if variances within subgroups are disregarded. A comparison of stratified vs. unstratified data illustrates this impact on statistical analysis.
Examine how the choice of different measures of central tendency can alter the narrative of a dataset. Provide examples from consumer behavior or environmental statistics.
Different measures highlight various aspects of data; choosing one over the other can create misleading impressions. Discuss case studies illustrating these effects in marketing or environmental modeling.
Critique the appropriateness of using mean deviation and standard deviation as measures of dispersion in highly skewed distributions. What alternatives might offer better insights?
In skewed distributions, mean deviation might not capture variability accurately; alternatives like trimmed means or robust measures provide improved insights. Discuss cases in finance or ecological studies.
Design a research proposal that incorporates various statistical measures discussed in class to address a real-world problem, justifying their use and application.
Outline a problem, calculate central tendencies and variabilities, and provide analysis. Justification hinges on the selected measures' robustness in offering insights into the research question.
Use this Class 11 Mathematics Statistics Formula Sheet for quick revision before school exams and CBSE exams. It brings together the important formulas, key concepts, and worked examples in one place so students can revise faster and download a printable PDF for offline study.
Important Formulas
Mean (x) = (Σxi) / n
x is the mean, Σxi is the sum of observations, and n is the number of observations. This formula calculates the average of a data set, summarizing central tendency.
Median (M) = {n+1}/2 th observation (if n is odd)
M is the median and n is the total number of observations. For an even number of observations, it is the average of the n/2 and (n/2 + 1) observations, providing the middle value.
Range = Maximum value - Minimum value
The range measures the spread of data by dividing the difference between the highest and lowest values, offering a quick sense of dispersion.
Mean Deviation (M.D.) = (Σ|xi - a|) / n
M.D. gives the average of absolute deviations from a central value a. It captures the dispersion of data points around this central point.
Variance (σ²) = (Σ(xi - x)²) / n
σ² represents the variance of the data, showing the average squared deviation from the mean x. It quantifies data variability.
Standard Deviation (σ) = √Variance
σ denotes the standard deviation, providing a measure of the average distance of data points from the mean, in the same units as the data.
M.D. (about Median) = (Σ|xi - M|) / n
Calculates mean deviation from the median, giving insights into how spread out the data is around the middle value.
Grouped Data M.D. = (Σf|xi - a|) / N
In this formula, f is frequency, xi are midpoints, and N is total frequency, used to find mean deviation for grouped data.
For Continuous Data: M.D. = (Σf|xi - a|) / N
Similar to grouped data, it helps in calculating the mean deviation by considering the midpoints of class intervals.
Coefficient of Variation (CV) = (σ / x) × 100
CV expresses the standard deviation as a percentage of the mean, allowing comparison of variability between different data sets.
Worked Examples
Σxi = (x1 + x2 + ... + xn)
This represents the sum of all observations in a data set, essential for calculating mean and other statistics.
Percentile = (n * p) / 100 th observation
p is the desired percentile. This gives the position of a value below which a given percentage of observations fall.
Interquartile Range (IQR) = Q3 - Q1
Q3 and Q1 are the third and first quartiles respectively, measuring the middle 50% of the data and providing insights into data spread.
Standard Score (Z) = (xi - μ) / σ
Z represents how many standard deviations an observation xi is from the mean μ, enabling comparison across different distributions.
Skewness = [3(Mean - Median)] / SD
This formula assesses data symmetry around the mean, indicating whether data is left or right-skewed.
Kurtosis = (Σ (xi - μ)⁴ / n) / (σ⁴)
Kurtosis measures the tailedness of the distribution, indicating how outlier-prone a distribution is.
z-score for grouped data = (xi - Mode) / SD
This transforms grouped data observations into a standardized form for comparing relative positions in distribution.
Frequencies: fi = Total Observations / Class Width
Used for classifying data into intervals, essential in statistics for creating histograms and frequency distributions.
Σf = N
Where Σf represents the total of the frequencies, indicating that the sum must equal the total number of observations.
Chebyshev’s Inequality: 1 - (1/k²)
This inequality provides a lower bound on the proportion of values that lie within k standard deviations of the mean.
Explore More Statistics Resources
Explore more chapter resources to strengthen your understanding and prepare for exams.
Explore vital concepts in statistics with our comprehensive chapter on measures of dispersion, including variance, standard deviation, and mean deviation. Ideal for Class 11 students.
Download worksheets, revision guides, formula sheets, and the official textbook PDF for Statistics.
Statistics Official Textbook PDF
Download the official NCERT/CBSE textbook PDF for Class 11 Mathematics.
Statistics Revision Guide
Use this one-page guide to revise the most important ideas from Statistics.
Statistics Formula Sheet
Download the Statistics formula sheet PDF with important formulas, worked examples, and quick revision support for exam preparation.
Statistics Practice Worksheet
Solve basic and application-based questions from Statistics.
Statistics Mastery Worksheet
Work through mixed Statistics questions to improve accuracy and speed.
Statistics Challenge Worksheet
Try harder Statistics questions that test deeper understanding.
Statistics Question Bank
Download important questions and exam-style prompts from Statistics.
Revise key terms and definitions from Statistics with interactive flashcards. Quick recall practice for CBSE Class 11 Mathematics.
Practice Statistics with Interactive Duels
Master Statistics via Live Academic Duels
Challenge your classmates or test your individual retention on the core concepts of CBSE Class 11 Mathematics (Mathematics). Compete in speed-recall question rounds matched explicitly to the latest syllabus milestones for Statistics.
Quick, competitive practice on Statistics with zero setup.