AQA GCSE Statistics: Topic Overview & Revision Guide | AQA GCSE 统计:知识点梳理与复习指南

📚 AQA GCSE Statistics: Topic Overview & Revision Guide | AQA GCSE 统计:知识点梳理与复习指南

This guide covers the essential topics for the AQA GCSE Statistics specification, organised into a clear revision sequence. Each section highlights the key definitions, formulas, and exam techniques you need to apply with confidence.

本指南按照清晰的复习顺序,梳理AQA GCSE统计考纲中的核心知识点。每一节都突出关键定义、公式和考试技巧,帮助你自信应对考试。


1. Statistical Enquiry Cycle | 统计调查循环

The statistical enquiry cycle is the overarching framework for all statistics questions. You should know its five stages: specify the problem, plan, collect, process and present, and interpret and evaluate. Every exam question will expect you to link your working back to this cycle.

统计调查循环是所有统计题目的总体框架。你需要掌握其五个阶段:明确问题、制定计划、收集数据、处理和呈现数据、解释与评估。考试中每道题都要求你把解题过程联系回这一循环。

When answering an extended question, always state the purpose of the investigation, explain how you collected or would collect the data, show your calculations, present a suitable chart, and then write a conclusion that answers the original question.

回答拓展题时,务必说明调查目的,解释你如何收集或将要收集数据,展示计算过程,绘制合适图表,然后写出回应原问题的结论。

  • Use context to decide which average or chart is most appropriate.

    根据实际背景判断哪种平均数或图表最合适。

  • Check whether predictions are reliable by considering sample size and possible bias.

    通过样本大小和潜在偏差来判断预测是否可靠。


2. Data Collection & Sampling | 数据收集与抽样

Primary data are collected directly by the researcher; secondary data are gathered from an existing source. You should be able to identify both and give advantages and disadvantages of each.

一手数据由研究者直接收集;二手数据来自已有来源。你需要能区分两者,并说明各自的优缺点。

Sampling methods include random, systematic, stratified, quota, and opportunity sampling. In a stratified sample, the number taken from each group is proportional to the group size.

抽样方法包括随机抽样、系统抽样、分层抽样、配额抽样和便利抽样。在分层抽样中,每组抽取的数量与该组在总体中所占比例相同。

Stratified sample size = (group size ÷ population size) × total sample size

For random sampling, use a random number generator or a random number table. For systematic sampling, select every nth item after a random start.

进行随机抽样时,可使用随机数生成器或随机数表。进行系统抽样时,在随机起点后每隔n个个体抽取一个。

  • Quota sampling: choose a fixed number of people with certain characteristics; it is non-random and faster.

    配额抽样:按特定特征选取固定人数;它非随机且速度更快。

  • Opportunity sampling: choose whoever is available; it is convenient but highly biased.

    便利抽样:选择方便获取的人;操作方便但偏差很大。


3. Types of Data | 数据类型

Quantitative data are numerical (discrete or continuous). Qualitative data are descriptive categories such as colour or brand.

定量数据是数值型数据(离散或连续)。定性数据是描述性类别,如颜色或品牌。

Discrete data can only take specific values, such as shoe size or number of siblings. Continuous data can take any value in a range, such as height or time.

离散数据只能取特定值,如鞋码或兄弟姐妹数量。连续数据可以取一定范围内的任意值,如身高或时间。

You may also need to convert grouped data into midpoints for estimating the mean, and understand how class boundaries are used in histograms.

你还需要会用组中值估计分组数据的平均数,并理解直方图中类界的使用方法。

  • Continuous data are often rounded; use class boundaries such as 10 ≤ x < 20.

    连续数据通常会被四舍五入;使用类似10 ≤ x < 20的类界。

  • Qualitative data cannot be averaged meaningfully.

    定性数据不能被有效求平均。


4. Data Representation | 数据表示

Different charts are suitable for different types of data. A bar chart is used for categorical data; a histograms is used for continuous data, often with unequal class widths.

不同类型的图表适用于不同类型的数据。条形图用于分类数据;直方图用于连续数据,通常各组宽度不等。

For a histogram with unequal class widths, the vertical axis is frequency density, calculated as frequency ÷ class width. The area of each bar represents the frequency.

对于组距不等的直方图,纵轴为频率密度,计算公式为:频率 ÷ 组距。每根条形的面积表示频率。

Frequency density = frequency ÷ class width

A stem-and-leaf diagram preserves the original data values and shows the shape of the distribution. A frequency polygon is drawn by plotting the midpoint of each class against its frequency and joining the points with straight lines.

茎叶图保留原始数据值并显示分布形状。频率多边形以各组中点对应频率描点,并用直线依次连接而成。

  • Pie charts show proportions; the angle for each category is (frequency ÷ total) × 360°.

    饼图显示比例;每个类别对应角度为(频率 ÷ 总数)× 360°。

  • Scatter graphs show the relationship between two variables.

    散点图显示两个变量之间的关系。


5. Measures of Central Tendency | 集中趋势度量

The mean, median, and mode are measures of central tendency. For ungrouped data, the mean is the sum of all values divided by the number of values.

平均数、中位数和众数是集中趋势的度量。对于未分组数据,平均数等于所有数值之和除以数值个数。

Mean = Σx ÷ n

For grouped data, estimate the mean using midpoints. The median is the middle value when data are ordered; the mode is the most frequent value.

对于分组数据,使用组中值来估计平均数。中位数是数据排序后的中间值;众数是出现次数最多的值。

When comparing distributions, always compare a measure of central tendency together with a measure of spread, not just one.

比较两组分布时,必须同时比较一个集中趋势量和一个离散程度量,不能只看其中一种。

  • If a data set has extreme values, the median is often more representative than the mean.

    如果数据存在极端值,中位数通常比平均数更具代表性。

  • Use the mode for qualitative data.

    对于定性数据使用众数。


6. Measures of Spread | 离散程度度量

The range, interquartile range, and standard deviation measure how spread out the data are. The range is the difference between the maximum and minimum values.

极差、四分位距和标准差衡量数据的离散程度。极差是最大值与最小值之差。

The lower quartile (Q1) is the median of the lower half; the upper quartile (Q3) is the median of the upper half. The interquartile range is Q3 − Q1.

下四分位数(Q1)是下半部分数据的中位数;上四分位数(Q3)是上半部分数据的中位数。四分位距为 Q3 − Q1。

Standard deviation measures the average distance of each value from the mean. For GCSE Statistics, you may use a calculator formula or the given formula.

标准差衡量每个数值与平均数的平均距离。在GCSE统计中,你可以使用计算器上的公式或题目给出的公式。

σ = √( Σ(x − x̄)² ÷ n )

  • A larger standard deviation means more variability.

    标准差越大表示变异性越大。

  • Use the interquartile range when the median is used, and the standard deviation when the mean is used.

    使用中位数时搭配四分位距;使用平均数时搭配标准差。


7. Cumulative Frequency & Box Plots | 累积频率与箱线图

A cumulative frequency table lists the running total of frequencies up to each class boundary. Plotting cumulative frequency against the upper boundaries gives a cumulative frequency graph.

累积频率表列出每个类界之前的频数累加值。以累积频率对每个类别的上界作图,得到累积频率图。

From this graph you can read the median (at the 50th percentile) and quartiles. The median is the value at 50% of the total frequency, Q1 at 25%, and Q3 at 75%.

从该图中可以读出中位数(第50百分位数)和四分位数。中位数对应总频数的50%,Q1对应25%,Q3对应75%。

A box plot displays the minimum, Q1, median, Q3, and maximum in a simple diagram. It is useful for comparing two distributions.

箱线图用简洁图示显示最小值、Q1、中位数、Q3和最大值,非常适合比较两组分布。

Percentiles divide the data into hundredths. For example, the 90th percentile is the value below which 90% of the data lie.

百分位数将数据分成一百等份。例如,第90百分位数表示90%的数据都小于该值。

  • When drawing a box plot, label the five-key summary clearly.

    绘制箱线图时,清楚标注五数概括。

  • Use the graph to estimate the number or percentage of data below a certain value.

    利用图表估计小于某值的数据数量或百分比。


8. Probability Basics | 概率基础

Probability measures the likelihood of an event occurring, always between 0 and 1. The sum of probabilities of all possible outcomes is 1.

概率衡量事件发生的可能性,取值总在0和1之间。所有可能结果的概率之和为1。

P(A) = number of favourable outcomes ÷ total number of outcomes

Complementary events: P(A’) = 1 − P(A). The complement of an event is the event not happening.

互补事件:P(A’) = 1 − P(A)。事件A的补事件表示A不发生。

Sample space diagrams show all possible outcomes for two events, helping you find probabilities systematically.

样本空间图列出两个事件所有可能结果,帮助系统计算概率。

  • For mutually exclusive events, P(A or B) = P(A) + P(B).

    对于互斥事件,P(A或B) = P(A) + P(B)。

  • If all outcomes are equally likely, use the formula above; otherwise use experimental frequency.

    如果所有结果等可能,用上述公式;否则使用实验频率。


9. Probability Rules | 概率法则

For independent events, the multiplication rule is P(A and B) = P(A) × P(B). Two events are independent if one occurring does not affect the probability of the other.

对于独立事件,乘法法则为 P(A且B) = P(A) × P(B)。若一个事件发生不影响另一个事件发生的概率,则两个事件独立。

For non-independent events, use conditional probability: P(A and B) = P(A) × P(B given A).

对于非独立事件,使用条件概率:P(A且B) = P(A) × P(在A条件下B发生)。

Tree diagrams are useful for multi-stage events. Each branch shows a probability, and the probability of a particular outcome is the product of probabilities along the branches.

树状图适用于多阶段事件。每条分支显示一个概率,某个结果的概率等于沿该路径各分支概率的乘积。

Venn diagrams show relationships between events. The intersection is A and B; the union is A or B. You should know how to enter frequencies or probabilities correctly.

维恩图显示事件之间的关系。交集是A且B;并集是A或B。你需要知道如何正确填入频数或概率。

  • Use the addition rule for non-mutually exclusive events: P(A or B) = P(A) + P(B) − P(A and B).

    对于非互斥事件使用加法法则:P(A或B) = P(A) + P(B) − P(A且B)。

  • Check all probabilities in a tree diagram sum to 1 at each set of branches.

    检查树状图中每组分支的概率之和为1。


10. Index Numbers | 指数

An index number compares the current value to a base value. The base is usually set to 100.

指数将当前值与基期值进行比较。基期通常设为100。

Index number = (current value ÷ base value) × 100

For example, if a price was £50 in 2015 and £60 in 2020, the index is (60 ÷ 50) × 100 = 120, meaning a 20% increase from the base year.

例如,某价格2015年为50英镑,2020年为60英镑,则指数为 (60 ÷ 50) × 100 = 120,表示比基期上涨20%。

A weighted index accounts for the importance of different items. Multiply each index by its weight and divide by the sum of weights.

加权指数考虑不同项目的重要性。将每个指数乘以其权重,再除以权重之和。

Weighted index = Σ(index × weight) ÷ Σweight

  • Index numbers are often used for prices and wages over time.

    指数常用于衡量物价和工资随时间的变化。

  • Be careful to interpret an index of 100 as no change from the base year.

    注意指数为100表示与基期相比没有变化。


11. Time Series & Moving Averages | 时间序列与移动平均

A time series is a set of data recorded at regular intervals over time. You should plot the data on a line graph and identify trends and seasonal variations.

时间序列是按固定时间间隔记录的一组数据。你需要将数据绘制成折线图,并识别趋势和季节性变化。

A moving average smooths out short-term fluctuations to show the underlying trend. For a 4-point moving average, add the first four data points, then the next four, etc.

移动平均消除短期波动,以显示潜在趋势。对于4点移动平均,先加前四个数据点,再依次向后移动四个数据点。

Moving average = (sum of a consecutive set of values) ÷ number of values

The seasonal effect is found by comparing the actual value to the moving average for that period. A positive difference indicates above-average seasonal effect.

季节性影响通过将实际值与同期的移动平均值进行比较得出。正差异表示季节性影响高于平均水平。

  • When the number of points is even, you may need to centre the moving average by averaging pairs.

    当点数为偶数时,可能需要通过再平均相邻两个移动平均值来得到居中移动平均。

  • Extrapolation beyond the given range can be unreliable; always state this limitation.

    超出给定范围的预测可能不可靠;务必说明这一局限。


12. Quality Control Charts | 质量控制图

Quality control charts are used in manufacturing to monitor whether a process is in control. You draw the target/mean value and warning and action limits.

质量控制图用于生产制造中监控过程是否处于受控状态。你需要绘制目标值/平均值以及警戒限和行动限。

Typically, warning limits are set at mean ± 2 standard deviations, and action limits at mean ± 3 standard deviations.

通常,警戒限设为平均数 ± 2个标准差,行动限设为平均数 ± 3个标准差。

Warning limit: x̄ ± 2σ

Action limit: x̄ ± 3σ

A process is considered out of control if a point lies beyond an action limit, or if there is a run of several points all above or below the mean. In these cases, remedial action is needed.

如果一点超过行动限,或者连续多个点都高于或低于平均值,则过程被认为失控,需要采取纠正措施。

  • One point beyond a warning limit warns of potential trouble; two consecutive points beyond the warning limit often indicate out-of-control.

    一个点超出警戒限警示可能存在问题;连续两个点超出警戒限通常表明失控。

  • Always label the limits clearly on your chart and state whether the process is in control.

    在图上清楚标注各界限,并判断过程是否受控。


Published by TutorHao | Statistics Revision Series | aleveler.com

更多咨询请联系16621398022(同微信)

Comments

屏轩国际教育cambridge primary/secondary checkpoint, cat4, ukiset,ukcat,igcse,alevel,PAT,STEP,MAT, ibdp,ap,ssat,sat,sat2课程辅导,国外大学本科硕士研究生博士课程论文辅导Cancel reply

This site uses Akismet to reduce spam. Learn how your comment data is processed.

Discover more from aleveler.com

Subscribe now to keep reading and get access to the full archive.

Continue reading

Exit mobile version