📚 GCSE WJEC Statistics: High-Frequency Topics and Common Mistake Analysis | GCSE WJEC 统计:高频考点与易错题分析
This revision guide breaks down the most frequently examined topics in GCSE WJEC Statistics and pinpoints the mistakes that cost students marks year after year. From drawing histograms with unequal class widths to interpreting cumulative frequency curves and avoiding probability tree blunders, each section pairs a clear English explanation with its Chinese equivalent so you can master both the concepts and the terminology.
本备考指南拆解 GCSE WJEC 统计学中最常考的主题,并指出年复一年让学生丢分的错误。从绘制不等组距的直方图到解读累积频率曲线,再到避免概率树图的失误,每个部分都提供了清晰的英文解释与对应的中文表达,帮助你同时掌握概念和术语。
1. Data Types and Question Design | 数据类型与问卷设计
WJEC papers frequently ask you to classify data as qualitative or quantitative, then further as discrete or continuous. Discrete data can only take certain values (e.g. number of siblings), while continuous data can take any value in a range (e.g. height). Getting this right determines which chart or average is appropriate.
WJEC 试题经常要求将数据分为定性数据或定量数据,再进一步划分为离散型或连续型。离散数据只能取特定值(如兄弟姐妹数),而连续数据可以取某个范围内的任何值(如身高)。正确分类决定了该选用哪种图表或平均数。
A classic pitfall is confusing ordinal data with continuous data. Rating scales (1 to 5 stars) produce ordinal data, so calculating a mean is questionable, yet many candidates do exactly that.
一个经典的陷阱是将有序数据与连续数据混淆。评分量表(1 到 5 星)产生的是有序数据,因此计算平均值是有问题的,然而许多考生恰恰这样做了。
When designing questionnaires, candidates lose marks by using overlapping response boxes, leading questions, or missing a time frame. Always include units, avoid ‘How often do you…?’ without a reference period, and ensure response options are exhaustive.
在设计问卷时,考生常因响应框重叠、问题具有导向性或缺少时间范围而失分。务必包含单位,避免没有参考时间段的“你多久……?”,并确保响应选项详尽无遗。
2. Sampling Methods and Capture-Recapture | 抽样方法与捕获‑再捕获
You must be able to describe and evaluate simple random, systematic, stratified, quota, and cluster sampling. Stratified sampling is a high-frequency topic: the sample size for a stratum is calculated as (stratum size ÷ population size) × total sample size.
你必须能够描述并评估简单随机抽样、系统抽样、分层抽样、配额抽样和整群抽样。分层抽样是高频考点:某一层的样本量计算公式为 (层大小 ÷ 总体大小) × 总样本量。
The most common mistake in stratified sampling questions is failing to round sensibly or forgetting that each stratum’s sample size must add to the total. Many candidates also confuse quota sampling with stratified sampling because both look at subgroups, but quota sampling does not involve random selection.
分层抽样题中最常见的错误是未能合理取整,或忘记各层样本量之和必须等于总样本量。许多考生还因两者都涉及子组而将配额抽样与分层抽样混淆,但配额抽样不采用随机抽取。
For capture-recapture (the Petersen method), the formula is N = (M × C) ÷ R, where M is the number marked and released, C is the total caught the second time, and R is the number recaptured. The key assumption is that the marked individuals mix evenly with the unmarked population.
对于捕获‑再捕获法(彼得森法),公式为 N = (M × C) ÷ R,其中 M 是标记并释放的数量,C 是第二次捕获的总数,R 是重捕数量。关键假设是标记个体与未标记个体均匀混合。
3. Bar Charts, Pie Charts and Pictograms | 条形图、饼图与象形图
Bar charts for discrete data must have gaps; frequency diagrams for continuous data should be drawn without gaps. Pie charts require angle = (frequency ÷ total) × 360°. Pictograms need a clear key and correct proportional areas when symbols represent more than one unit.
离散数据的条形图必须有间隙;连续数据的频率图则应无间隙绘制。饼图需要角度 = (频率 ÷ 总计) × 360°。象形图需要一个清晰的图例,且当符号代表多于一个单位时,面积须成正确比例。
A frequent error in pie charts is using sector areas rather than angles, causing distorted representations. For pictograms, drawing partial symbols inaccurately or failing to give the key costs marks.
饼图中常见的错误是使用扇形面积而非角度,导致表现失真。对于象形图,不准确地画部分符号或未提供图例都会失分。
Misleading graphs regularly appear in exam criticism questions. Look out for truncated scales, inconsistent bin widths, and 3D effects that exaggerate differences.
误导性图表经常出现在考试中的批评性问题里。注意截断的坐标轴刻度、不一致的组距宽度,以及夸大差异的三维效果。
4. Histograms and Frequency Density | 直方图与频率密度
A histogram plots frequency density against the variable. Frequency density = frequency ÷ class width. If all class widths are equal, a histrogram reduces to a frequency diagram; when widths are unequal, you must use frequency density.
直方图以频率密度作为纵轴绘图。频率密度 = 频率 ÷ 组距宽度。如果所有组距宽度相等,直方图就退化为频率图;当宽度不等时,必须使用频率密度。
The single most penalised error is forgetting to work out frequency density. Students often plot raw frequencies on the vertical axis and produce a distorted shape that loses marks for both the graph and any subsequent estimates of mode or median.
扣分最多的一个错误是忘记计算频率密度。学生往往在纵轴上直接标出原始频率,得出扭曲的图形,导致图表本身以及随后对众数或中位数的估计都失分。
Always label axes clearly: ‘Frequency density’ on the y-axis and the variable name on the x-axis. Scales must be uniform and the area of each bar is proportional to frequency.
务必清晰标注坐标轴:纵轴标“频率密度”,横轴标变量名。刻度必须均匀,且每个直条的面积与频率成正比。
5. Cumulative Frequency and Box Plots | 累积频率与箱线图
A cumulative frequency curve is plotted at the upper class boundary. From it you can read off the median, lower quartile (Q₁) and upper quartile (Q₃). The interquartile range (IQR) = Q₃ – Q₁.
累积频率曲线在组距的上限处描点。你可以从中读取中位数、下四分位数(Q₁)和上四分位数(Q₃)。四分位距(IQR)= Q₃ – Q₁。
A very common slip is taking the median at ½n instead of ½(n+1) when using a cumulative frequency table. With a graph, the convention is to read at ½n, ¼n, and ¾n for median and quartiles. Check your specification: WJEC typically uses ½n for grouped data curves, but the safest approach is to draw lines at those fractions of the total frequency.
一个很常见的失误是使用累积频率表时将中位数取在 ½n 而非 ½(n+1) 处。对于图形,惯例是在 ½n、¼n 和 ¾n 处读取中位数和四分位数。请核对你的课程大纲:WJEC 对于分组数据曲线通常使用 ½n,但最稳妥的做法是在总频率的相应分数处画线。
For box plots, remember the whisker endpoints are the minimum and maximum within 1.5 × IQR of the quartiles; any point beyond is an outlier and should be marked with a cross. Many candidates extend whiskers to extreme values without checking for outliers, losing accuracy marks.
绘制箱线图时,须记住须线的端点是在四分位数 1.5 × IQR 范围内的最小值和最大值;任何超出该范围的点都是异常值,应用叉号标出。许多考生未检查异常值便将须线延伸至极值,从而失掉准确性分数。
6. Measures of Average and Spread | 平均数与离散度的度量
You need to choose between mean, median, and mode, and between range, IQR, and standard deviation. The mean is sensitive to extreme values, so when data is skewed the median is often a better measure of centrality. For spread, IQR is robust, while standard deviation uses all values.
你需要在平均值、中位数和众数之间,以及在极差、IQR 和标准差之间进行选择。平均值对极端值敏感,因此当数据偏斜时,中位数往往是更好的集中量数。对于离散度,IQR 是稳健的,而标准差使用了所有的数值。
A typical WJEC question asks, ‘Which average is most suitable? Give a reason.’ Many candidates simply say ‘the mean because it uses all the data’ without considering outliers. If the data contains a clear outlier, the median is more suitable, and you must state that it is unaffected by extreme values.
典型的 WJEC 题目会问:“哪种平均数最合适?请说明理由。”许多考生只回答“平均值,因为它使用了所有数据”,却不考虑异常值。如果数据包含明显的异常值,中位数更合适,并且你必须说明它不受极端值影响。
When calculating standard deviation for a sample, use the formula s = √[Σ(x – x̄)² ÷ (n – 1)]. For a population, divide by n. Candidates often mix up n and n – 1, especially on calculator tests.
计算样本标准差时,使用公式 s = √[Σ(x – x̄)² ÷ (n – 1)]。对于总体,则除以 n。考生常常混淆 n 和 n – 1,尤其是在计算器考试中。
7. Scatter Diagrams and Correlation | 散点图与相关性
Plotting a scatter diagram and drawing a line of best fit by eye is a staple skill. Correlation can be described as positive, negative or none, and further qualified as strong, moderate or weak. You must be able to interpret correlation in context.
绘制散点图并凭目测画出最佳拟合线是一项基本技能。相关性可描述为正相关、负相关或无相关,并可进一步限定为强、中或弱。你必须能够在上下文情境中解读相关性。
The most serious misconception is claiming causation from correlation. WJEC examiners frequently embed a question like ‘Does a strong positive correlation between ice cream sales and drowning incidents mean ice cream causes drowning?’ Students must explain that a third factor (hot weather) influences both.
最严重的误解是从相关性推断出因果关系。WJEC 考官经常嵌入类似“冰淇淋销量与溺水事件之间的强正相关是否意味着冰淇淋导致溺水?”的问题。学生必须解释第三个因素(炎热天气)同时影响着两者。
When drawing the line of best fit, ensure it passes through the mean point (x̄, ȳ) and balances points above and below. Avoid forcing the line through the origin unless the context justifies it.
绘制最佳拟合线时,确保它通过均值点(x̄, ȳ)并平衡线上下的点。除非上下文有合理理由,否则应避免强制直线通过原点。
8. Probability and Tree Diagrams | 概率与树状图
GCSE WJEC probability covers the addition rule (for OR, when events are mutually exclusive, P(A or B) = P(A) + P(B)), and the multiplication rule (for AND, when independent, P(A and B) = P(A) × P(B)). Tree diagrams are examined almost every series.
GCSE WJEC 概率部分涵盖加法法则(当事件互斥时,P(A 或 B) = P(A) + P(B)),以及乘法法则(当事件独立时,P(A 且 B) = P(A) × P(B))。树状图几乎每个考季都会出现。
Mistakes on tree diagrams include forgetting to label branches with probabilities, or not checking that probabilities on each pair of branches sum to 1. When events are not independent, students often misapply the multiplication rule by failing to adjust the conditional probabilities on the second set of branches.
树状图的错误包括忘记在分支上标注概率,或未检查每对分支的概率之和是否为 1。当事件不独立时,学生常因未能在第二组分枝上调整条件概率而误用乘法法则。
Conditional probability frequently trips students up. Remember P(A|B) = P(A and B) ÷ P(B). A two-way table can simplify problems that look complex as tree diagrams.
条件概率经常让学生栽跟头。记住 P(A|B) = P(A 且 B) ÷ P(B)。对于看似复杂的树状图问题,双向表可以简化求解。
9. Time Series and Moving Averages | 时间序列与移动平均
A time series graph plots data over time. To smooth out fluctuations and identify the trend, you calculate moving averages. For a 4-point moving average, plot the average against the midpoint of the four time periods.
时间序列图是将数据按时间绘制。为了消除波动并识别趋势,需要计算移动平均。对于 4 点移动平均,将平均值绘制在四个时间段的中点位置上。
A common error is plotting the moving average at the last time point instead of the midpoint. This shifts the trend line and leads to inaccurate predictions. Also, when calculating seasonal variations, remember that the sum of all variations should be zero (or very close).
常见的错误是将移动平均绘制在最后一个时间点上,而不是中点。这会使趋势线发生偏移并导致预测不准确。此外,在计算季节变动时,切记所有变动之和应为零(或非常接近零)。
On WJEC papers, you may be asked to make a prediction using the trend line, then adjust it by a seasonal effect. Always state clearly which seasonal adjustment you are adding or subtracting.
在 WJEC 的试卷中,你可能会被要求使用趋势线做出预测,然后通过季节效应进行调整。始终清楚地说明你是加上还是减去哪个季节调整量。
10. Index Numbers | 指数
Index numbers compare the price or quantity of an item in a given period with its value in a base period. The simple price relative formula is (current price ÷ base price) × 100. Weighted aggregate indices like Laspeyres and Paasche are also tested.
指数用于将某一时期商品的价格或数量与其在基期的值进行比较。简单价比的公式为(现价 ÷ 基价) × 100。考试也会涉及加权综合指数,如拉氏指数和帕氏指数。
Weighted index errors usually stem from using the wrong base year quantities or failing to multiply weights correctly. When asked to interpret an index value, don’t just state ‘it increased’; say ‘prices rose by 15% compared with the base year’.
加权指数的错误通常源于使用错误的基期数量,或未能正确地乘以权重。当被要求解释一个指数值时,不要只说“它上升了”;而要说“与基年相比,价格上涨了 15%”。
Ensure you can change the base year and chain-link indices. A question might give index values with different base years and ask you to splice them together.
确保你能够更换基年并将指数链接起来。题目可能会给出不同基年的指数值,并要求你将它们拼接在一起。
11. Quality Assurance and Control Charts | 质量保证与控制图
WJEC includes basic statistical process control, often using mean and range charts with warning limits (±2 standard errors) and action limits (±3 standard errors). You must state whether a process is ‘in control’ or identify points beyond action limits.
WJEC 涵盖基本的统计过程控制,常使用带有警戒限(±2 标准误)和行动限(±3 标准误)的均值图和极差图。你必须说明过程是否“受控”,或识别出超出行动限的点。
A recurrent error is misidentifying a run of points on one side of the target as random variation. A run of 7 or more points above or below the centre line is a signal that the process may be going out of control, even if none exceed the action limits.
一个反复出现的错误是将目标线一侧的一连串点误认为是随机变异。即使没有点超出行动限,如果连续 7 个或更多点出现在中心线的上方或下方,也标志着过程可能正在失控。
12. Exam Craft: Common Slips and Top Marks | 考场技艺:常见笔误与高分诀窍
Lost marks in WJEC Statistics often come from not showing working, omitting units, or giving answers to an inappropriate degree of accuracy. Always round measures sensibly: money to 2 decimal places, probabilities to 2 or 3 significant figures, and averages to 1 decimal place more than the original data unless instructed otherwise.
WJEC 统计学中丢分常来自未展示计算过程、遗漏单位或给出不当精确度的答案。务必合理取整:金额保留两位小数,概率保留 2 或 3 位有效数字,平均数比原始数据多一位小数,除非另有说明。
When comparing distributions using box plots or summary statistics, you must make a comparative statement using both a measure of central tendency (e.g. ‘the median is higher for A’) and a measure of spread (‘the IQR is smaller for A, so it is more consistent’). Just copying numbers from the graph earns no credit.
使用箱线图或汇总统计量比较分布时,必须同时使用集中量数(如“A 的中位数较高”)和离散量数(“A 的 IQR 较小,因此更稳定”)进行比较说明。仅从图中抄下数字不会得分。
A final piece of advice: before tackling a probability question, write ‘Assume events are independent’ or note where conditional probabilities change. This simple habit prevents the most common deduction-happy error on the paper.
最后一条建议:在解答概率题之前,写下“假设事件相互独立”或注明条件概率在何处发生变化。这个简单的习惯可以避免试卷上最常见的扣分点。
Published by TutorHao | Statistics Revision Series | aleveler.com
更多咨询请联系16621398022(同微信)
屏轩国际教育cambridge primary/secondary checkpoint, cat4, ukiset,ukcat,igcse,alevel,PAT,STEP,MAT, ibdp,ap,ssat,sat,sat2课程辅导,国外大学本科硕士研究生博士课程论文辅导