📚 Year 8 CIE Statistics: Quick Vocabulary Memorisation Guide | 8年级 CIE 统计:词汇术语速记指南
Mastering statistical vocabulary is the first step to confidence in data handling. This guide breaks down essential terms for Year 8 CIE Statistics into easy-to-remember groups with paired explanations, memory hooks and visual associations.
掌握统计学术语是自信处理数据的第一步。这份指南将8年级 CIE 统计学核心术语拆解为易记的词组,配有配对解释、记忆口诀和视觉联想,帮助你快速形成长期记忆。
1. What is Statistics? | 什么是统计学?
‘Statistics’ is the science of collecting, organising, presenting, analysing and interpreting numerical data. The word comes from ‘state’ because rulers once used numbers to manage countries.
“统计学”是收集、整理、展示、分析和解读数字数据的科学。该词源自“国家”,因为古代统治者曾用数字来管理国家。
‘Data’ simply means pieces of information. Data can be numbers, words or observations. ‘Raw data’ is the unprocessed, first-hand information before any sorting or cleaning.
“数据”简单来说就是信息片段。数据可以是数字、文字或观察结果。“原始数据”指未经任何分类或整理前的一手信息。
2. Data Types: Qualitative vs. Quantitative | 数据类型:定性数据与定量数据
‘Qualitative data’ describes qualities or categories that cannot be measured with numbers in a meaningful way, such as eye colour, favourite food or type of pet. Think ‘quality’.
“定性数据”描述无法用数字有意义地测量的属性或类别,例如眼睛颜色、最爱的食物或宠物类型。想想“性质”这个词。
‘Quantitative data’ deals with quantities – things you can count or measure. Height in cm, mass in kg, number of students in a class. It is further split into ‘discrete’ and ‘continuous’.
“定量数据”处理的是数量——你可以计数或测量的对象。比如以厘米为单位的身高、以千克为单位的质量、班级学生人数。它进一步分为“离散”和“连续”。
‘Discrete data’ can only take specific, separate values – often whole numbers. Examples: shoe sizes, number of books, goals scored. You cannot have 3.2 goals in a football match.
“离散数据”只能取特定的、独立的值——通常是整数。例如:鞋码、书的数量、进球数。一场足球赛不可能有3.2个进球。
‘Continuous data’ can take any value within a range. Measurements like length, time and temperature are continuous because they can be fractions. A race time could be 10.45 seconds.
“连续数据”可以取某个范围内的任意值。长度、时间和温度等测量结果是连续的,因为它们可以是小数。比赛时间可能是10.45秒。
- Memory hook: Discrete = dots on a number line; Continuous = smooth line with no gaps.
- 记忆口诀: 离散=数轴上的点;连续=无缝隙的平滑线。
3. Measures of Central Tendency: The Three Averages | 集中趋势度量:三种平均数
The ‘mean’ is what most people call the average. Add all values and divide by the number of values. Formula: Mean = Σx / n. The Σx means ‘sum of all data values’.
“均值”就是多数人所说的平均数。将所有数值相加,再除以数值个数。公式:均值 = Σx / n。Σx 意为“所有数据值的总和”。
The ‘median’ is the middle value when data is ordered from smallest to largest. For an even number of items, find the mean of the two middle numbers. The median splits data into two equal halves.
“中位数”是将数据从小到大排序后的中间值。如果有偶数个数,则取中间两个数的均值。中位数将数据均分为两半。
The ‘mode’ is the value that appears most often. A data set can have one mode, more than one mode (bimodal or multimodal) or no mode at all if all values occur equally.
“众数”是出现次数最多的值。数据集可以有一个众数、多个众数(双众数或多众数),或者当所有值出现次数相同时没有众数。
| Term / 术语 | Memory trick / 记忆技巧 |
| Mean / 均值 | ‘Mean teacher makes you add everything and divide.’ / “严厉的老师让你全部加起来再除。” |
| Median / 中位数 | ‘Median’ sounds like ‘middle’ — think of the median strip on a road. / ‘Median’ 音似 ‘middle’——想想马路中央的分隔带。 |
| Mode / 众数 | ‘Mode’ is ‘most’. The fashion mode is the most popular style. / ‘Mode’ 就是 ‘most’(最多)。流行时尚就是最受欢迎的款式。 |
4. Measures of Spread: Range and Variability | 离散程度度量:极差与变异性
The ‘range’ tells you how spread out the data is. It is the difference between the largest value and the smallest value. Range = max – min. A small range means data points are close together.
“极差”告诉你数据分散的程度。它是最大值与最小值之间的差。极差 = 最大值 – 最小值。极差小意味着数据点聚集紧密。
‘Spread’ or ‘variability’ is a general term describing how data deviates from the centre. Two classes could have the same mean score but very different spread – meaning one class has much more mixed results.
“离散程度”或“变异性”是描述数据偏离中心程度的通用术语。两个班级可能有相同的平均分,但离散程度大不相同——这表明一个班级学生成绩参差不齐。
‘Outliers’ are extreme values that lie well away from the rest of the data. They can dramatically affect the mean but do not influence the median. Always check for outliers before choosing your average.
“异常值”是远离其他数据的极端值。它们会极大地影响均值,但不影响中位数。选择平均数前务必检查是否存在异常值。
5. Frequency and Frequency Tables | 频数与频数表
‘Frequency’ is simply the number of times an event or data value occurs. If you asked 20 friends their favourite fruit, the frequency of ‘apple’ might be 8. The total frequency is always the sum of all frequencies.
“频数”就是某个事件或数据值出现的次数。如果你问20位朋友最喜欢的水果,“苹果”的频数可能是8。总频数始终是所有频数的总和。
A ‘frequency table’ organises data into a table showing each category or group alongside its frequency. Columns typically include ‘Tally’ (using strokes to count) and ‘Frequency’.
“频数表”将数据整理到表格中,显示每个类别或组别及其频数。列通常包括“计数”(用画“正”字方式)和“频数”。
‘Grouped frequency table’ is used when data is continuous or has many distinct values. We group data into class intervals, e.g., 0–9, 10–19, 20–29. Be careful with the notation 0 ≤ x < 10.
“分组频数表”用于连续数据或有大量不重复值的情况。我们将数据分组到区间内,如 0–9、10–19、20–29。留意使用 0 ≤ x < 10 这样的区间记号。
6. Charts and Graphs: Visual Vocabulary | 图表:视觉化术语
A ‘bar chart’ uses rectangular bars to represent categorical data. The height or length of each bar equals the frequency. Bars do not touch; this highlights that categories are separate.
“条形图”用矩形条表示类别数据。每个条的高度或长度等于频数。条与条之间不接触,以突出类别是独立的。
A ‘pie chart’ is a circle divided into sectors. Each sector’s angle is proportional to the frequency it represents. Angle = (frequency ÷ total frequency) × 360°.
“饼图”是被划分成扇形的圆形。每个扇形的角度与其表示的频数成正比。角度 = (频数 ÷ 总频数) × 360°。
A ‘line graph’ uses points connected by straight lines, often to show changes over time. The horizontal axis (x-axis) is usually time, and the vertical axis (y-axis) is the measurement.
“折线图”用点并通过直线连接,通常用于显示随时间变化的趋势。横轴(x轴)通常是时间,纵轴(y轴)是测量值。
A ‘pictogram’ uses small images or symbols to represent data. Each picture could stand for 1, 2, 5 or 10 items. Always include a key explaining what one symbol means.
“象形图”用小图像或符号来表示数据。每个图案可以代表1、2、5或10个项目。务必附上图例说明一个符号的含义。
‘Scatter graph’ (or scatter plot) shows pairs of related data as points on a grid. It helps us see if there is a relationship (correlation) between two variables.
“散点图”将成对的相关数据以点的形式绘制在网格上。它帮助我们观察两个变量之间是否存在关系(相关性)。
7. Probability Language: From Impossible to Certain | 概率用语:从不可能到必然
‘Probability’ is the chance that something will happen. It is always a number between 0 and 1. 0 means impossible, 1 means certain. Probability = (number of favourable outcomes) / (total number of possible outcomes).
“概率”是某事件发生的可能性大小。它总是一个介于0到1之间的数字。0表示不可能,1表示必然。概率 = (有利结果数) / (所有可能结果数)。
‘Experiment’ is any repeatable process that gives rise to an outcome, e.g., rolling a die, flipping a coin, drawing a card. A ‘trial’ is one single performance of the experiment.
“试验”是任何产生结果的、可重复的过程,如掷骰子、抛硬币、抽牌。“一次尝试”指试验的单次执行。
‘Outcome’ is the result of a single trial, such as getting a ‘3’ when rolling a die. The ‘sample space’ is the set of all possible outcomes, often shown as a list, table or tree diagram.
“结果”是单次尝试的结果,例如掷骰子得到“3”。“样本空间”是所有可能结果的集合,通常以列表、表格或树状图展示。
An ‘event’ is a set of one or more outcomes we are interested in. For example, ‘rolling an even number’ on a die is an event containing outcomes {2, 4, 6}. Events can be ‘equally likely’, ‘biased’ or ‘random’.
“事件”是我们感兴趣的一个或多个结果的集合。例如,骰子掷出“偶数”就是一个包含结果{2,4,6}的事件。事件可能是“等可能的”、“有偏的”或“随机的”。
- 0 probability = impossible / 0概率 = 不可能
- 0.5 probability = even chance / 0.5概率 = 均等机会
- 1 probability = certain / 1概率 = 必然
8. Sampling: Population and Sample | 抽样:总体与样本
The ‘population’ is the entire group we want to find out about. It could be all students in a school, all cars in a city or all trees in a forest. We rarely test the whole population.
“总体”是我们想要了解的全部对象。它可能是全校学生、全城汽车或森林中的所有树木。我们很少对整个总体进行调查。
A ‘sample’ is a smaller, manageable selection taken from the population. The goal is for the sample to be ‘representative’ so that conclusions can be generalised back to the whole population.
“样本”是从总体中选出的、规模较小且便于处理的一部分。目标是样本具有“代表性”,这样结论才能推广回整个总体。
‘Random sampling’ means every member of the population has an equal chance of being chosen. This avoids ‘bias’, which is when the sample is not fair and favours certain groups.
“随机抽样”意味着总体中的每个成员都有均等的机会被选中。这避免了“偏差”,即样本不公平、偏向某些群体的情况。
9. Correlation: Seeing Relationships in Scatter Graphs | 相关性:从散点图中观察关系
‘Correlation’ describes the strength and direction of a relationship between two variables. If both increase together, we call it ‘positive correlation’. If one increases while the other decreases, it is ‘negative correlation’.
“相关性”描述了两个变量之间关系的强度和方向。如果两者同时上升,称为“正相关”。如果一个上升而另一个下降,则是“负相关”。
‘No correlation’ means the points on the scatter graph are scattered randomly with no clear pattern. Changes in one variable do not predict changes in the other.
“无相关”意味着散点图上的点随机分布,没有明显模式。一个变量的变化无法预测另一个变量的变化。
‘Line of best fit’ is a straight line drawn through the centre of the data points on a scatter graph. It helps to estimate unknown values. The line does not have to pass through all points.
“最佳拟合线”是通过散点图数据中心的一条直线。它有助于估算未知数值。该线不必经过所有点。
‘Outliers’ on a scatter graph are points that lie far from the general pattern. They should be investigated; sometimes they occur due to measurement error.
散点图上的“异常值”是远离整体模式的点。应该对这些点进行调查;有时它们是由于测量误差产生的。
10. Quickfire Glossary: Must-know Terms | 速记词汇表:必会术语
Here is a last-chance check of terms you need to recognise instantly in an exam.
以下是考前你必须能立即识别的术语清单。
| Term / 术语 | Definition / 定义 |
| Hypothesis / 假设 | A testable statement about what you expect to find in an investigation. / 关于你在调查中预期发现的可验证陈述。 |
| Variable / 变量 | A characteristic or quantity that can change or take different values. / 可以变化或取不同值的特征或数量。 |
| Primary data / 一手数据 | Data you collect yourself for a specific purpose. / 为你自己的特定目的而收集的数据。 |
| Secondary data / 二手数据 | Data collected by someone else that you then use. / 由他人收集、你接着使用的数据。 |
| Census / 普查 | A survey that collects data from the entire population. / 从整个总体收集数据的调查。 |
| Tally chart / 计数表 | A simple way to record frequency using strokes, crossed in groups of five. / 用画“正”字记录频数的简单方法,每五划一束。 |
11. Memory Strategies That Work | 有效的记忆策略
Use ‘word families’: link mean, median and mode together as the three ‘Ms’ of averages. Associate ‘range’ with ‘max – min’ like a number line showing spread.
使用“词汇家族”:将 mean、median 和 mode 归为平均数的三个“M”。将“range 极差”与“max – min”关联,好比数轴上显示的跨度。
Draw a ‘statistics scene’ in your mind: picture a school experiment. The population is the whole playground; the sample is a small group wearing wristbands. The mean is everyone standing at a balance point; the median is the child in the middle of a sorted queue.
在脑海中绘制一幅“统计场景”:想象一个校园实验。总体是整个操场上的学生;样本是戴着腕带的一小组人。均值是所有人站在平衡点上;中位数是排好队后站在正中间的孩子。
Create bilingual flashcards with the English term on one side and the Chinese definition plus a visual icon on the other. Build a glossary bookmark to keep inside your exercise book.
制作双语闪卡,一面写英文术语,另一面写中文定义并配上一个视觉图标。制作一个词汇书签,夹在练习册里随时翻阅。
12. Final Checklist Before the Exam | 考前最终自查清单
- Can I tell the difference between discrete and continuous data? / 我能区分离散数据与连续数据吗?
- Do I know when to use mean, median or mode? / 我知道何时该用均值、中位数或众数吗?
- Can I read and construct a frequency table? / 我能读懂和制作频数表吗?
- Can I calculate probability as a fraction, decimal or percentage? / 我能将概率表示为分数、小数或百分数吗?
- Can I describe correlation from a scatter graph? / 我能根据散点图描述相关性吗?
Published by TutorHao | Statistics Revision Series | aleveler.com
更多咨询请联系16621398022(同微信)
屏轩国际教育cambridge primary/secondary checkpoint, cat4, ukiset,ukcat,igcse,alevel,PAT,STEP,MAT, ibdp,ap,ssat,sat,sat2课程辅导,国外大学本科硕士研究生博士课程论文辅导