📚 KS3 WJEC Statistics: Vocabulary & Terminology Quick Guide | KS3 WJEC 统计:词汇术语速记指南
Mastering statistics starts with understanding the language of data. In KS3 WJEC Statistics, you will encounter a wide range of terms — from ‘discrete data’ to ‘interquartile range’ — that form the foundation for data handling, probability, and problem-solving. This guide breaks down the essential vocabulary you need to know, with twin-language explanations to help you learn, remember, and apply each term confidently. Every term is paired with its Chinese equivalent and a clear definition, so you can build your statistical fluency in both languages.
掌握统计学的关键在于理解数据的语言。在 KS3 WJEC 统计学中,你会遇到大量术语——从 “离散数据” 到 “四分位距”——它们是数据处理、概率和问题解决的基础。这本指南将分解你需掌握的核心词汇,通过双语解释帮助你轻松学习、牢固记忆并自信运用每个术语。每个术语都配有中文对应和清晰的定义,让你能够在两种语言中建立统计流利度。
1. Types of Data | 数据类型
Understanding the type of data you are working with is the first step in statistics. Data can be divided into two main families: qualitative data and quantitative data. Qualitative data describes qualities or categories; quantitative data deals with numbers you can measure or count.
明确你处理的数据类型是统计学的第一步。数据可以分为两大族:定性数据和定量数据。定性数据描述属性或类别;定量数据涉及你可以测量或计数的数字。
Qualitative data (also called categorical data) groups information into non-numerical categories. Examples include favourite colour, type of transport to school, or yes/no answers to a survey question. Because it describes a ‘quality’, we often use bar charts or pie charts to display it, and we summarise it with frequencies or modes — never a mean.
定性数据(又称分类数据)将信息归入非数值类别。例如最喜欢的颜色、上学交通方式或问卷中的是/否答案。由于它描述 “性质”,我们通常用条形图或饼图来展示,并用频数或众数来概括——绝不使用平均数。
Quantitative data is numerical. It splits into two crucial subtypes: discrete data and continuous data. Recognising the difference is essential for choosing the right graph and avoiding common mistakes.
定量数据是数字型的,分为两个关键子类型:离散数据和连续数据。识别二者的差异对于选择正确的图表和避免常见错误至关重要。
Discrete data can only take specific, separate values — almost always whole numbers that you get by counting. Examples: the number of books in a bag, the score on a dice, or the number of students late for class. You cannot have 3.7 books or 4.2 students in the context of counting.
离散数据只能取特定的、分离的数值——几乎总是通过计数得到的整数。例如:书包里的书本数量、骰子点数或迟到学生人数。在计数的语境下,你不可能有 3.7 本书或 4.2 个学生。
Continuous data can take any value within a given range and is obtained by measuring. Height, mass, temperature, time, and length are continuous because they can include fractions and decimals. Even if you record height to the nearest centimetre, the underlying quantity is continuous.
连续数据可以在给定范围内取任意值,是通过测量获得的。身高、质量、温度、时间和长度都是连续的,因为它们可以包含分数和小数。即使你把身高记录到最接近的厘米,其底层量仍是连续的。
Quick memory aid: ‘discrete’ sounds like ‘discrete’ (separate) numbers you can count; ‘continuous’ flows like a continuous measurement scale.
快速记忆: “discrete” 发音像 “discrete”(分离的)数字,可以计数;”continuous” 像连续不断的测量尺度一样流动。
2. Collecting Data | 数据收集
Primary data is information you collect yourself, first-hand, for a specific purpose. For example, conducting a survey of your classmates’ screen time, measuring plant growth in a science experiment, or timing how long it takes to run 100 metres. Primary data is reliable because you control the collection method, but it can be time-consuming.
原始数据是你自己为特定目的而亲手收集的信息。例如,调查同班同学的屏幕使用时间、在科学实验中测量植物生长,或计时跑 100 米所需的时间。原始数据可靠是因为你控制收集方法,但可能耗时较长。
Secondary data is information that someone else has already gathered. Sources include the internet, government statistics, textbooks, or data from a previous year’s experiment. Using secondary data saves time and resources, but you must check whether it is trustworthy and relevant to your question.
二手数据是别人已经收集好的信息。来源包括互联网、政府统计数据、教科书或往年实验的数据。使用二手数据节省时间和资源,但你必须检查它是否可信且与你的问题相关。
A survey is a method of gathering information by asking people questions. A questionnaire is the printed or online set of questions used in a survey. Questions must be clear, unbiased, and easy to answer. Closed questions give limited options (e.g., Yes/No, multiple choice), while open questions allow free-text responses.
调查是通过向人们提问来收集信息的方法。问卷是调查中使用的印刷版或在线问题集。问题必须清晰、无偏见且易于回答。封闭式问题提供有限选项(如 是/否、选择题),而开放式问题允许自由文本回答。
In statistics, a population is the entire group of individuals or items that you want to study — for example, all Year 8 students in Wales. A sample is a smaller, manageable subset selected from the population. A census collects data from every member of the population, whereas a sample surveys only a portion. A census gives accurate results but is often impractical for large populations.
在统计学中,总体是你想研究的全部个体或项目——例如,威尔士所有 8 年级学生。样本是从总体中选出的较小、可管理的子集。普查从总体的每个成员收集数据,而抽样调查只调查一部分。普查给出精确结果,但对于大总体往往不现实。
When you design a data collection sheet, include a title, clear column headings, and space for tallies and frequencies. Always pilot your questionnaire on a small group before the main survey to spot confusing questions.
设计数据收集表时,要包含标题、清晰的列标题以及用于划记和记录频数的空间。在正式调查前,先在小范围内试用你的问卷,以发现令人困惑的问题。
3. Presenting Data | 数据展示
A tally chart helps you record data as you collect it. You use vertical strokes, and every fifth stroke crosses the previous four to make a gate of five — for example, |||| becomes ~~||. This makes counting frequencies quick and reduces errors.
计数表帮助你在收集数据时进行记录。你使用竖线,每第五笔划穿前四笔形成一组五——例如 |||| 变成 ~~||。这使得计算频数快捷并减少错误。
Frequency is simply the number of times a particular value or category occurs. A frequency table lists all categories alongside their frequencies. It is the foundation for most statistical diagrams and calculations.
频数就是某个特定值或类别出现的次数。频数表列出了所有类别及其对应的频数。它是大多数统计图表和计算的基础。
A pictogram uses small pictures or symbols to represent data. Each picture stands for a certain number of items (the key). Always check the key carefully — a half picture represents half the key’s value. Pictograms are eye-catching but can be misleading if symbols are not scaled correctly.
象形图使用小图片或符号来呈现数据。每个图片代表一定数量的项目(图例)。务必仔细查看图例——半个图片代表图例值的一半。象形图引人注目,但如果符号比例不正确,可能会产生误导。
A bar chart displays categorical data with rectangular bars. The height or length of each bar shows the frequency. Bars are usually separated by equal gaps to remind you that the data is discrete or categorical. A bar-line chart is similar but uses a thin line instead of a wide bar.
条形图用矩形条展示分类数据。每个条形的高度或长度表示频数。条形之间通常有相等的间隙,提醒你数据是离散或分类的。条形折线图类似,但用细线代替宽条。
A pie chart is a circle divided into sectors, each representing a category. The angle of each sector is proportional to the frequency. Use the formula: sector angle = (frequency ÷ total frequency) × 360°. Pie charts are best for showing proportions of a whole.
饼图是一个被划分为扇形的圆,每个扇形代表一个类别。扇形的角度与频数成比例。使用公式:扇形角度 = (频数 ÷ 总频数) × 360°。饼图最适合展示各部分在整体中所占的比例。
A line graph uses points connected by straight lines to show how a variable changes over time or across ordered categories. The horizontal axis often represents time. Line graphs highlight trends and are excellent for continuous data.
折线图用直线连接各点来显示变量如何随时间或有序类别变化。横轴通常表示时间。折线图能突出趋势,非常适合连续数据。
A scatter graph (or scatter plot) shows the relationship between two sets of data. Each point represents a pair of values. Plot the independent variable on the x-axis and the dependent variable on the y-axis. Scatter graphs help us spot correlation.
散点图展示两组数据之间的关系。每个点代表一对数值。将自变量标在 x 轴,因变量标在 y 轴。散点图帮助我们发现相关性。
4. Measures of Central Tendency | 集中趋势量数
The mean is the arithmetic average of a set of numbers. To calculate the mean, add up all the data values and then divide by how many values there are. The formula is:
Mean = ∑x ÷ n
where ∑x is the sum of all data values and n is the number of values. The mean includes every piece of data, so it can be pulled up or down by extreme values (outliers).
平均值是一组数字的算术平均。要计算平均值,先把所有数据值相加,然后除以数据值的个数。公式为:平均 = ∑x ÷ n,其中 ∑x 是所有数据值的总和,n 是数据值的个数。因为平均值包含每个数据,所以它可能会被极端值(异常值)拉高或拉低。
The median is the middle value when the data is arranged in order from smallest to largest. If there is an odd number of values, the median is the exact centre. If there is an even number, it is the mean of the two middle numbers. The median is not affected by outliers, which makes it a better average for skewed data.
中位数是将数据按从小到大排列后处于中间位置的值。如果数据个数为奇数,中位数就是正中间的那个数;如果为偶数,则是中间两个数的平均值。中位数不受异常值影响,因此对于偏态数据,它是更好的平均数。
The mode (or modal value) is the value that appears most frequently. A data set can have one mode, more than one mode (bimodal or multimodal), or no mode if all values occur equally often. The mode is the only average that can be used with qualitative data — you can find the modal favourite colour, but not the mean colour.
众数(或模态值)是出现最频繁的值。一组数据可以有一个众数、多个众数(双众数或多众数),或者如果所有值出现次数相同则没有众数。众数是唯一可用于定性数据的平均数——你可以找到最喜欢的颜色的众数,但不能找到颜色的平均值。
Remember the three ‘M’s: Mean = average you calculate, Median = middle you locate, Mode = most frequent you spot. For a quick check, if you have ever heard ‘the average person’, it is usually the mean or median depending on context, but the mode tells you the most typical category.
记住三个 “M”:Mean(平均值)= 需要计算的平均数,Median(中位数)= 定位得到的中间值,Mode(众数)= 观察出现最多的值。快捷检查:如果你听过 “普通人” 这个概念,通常在不同情境中指均值或中位数,而众数告诉你最典型的类别。
5. Measures of Spread | 离散程度量数
The range measures how spread out the data is. It is the difference between the largest and smallest values:
Range = Maximum value – Minimum value
A small range means the data are clustered closely together; a large range suggests greater variability. Always write the range as a single number, not as ‘from 2 to 10’.
极差衡量数据的分散程度。它是最大值和最小值的差:极差 = 最大值 − 最小值。极差小表示数据紧密聚集;极差大表明变异性更大。始终将极差写成一个单一数字,而不是 “从 2 到 10″。
Consistency is linked to spread. In practical contexts, two sets of data can have the same mean but very different ranges. The set with the smaller range is more consistent. For example, a netball shooter who scores 8, 9, 10 goals across three games (range 2) is more consistent than one who scores 2, 10, 15 (range 13), even if their means are similar.
一致性与离散程度相关。在实际情境中,两组数据可能有相同的均值,但极差差异很大。极差较小的那一组更一致。例如,一位投球手在三场比赛中分别进了 8、9、10 个球(极差为 2),比另一位进了 2、10、15 个球(极差为 13)的
Published by TutorHao | KS3 统计 Revision Series | aleveler.com
更多咨询请联系16621398022(同微信)
屏轩国际教育cambridge primary/secondary checkpoint, cat4, ukiset,ukcat,igcse,alevel,PAT,STEP,MAT, ibdp,ap,ssat,sat,sat2课程辅导,国外大学本科硕士研究生博士课程论文辅导Cancel reply