📚 Year 9 CCEA Statistics: Quick Guide to Vocabulary and Terminology | Year 9 CCEA 统计:词汇术语速记指南
Mastering statistics begins with a solid grasp of its unique language. This guide walks you through the essential vocabulary and terminology for the Year 9 CCEA Statistics course, breaking down each term with clear examples. By using this quick-reference resource, you will build confidence in collecting, analysing, and interpreting data, as well as understanding probability.
掌握统计学要从扎实理解其独特的语言开始。本指南将带你回顾Year 9 CCEA统计学课程的核心词汇与术语,并配以清晰的示例进行解析。通过使用这份速记资源,你将在数据的收集、分析和解释以及概率的理解方面建立信心。
1. Data Types | 数据类型
Data refers to pieces of information gathered through observation, measurement, or research. In statistics, we classify data into two broad categories: quantitative (numerical) and qualitative (categorical). Quantitative data can be further split into discrete and continuous types. Understanding the type of data you are handling is crucial because it determines which statistical methods and charts are appropriate.
数据是指通过观察、测量或研究收集到的信息。在统计学中,我们将数据分为两大类:定量(数值型)数据和定性(类别型)数据。定量数据又可细分为离散数据和连续数据。了解你所处理的数据类型至关重要,因为它决定了哪些统计方法和图表是适用的。
Discrete data can only take specific, separate values. They are usually counted, such as the number of students in a class or the score on a dice. Continuous data, on the other hand, can take any value within a range and are often measured, like a person’s height (which could be 162.5 cm) or the time taken to run a race.
离散数据只能取特定的、分开的值。它们通常是计数的结果,例如一个班级的学生人数或掷骰子的分数。另一方面,连续数据可以取范围内任意值,通常是测量得到的,比如一个人的身高(可能是162.5厘米)或跑步比赛所花的时间。
Qualitative data describes attributes or characteristics that cannot be measured numerically. Examples include favourite colour, hair type, or the brand of a mobile phone. This data is often grouped into categories for analysis.
定性数据描述的是无法用数字衡量的属性或特征。例如最喜欢的颜色、头发类型或手机品牌。这类数据通常被归入不同类别进行分析。
2. Variables: Independent and Dependent | 变量:自变量与因变量
A variable is any characteristic, number, or quantity that can change or vary across individuals or situations. In an experiment or survey, we often look for a relationship between two variables: the independent variable (the one we change or control) and the dependent variable (the one we measure or observe).
变量是指任何在个体或情境之间可能发生变化或不同的特征、数字或数量。在实验或调查中,我们经常寻找两个变量之间的关系:自变量(我们改变或控制的变量)和因变量(我们测量或观察的变量)。
For example, a student might investigate whether the temperature of water (independent variable) affects how quickly a sugar cube dissolves (dependent variable). The independent variable is plotted on the x-axis of a scatter graph, while the dependent variable is plotted on the y-axis. Recognising this pairing helps you design investigations and interpret graphs correctly.
例如,一名学生可能研究水温(自变量)是否影响方糖溶解的速度(因变量)。在散点图中,自变量通常画在x轴上,因变量画在y轴上。识别这种配对有助于你正确设计调查和解读图表。
3. Mean, Median, and Mode | 平均数、中位数和众数
These three measures are known as averages or measures of central tendency. They summarise a set of data with a single representative value.
这三种度量被称为平均数或集中趋势的度量。它们用一个代表性数值来概括一组数据。
The mean is the sum of all data values divided by the number of values. It is often called the arithmetic average.
平均数是将所有数据值相加后除以数值的个数。它常被称为算术平均数。
Mean = (Sum of all data values) ÷ (Number of data values)
For example, the mean of 3, 7, 8, 5, 2 is (3+7+8+5+2) ÷ 5 = 25 ÷ 5 = 5. The mean can be affected by extreme values, or outliers.
例如,3、7、8、5、2的平均数是(3+7+8+5+2)÷5 = 25÷5 = 5。平均数可能会受到极端值(即离群值)的影响。
The median is the middle value when the data are arranged in order. For an odd number of values, it is simply the central number; for an even number, it is the mean of the two central numbers.
中位数是将数据按顺序排列后处于中间位置的值。当数据个数为奇数时,它就是正中间的数;当为偶数时,则是中间两个数的平均数。
If we have 2, 3, 5, 7, 8, the median is 5. For 2, 3, 5, 7, 8, 10, the median is (5+7)÷2 = 6. The median is not affected by outliers, making it useful for skewed data.
假设数据为2、3、5、7、8,中位数是5。对于2、3、5、7、8、10,中位数是(5+7)÷2 = 6。中位数不受离群值影响,因此适用于偏斜的数据。
The mode is the value that appears most frequently. A data set can have one mode, more than one mode (bimodal or multimodal), or no mode at all if all values appear equally often. The mode is particularly useful for qualitative data.
众数是出现次数最多的值。一组数据可能有一个众数、多个众数(双众数或多众数),或者如果没有值出现率最高则没有众数。众数对于定性数据特别有用。
4. Range and Spread | 极差与离散度
While averages give a central value, the range measures how spread out the data are. It is the difference between the highest and lowest values in the set.
平均数给出中心值,而极差测量的是数据的分散程度。它是一组数据中最大值与最小值之间的差值。
Range = Largest value – Smallest value
A larger range indicates greater variability. For example, two classes might both have a mean test score of 60, but Class A with scores ranging from 55 to 65
Published by TutorHao | Year 9 统计 Revision Series | aleveler.com
更多咨询请联系16621398022(同微信)