语义特征规范:一种交叉方法和跨语言的比较
Sasa L Kivisaari1, Annika Hultén2, Marijn van Vliet2
1Department of Neuroscience and Biomedical Engineering, Aalto University, School of Science, PO Box 12200, FI-00076, Espoo, Finland. sasa.kivisaari@aalto.fi.
Behavior research methods
|December 20, 2023
概括
这项研究比较了芬兰语和英语的行为和语料库基础的语义规范. 结果显示,这两种方法在很大程度上对具体的对象一致,验证了语料库规范,以便更容易进行语义分析.
科学领域:
- 认知科学 认知科学
- 语言学的语言学.
- 计算语言学 计算语言学
背景情况:
- 赋予刺激的意义是人类行为和语言的基础.
- 语义特征向量通常来自行为生产规范或语料库统计.
- 数据收集方法和语言对语义表示的影响尚不清楚.
研究的目的:
- 在芬兰语和英语中比较基于行为和语料库的语义规范.
- 评估数据收集方法对语义特征向量的影响.
- 验证用于语义研究的语料库衍生规范的使用.
主要方法:
- 开发了新的芬兰行为生产规范,用于抽象和具体的概念.
- 采用了基于行为和语料库的规范之间的全对全比较方法.
- 分析了来自两种方法的语义特征向量,跨两个语言.
主要成果:
- 基于行为和语料库的规范在很大程度上为具体对象提供了类似的语义信息.
- 在不同的规范集中,项目级映射是可行的.
- 对具体概念而言,语义表示的差异是最小的.
结论:
- 在许多语义研究中,体源规范是劳动密集型行为规范的有效和更容易获得的替代方案.
- 选择规范收集方法对具体对象的语义表示的影响有限.
- 未来的研究可以利用语料库规范进行大规模的语义分析.
相关概念视频
Multiple Comparison Tests
3.9K
Multiple comparison test, abbreviated as MCT, is a post hoc analysis generally performed after comparing multiple samples with one or more tests. An MCT will help identify a significantly different sample among multiple samples or a factor among multiple factors.
It would be easy to compare two samples using a significance alpha level of 0.05. In other words, there is only one sample pair to be compared. However, it would be difficult to identify a significantly different sample if the number...
It would be easy to compare two samples using a significance alpha level of 0.05. In other words, there is only one sample pair to be compared. However, it would be difficult to identify a significantly different sample if the number...
3.9K
Nominal Level of Measurement
28.6K
The way a set of data is measured is called its level of measurement. Correct statistical procedures depend on a researcher being familiar with levels of measurement. Not every statistical operation can be used with every set of data. For analysis, data are classified into four levels of measurement—nominal, ordinal, interval, and ratio.
The data that cannot be measured but can be grouped into categories fall under the nominal level of measurement. Data that is measured using a nominal...
The data that cannot be measured but can be grouped into categories fall under the nominal level of measurement. Data that is measured using a nominal...
28.6K
Sign Test for Nominal Data
99
The sign test is a nonparametric method used to evaluate hypotheses about the median of a single sample or to compare the medians of two related samples. The sign test is particularly useful when dealing with nominal data, which includes distinct categories without an inherent order, such as names, labels, and preferences. Nominal data restricts statistical analysis to evaluating population proportions rather than mean or median values that require continuous data.
For example, consider a...
For example, consider a...
99
Stereotype Content Model
14.7K
The Stereotype Content Model (SCM) was first proposed by Susan Fiske and her colleagues (Fiske, Cuddy, Glick & Xu, 2002; see also Fiske, 2012 and Fiske, 2017). The SCM specifies that when someone encounters a new group, they will stereotype them based on two metrics: warmth—or that group’s perceived intent, and how likely they are to provide help or inflict harm—and competence—or their ability to carry out that objective. Depending on the warmth-competence...
14.7K
Empirical Method to Interpret Standard Deviation
5.3K
The empirical rule, also known as the three-sigma rule, allows a statistician to interpret the standard deviation in a normally distributed dataset. The rule states that 68% of the data lies within one standard deviation from the mean, 95% lies within two standard deviations from the mean, and 99.7% lies within three standard deviations from the mean. Additionally, this rule is also called the 68-95-99.7 rule.
This rule is used widely in statistics to calculate the proportion of data values...
This rule is used widely in statistics to calculate the proportion of data values...
5.3K
Expected Frequencies in Goodness-of-Fit Tests
2.5K
A goodness-of-fit test is conducted to determine whether the observed frequency values are statistically similar to the frequencies expected for the dataset. Suppose the expected frequencies for a dataset are equal such as when predicting the frequency of any number appearing when casting a die. In that case, the expected frequency is the ratio of the total number of observations (n) to the number of categories (k).
2.5K


