相关实验视频
Updated: Jul 6, 2025

06:45
Loneliness Assuaged: Eye-Tracking an Audience Watching Barrage Videos
Published on: May 29, 2020
4.2K
来自YouTube的孟加拉语新闻和论数据集
Lomat Haider Chowdhury1, Salekul Islam2, Swakkhar Shatabda2
1Department of Computer Science and Engineering, Ahsanullah University of Science and Technology, Bangladesh.
Data in brief
|January 4, 2024
概括
这项研究引入了大量的孟加拉语YouTube新闻评论数据集,为公众论动态提供了有价值的见解. 研究人员现在可以分析在线讨论,并跟踪在孟加拉新闻领域不断演变的公众情绪.
科学领域:
- 社会科学 社会科学 社会科学
- 媒体研究 媒体研究
- 计算语言学 计算语言学
背景情况:
- 传统的新闻传播正在发展,在线平台和YouTube频道成为视觉新闻的主要来源.
- 通过评论和回复,公众对在线新闻的参与显而易见,从而创造了丰富的数据供分析.
- 关于孟加拉新闻和相关公众评论的研究存在重大差距.
研究的目的:
- 为了研究目的,提供来自YouTube的孟加拉新闻评论和回复的综合数据集.
- 为了促进对公众论模式及其在孟加拉新闻背景下随着时间的推移而发生的演变的分析.
- 为了解决对孟加拉语新闻的在线公共话语理解的研究缺口.
主要方法:
- 收集了762,678条公众评论和回复的数据集,来自2017年至2023年间发表的16,016个孟加拉新闻视频.
- 提取了每个新闻项目和评论的15个关键属性,包括元数据,参与度量和文本内容.
- 通过编码作者名字来确保评论者的隐私,并提供了一个翻译的文件以使其易于访问.
主要成果:
- 孟加拉语新闻评论和回复的大规模数据集现在可用于学术研究.
- 数据集包括关于新闻视频和相关用户参与度的详细信息,例如喜欢和浏览量.
- 评论员隐私是通过数据匿名化技术来维护的.
结论:
- 发布的数据集为研究孟加拉语数字领域的论提供了宝贵的资源.
- 学者可以利用这些数据来识别趋势,分析情绪,并了解在线讨论的动态.
- 这项工作通过提供独特的数据资源,为不断增长的计算社会科学领域做出了贡献.
相关概念视频
Bias
4.2K
Bias refers to any tendency that prevents a question from being considered unprejudiced. In research, bias occurs when one outcome or answer is selected or encouraged over others in sampling or testing. Bias can occur during any research phase, including study design, data collection, analysis, and publication.
In statistics, a sampling bias is created when a sample is collected from a population, and some members of the population are not as likely to be chosen as others (remember, each member...
In statistics, a sampling bias is created when a sample is collected from a population, and some members of the population are not as likely to be chosen as others (remember, each member...
4.2K
Data Collection by Observations
12.0K
Data collection refers to a systematic way of obtaining, observing, measuring, and analyzing accurate information. Observational studies are one of the most widely used methods of data collection. It involves collecting data by observing the behavior and physical characteristics of a sample without making any modifications to the sample.
An astronomer viewing the motion and brightness of stars in the sky and recording the data is an example of observational data collection. A botanist recording...
An astronomer viewing the motion and brightness of stars in the sky and recording the data is an example of observational data collection. A botanist recording...
12.0K
Mean From a Frequency Distribution
16.7K
Sometimes, data gathered from an experiment on a large sample or population are organized into concise tables. In such cases, the frequency of the quantitative data set is plotted in the form of a table. Or else, the data values are grouped into the quantity’s intervals, which form classes, and their respective frequencies are known. That is, the data values are distributed over different categories or classes. This is known as frequency distribution.
When such a data set is encountered,...
When such a data set is encountered,...
16.7K
Data Collection by Survey
6.5K
The systematic method of obtaining and analyzing accurate information of a population is called data collection. A survey is a standard method of data collection that involves collecting information from a target human population about their experience, opinion, or knowledge of a product, service, or process. The responses are recorded and interpreted. The most common survey examples are written questionnaires, face-to-face or telephonic conversations, focus groups, and electronic (e-mail or...
6.5K
Surveys
14.8K
Often, psychologists develop surveys as a means of gathering data. Surveys are lists of questions to be answered by research participants, and can be delivered as paper-and-pencil questionnaires, administered electronically, or conducted verbally. Generally, the survey itself can be completed in a short time, and the ease of administering a survey makes it easy to collect data from a large number of people.
14.8K
Prediction Intervals
2.3K
The interval estimate of any variable is known as the prediction interval. It helps decide if a point estimate is dependable.
However, the point estimate is most likely not the exact value of the population parameter, but close to it. After calculating point estimates, we construct interval estimates, called confidence intervals or prediction intervals. This prediction interval comprises a range of values unlike the point estimate and is a better predictor of the observed sample value, y.
However, the point estimate is most likely not the exact value of the population parameter, but close to it. After calculating point estimates, we construct interval estimates, called confidence intervals or prediction intervals. This prediction interval comprises a range of values unlike the point estimate and is a better predictor of the observed sample value, y.
2.3K

