Related Experiment Video
Updated: May 21, 2025

Eye-tracking to Distinguish Comprehension-based and Oculomotor-based Regressive Eye Movements During Reading
Published on: October 18, 2018
NOIRBETTIK: A reading comprehension based multiple choice question answering dataset in Bangla language
Tanjim Taharat Aurpa1, Md Shahriar Hossain Apu2, Farzana Akter2
1Department of Data Science and Engineering, Bangabandhu Sheikh Mujibur Rahman Digital, University, Bangladesh.
A new Bangla dataset, NOIRBETTIK, was created for automated reading comprehension and multiple-choice question (MCQ) answering. This resource supports advancements in educational technology for Bangla language learners.
Area of Science:
- Natural Language Processing
- Educational Technology
- Computational Linguistics
Background:
- The COVID-19 pandemic accelerated online education, increasing the need for automated learning and assessment tools.
- Multiple-choice questions (MCQs) are crucial for evaluating comprehension in digital learning environments.
- Existing Bangla language resources for context-based NLP tasks, particularly reading comprehension, are scarce.
Purpose of the Study:
- To introduce NOIRBETTIK, a novel, human-made dataset for reading comprehension-based MCQ answering in Bangla.
- To address the critical gap in high-quality Bangla datasets for NLP and educational applications.
- To facilitate the development of advanced automated systems for Bangla language learning and evaluation.
Main Methods:
- Dataset creation involved sourcing authentic Bangla materials (books, articles, biographies).
- Passages were paired with multiple-choice questions, each having four alternatives.
- The dataset's creation and annotation processes were meticulously documented and compared to existing resources.
Main Results:
- The NOIRBETTIK dataset comprises longer passages and contextually relevant MCQs.
- It is a unique, human-generated resource specifically designed for Bangla reading comprehension tasks.
- The dataset's structure and creation methodology are detailed, ensuring transparency and usability.
Conclusions:
- The release of NOIRBETTIK provides a valuable resource for Bangla NLP research.
- This dataset can significantly enhance reading comprehension systems for Bangla-speaking students.
- NOIRBETTIK is poised to drive innovation in educational technologies for the Bangla language community.
More Related Videos
09:00Author Spotlight: Validation of SICOLE-R for Assessing Cognitive and Reading Skills in Spanish-Speaking Children and Its Role in Personalized Education
Published on: August 16, 2024
06:48Lexical Decision Task for Studying Written Word Recognition in Adults with and without Dementia or Mild Cognitive Impairment
Published on: June 25, 2019
Related Concept Videos
Data Collection by Observations
An astronomer viewing the motion and brightness of stars in the sky and recording the data is an example of observational data collection. A botanist recording...
Multiple Comparison Tests
It would be easy to compare two samples using a significance alpha level of 0.05. In other words, there is only one sample pair to be compared. However, it would be difficult to identify a significantly different sample if the number...
Data Collection by Survey
Classification of Systems-II
Surveys
Multiple Regression
Farmers can use multiple regression to determine the crop yield based on more than one factor, such as water availability, fertilizer, soil properties, etc. Here, the crop yield is the response or dependent variable as it depends on the other independent variables. The analysis requires the construction of a scatter plot...