Related Experiment Video
Updated: Dec 5, 2025

07:35
Selecting Multiple Biomarker Subsets with Similarly Effective Binary Classification Performances
Published on: October 11, 2018
7.9K
Using the contextual language model BERT for multi-criteria classification of scientific articles
Ashwin Karthik Ambalavanan1, Murthy V Devarakonda1
1Arizona State University, United States.
Journal of Biomedical Informatics
|October 15, 2020
Summary
This study compares ensemble models for biomedical article screening. A cascade ensemble achieved higher precision, while a single integrated model offered better recall for systematic reviews.
Area of Science:
- Biomedical Natural Language Processing
- Information Retrieval
- Machine Learning
Background:
- Efficiently finding scientific articles is a key challenge in biomedical NLP.
- Applications like systematic reviews and interactive search require effective article screening.
- The optimal approach for modeling multiple selection criteria in ensemble models is unclear.
Purpose of the Study:
- To compare different ensemble architectures for screening scientific articles based on multiple criteria.
- To evaluate the performance of a novel cascade ensemble model against other approaches.
- To assess the impact of modern contextual language models like SciBERT on this task.
Main Methods:
- Framed article screening as a text classification problem.
- Developed and compared a cascade ensemble model with an individual task learner (ITL) and other ensemble methods.
- Utilized SciBERT on a dataset of ~49K MEDLINE abstracts (Clinical Hedges).
Main Results:
- The cascade ensemble demonstrated superior precision and F-measure compared to ITL and other models.
- The individual task learner (ITL) achieved significantly higher recall.
- ITL showed improved precision at high recall levels in fixed high recall studies.
Conclusions:
- Contextual language models like SciBERT are effective for scientific article screening.
- The ITL model is suitable for systematic reviews requiring high recall.
- The cascade ensemble excels in interactive search applications due to its high F-measure.
Keywords:
BERTBiomedical natural language processingMachine learningNeural networksSciBERTScreening scientific articlesText classificationMore Related Videos
Related Concept Videos
Classification of Systems-II
400
Continuous-time systems have continuous input and output signals, with time measured continuously. These systems are generally defined by differential or algebraic equations. For instance, in an RC circuit, the relationship between input and output voltage is expressed through a differential equation derived from Ohm's law and the capacitor relation,
400
Classification of Systems-I
468
Linearity is a system property characterized by a direct input-output relationship, combining homogeneity and additivity.
Homogeneity dictates that if an input x(t) is multiplied by a constant c, the output y(t) is multiplied by the same constant. Mathematically, this is expressed as:
Homogeneity dictates that if an input x(t) is multiplied by a constant c, the output y(t) is multiplied by the same constant. Mathematically, this is expressed as:
468
How Data are Classified: Categorical Data
40.3K
A variable, usually notated by capital letters such as X and Y, is a characteristic or measurement that can be determined for each member of a population. Data are the actual values of variables. They may be numbers, or they may be words. Datum is a single value.
Data are classified based on whether they are measurable or not. Categorical data cannot be measured; instead, it can be divided into categories. For example, if Y denotes a person's party affiliation, some examples of Y include...
Data are classified based on whether they are measurable or not. Categorical data cannot be measured; instead, it can be divided into categories. For example, if Y denotes a person's party affiliation, some examples of Y include...
40.3K
Classification of Signals
1.2K
In signal processing, signals are classified based on various characteristics: continuous-time versus discrete-time, periodic versus aperiodic, analog versus digital, and causal versus noncausal. Each category highlights distinct properties crucial for understanding and manipulating signals.
A continuous-time signal holds a value at every instant in time, representing information seamlessly. In contrast, a discrete-time signal holds values only at specific moments, often denoted as x(n), where...
A continuous-time signal holds a value at every instant in time, representing information seamlessly. In contrast, a discrete-time signal holds values only at specific moments, often denoted as x(n), where...
1.2K
Aggregates Classification
587
Aggregate classification is generally based on its size, petrographic characteristics, weight, and source. Size classification ranges from coarse to fine aggregates, defined by the size of the particles. Coarse aggregates are particles that do not pass through ASTM sieve No. 4, and aggregates that pass through the sieve are fine aggregates.
Petrographic classification groups aggregates based on common mineralogical characteristics. Some of the common mineral groups found in aggregates are...
Petrographic classification groups aggregates based on common mineralogical characteristics. Some of the common mineral groups found in aggregates are...
587
Multiple Regression
3.6K
Multiple regression assesses a linear relationship between one response or dependent variable and two or more independent variables. It has many practical applications.
Farmers can use multiple regression to determine the crop yield based on more than one factor, such as water availability, fertilizer, soil properties, etc. Here, the crop yield is the response or dependent variable as it depends on the other independent variables. The analysis requires the construction of a scatter plot...
Farmers can use multiple regression to determine the crop yield based on more than one factor, such as water availability, fertilizer, soil properties, etc. Here, the crop yield is the response or dependent variable as it depends on the other independent variables. The analysis requires the construction of a scatter plot...
3.6K

