Jove
Visualize
Contact Us
JoVE
x logofacebook logolinkedin logoyoutube logo
ABOUT JoVE
OverviewLeadershipBlogJoVE Help Center
AUTHORS
Publishing ProcessEditorial BoardScope & PoliciesPeer ReviewFAQSubmit
LIBRARIANS
TestimonialsSubscriptionsAccessResourcesLibrary Advisory BoardFAQ
RESEARCH
JoVE JournalMethods CollectionsJoVE Encyclopedia of ExperimentsArchive
EDUCATION
JoVE CoreJoVE BusinessJoVE Science EducationJoVE Lab ManualFaculty Resource CenterFaculty Site
Terms & Conditions of Use
Privacy Policy
Policies

Related Concept Videos

Facial Feedback Hypothesis01:24

Facial Feedback Hypothesis

243
Charles Darwin proposed that facial expressions are an evolutionary adaptation for communication. He argued that these expressions are not influenced by culture but are universal across species. For example, a snarling expression with exposed teeth signals a threat in many animals, including humans. Darwin also suggested that displaying an emotion can intensify the feeling. Smiling, for example, could enhance one's sense of happiness. This idea laid the foundation for understanding the role...
243
Nonconscious Mimicry01:13

Nonconscious Mimicry

4.6K
Nonconscious mimicry occurs when individuals alter their mannerisms to match the behaviors and expressions of those nearby, without intention.
4.6K
Perceiving Loudness, Pitch, and Location01:21

Perceiving Loudness, Pitch, and Location

418
The human brain perceives pitch through two primary mechanisms reflected in place theory and frequency theory. Each mechanism describes how sound waves are interpreted as specific pitches by the brain, offering insights into the intricate processes of auditory perception.
Place theory, or place coding, suggests that different pitches are heard because various sound waves activate specific locations along the cochlea's basilar membrane. The brain determines the pitch of a sound by...
418
Masking and Demasking Agents01:19

Masking and Demasking Agents

2.6K
EDTA titrations may necessitate masking and demasking agents to temporarily protect a particular metal ion in a mixture from the EDTA reaction. These agents facilitate the sequential analysis of the metal ions by forming stable complexes with some—but not all—metal ions during certain steps.
There are many masking agents, such as cyanide, fluoride, triethanolamine, thiourea, and 2,3-bis(sulfanyl)propan-1-ol (formerly 2,3-dimercapto-1-propanol), with the masking agent chosen based on...
2.6K
Linear Approximation in Frequency Domain01:26

Linear Approximation in Frequency Domain

130
Linear systems are characterized by two main properties: superposition and homogeneity. Superposition allows the response to multiple inputs to be the sum of the responses to each individual input. Homogeneity ensures that scaling an input by a scalar results in the response being scaled by the same scalar.
In contrast, nonlinear systems do not inherently possess these properties. However, for small deviations around an operating point, a nonlinear system can often be approximated as linear....
130
¹H NMR: Interpreting Distorted and Overlapping Signals01:02

¹H NMR: Interpreting Distorted and Overlapping Signals

1.1K
Spin systems where the difference in chemical shifts of the coupled nuclei is greater than ten times J are called first-order spin systems. These nuclei are weakly coupled, and their chemical shifts and coupling constant can generally be estimated from the well-separated signals in the spectrum.
As Δν decreases and the signals move closer, the doublets appear increasingly distorted. The intensities of the inner lines increase at the cost of those of the outer lines as the signals are...
1.1K

You might also read

Related Articles

Articles linked to this work by shared authors, journal, and citation graph.

Sort by
Same author

Improved prediction of chlorophyll-a concentrations using advancing graph neural network variants.

The Science of the total environment·2025
Same author

Data-driven analysis of climate impact on tomato and apple prices using machine learning.

Heliyon·2025
Same author

Directional and Strain-Specific Interaction Between <i>Lactobacillus plantarum</i> and <i>Staphylococcus aureus</i>.

Microorganisms·2025
Same author

Development of high-performance inducible and secretory expression vector and host system for enhanced recombinant protein production.

Scientific reports·2024
Same author

Microbiology of tattoo-associated infections since 1820.

The Lancet. Microbe·2024
Same author

Causes, patterns, and epidemiology of tattoo-associated infections since 1820.

The Lancet. Microbe·2024

Related Experiment Video

Updated: Sep 6, 2025

Author Spotlight: Investigating the Impact of Emotional Prosodies on Voice Recognition and Perception
05:48

Author Spotlight: Investigating the Impact of Emotional Prosodies on Voice Recognition and Perception

Published on: August 9, 2024

1.6K

BPCNN: Bi-Point Input for Convolutional Neural Networks in Speaker Spoofing Detection.

Sunghyun Yoon1, Ha-Jin Yu2

  • 1Department of Artificial Intelligence, Kongju National University, Cheonan 31080, Korea.

Sensors (Basel, Switzerland)
|June 24, 2022
PubMed
Summary

We introduce bi-point input for convolutional neural networks (CNNs) to process variable-length features like speech. This method improves performance by feeding pairs of segments, enhancing information available to the CNN.

Keywords:
bi-point inputbidirectional feature segmentationconvolutional neural network (CNN)spoofing detectionvariable-length features

More Related Videos

Foreign Accent and Forensic Speaker Identification in Voice Lineups: The Influence of Acoustic Features Based on Prosody
09:09

Foreign Accent and Forensic Speaker Identification in Voice Lineups: The Influence of Acoustic Features Based on Prosody

Published on: September 27, 2024

522
A Lightweight, Headphones-based System for Manipulating Auditory Feedback in Songbirds
10:13

A Lightweight, Headphones-based System for Manipulating Auditory Feedback in Songbirds

Published on: November 26, 2012

14.5K

Related Experiment Videos

Last Updated: Sep 6, 2025

Author Spotlight: Investigating the Impact of Emotional Prosodies on Voice Recognition and Perception
05:48

Author Spotlight: Investigating the Impact of Emotional Prosodies on Voice Recognition and Perception

Published on: August 9, 2024

1.6K
Foreign Accent and Forensic Speaker Identification in Voice Lineups: The Influence of Acoustic Features Based on Prosody
09:09

Foreign Accent and Forensic Speaker Identification in Voice Lineups: The Influence of Acoustic Features Based on Prosody

Published on: September 27, 2024

522
A Lightweight, Headphones-based System for Manipulating Auditory Feedback in Songbirds
10:13

A Lightweight, Headphones-based System for Manipulating Auditory Feedback in Songbirds

Published on: November 26, 2012

14.5K

Area of Science:

  • Artificial Intelligence
  • Machine Learning
  • Signal Processing

Background:

  • Convolutional Neural Networks (CNNs) typically require fixed-size inputs, posing challenges for variable-length data like speech.
  • Existing methods like feature segmentation limit CNNs to processing one segment at a time, hindering comprehensive analysis.

Purpose of the Study:

  • To propose a novel input method, "bi-point input," for CNNs to effectively handle variable-length features.
  • To enhance the information processing capacity of CNNs by allowing them to consider multiple segments simultaneously.

Main Methods:

  • The proposed bi-point input method feeds pairs of segments from a variable-length feature into a CNN concurrently.
  • Various combination strategies for these segments are explored, with guidance provided for optimal segment length selection.
  • The method was evaluated on spoofing detection tasks using the ASVspoof 2019 database.

Main Results:

  • The bi-point input method significantly improved performance in spoofing detection tasks.
  • A relative reduction in equal error rate (EER) of approximately 17.2% for logical access (LA) and 43.8% for physical access (PA) was achieved.

Conclusions:

  • The bi-point input method offers a more effective way for CNNs to process variable-length features compared to traditional segmentation.
  • This approach enhances the contextual information available to the model, leading to improved accuracy in tasks like speaker verification and anti-spoofing.