Related Experiment Video
Updated: Jan 11, 2026

Measuring Attention and Visual Processing Speed by Model-based Analysis of Temporal-order Judgments
Published on: January 23, 2017
Ratio divergence learning using target energy in restricted Boltzmann machines: Beyond Kullback-Leibler divergence
Yuichi Ishida1, Yuma Ichikawa1,2, Aki Dote1
1Fujitsu Ltd., Kawasaki, Japan.
Abstract:
We propose ratio divergence (RD) learning for discrete energy-based models, a method that utilizes both training data and a tractable target energy function. We apply RD learning to restricted Boltzmann machines (RBMs), which are a minimal model that satisfies the universal approximation theorem for discrete distributions. RD learning combines the strength of both forward and reverse Kullback-Leibler divergence (KLD) learning, effectively addressing the "notorious" issues of underfitting with the forward KLD and mode collapse with the reverse KLD. Since the summation of forward and reverse KLD seems to be sufficient to combine the strength of both approaches, we include this learning method as a direct baseline in numerical experiments to evaluate its effectiveness. Numerical experiments demonstrate that RD learning outperforms other learning methods in terms of energy function fitting, mode-covering, and learning stability across various discrete energy-based models. Moreover, the performance gaps between RD learning and the other learning methods become more pronounced as the dimensions of target models increase.
Related Concept Videos
Maxwell-Boltzmann Distribution: Problem Solving
This distribution function f(v) is defined by saying that the expected number N (v1,v2) of particles with speeds between v1 and v2 is given by
Divergence and Stokes' Theorems
Energy Conservation and Bernoulli's Equation
All the terms in the equation have the dimension of energy per unit volume. The kinetic energy per unit volume is called the kinetic energy density, and the potential energy per unit volume is...
Conservation of Energy in Control Volume
For steady flow systems, the time derivative of the stored energy becomes zero since there is no energy accumulation within the control volume. This simplifies the energy equation to:
Difference from Background: Limit of Detection
The LOD indicates the presence or absence...
Conservation of Energy: Application