Learning coefficient of generalization error in Bayesian estimation and vandermonde matrix-type singularity

Miki Aoyagi1, Kenji Nagata

  • 1Department of Mathematics, College of Science and Technology, Nihon University, Kanda, Chiyoda-ku, 101-8308, Japan. aoyagi.miki@nihon-u.ac.jp

Neural Computation
|February 3, 2012
PubMed
Summary

This study introduces a new algebraic geometry method to calculate learning coefficients, crucial for understanding generalization error in machine learning models. The findings provide tighter bounds and explicit values for complex models like neural networks.

Related Concept Videos

One-Compartment Open Model: Wagner-Nelson and Loo Riegelman Method for ka Estimation01:24

One-Compartment Open Model: Wagner-Nelson and Loo Riegelman Method for ka Estimation

This lesson introduces two critical methods in pharmacokinetics, the Wagner-Nelson and Loo-Riegelman methods, used for estimating the absorption rate constant (ka) for drugs administered via non-intravenous routes. The Wagner-Nelson method relates ka to the plasma concentration derived from the slope of a semilog percent unabsorbed time plot. However, it is limited to drugs with one-compartment kinetics and can be impacted by factors like gastrointestinal motility or enzymatic degradation.
On...
Propagation of Uncertainty from Systematic Error01:10

Propagation of Uncertainty from Systematic Error

The atomic mass of an element varies due to the relative ratio of its isotopes. A sample's relative proportion of oxygen isotopes influences its average atomic mass. For instance, if we were to measure the atomic mass of oxygen from a sample, the mass would be a weighted average of the isotopic masses of oxygen in that sample. Since a single sample is not likely to perfectly reflect the true atomic mass of oxygen for all the molecules of oxygen on Earth, the mass we obtain from this particular...
Singularity Functions for Shear01:26

Singularity Functions for Shear

In structural analysis, singularity functions are crucial in simplifying the representation of shear forces in beams under discontinuous loading. These functions describe discontinuous variations in shear force across a beam with varying loads by using a single mathematical expression, regardless of the complexity of the loading conditions. The singularity functions are derived from creating a free-body diagram of the beam and then making conceptual cuts at specific points to examine the shear...
Propagation of Uncertainty from Random Error00:59

Propagation of Uncertainty from Random Error

An experiment often consists of more than a single step. In this case, measurements at each step give rise to uncertainty. Because the measurements occur in successive steps, the uncertainty in one step necessarily contributes to that in the subsequent step. As we perform statistical analysis on these types of experiments, we must learn to account for the propagation of uncertainty from one step to the next. The propagation of uncertainty depends on the type of arithmetic operation performed on...
Singularity Functions for Bending Moment01:18

Singularity Functions for Bending Moment

Singularity functions simplify the representation of bending moments in beams subjected to discontinuous loading, allowing the use of a single mathematical expression. For a supported beam AB, with uniform loading from its midpoint M to the right side end B, the approach involves conceptual 'cuts' at specific points to determine the bending moment in each segment. By cutting the beam at a point between A and M, the bending moment for the segment before reaching midpoint M is represented using a...
Linearization and Approximation01:26

Linearization and Approximation

Linearization is a mathematical technique used to approximate complex, nonlinear functions with simpler linear models in the vicinity of a chosen reference point. The method is based on the idea that, although a function may be difficult to evaluate exactly, its behavior near a specific input value can often be closely approximated by the tangent line at that point. This approach is particularly useful when small deviations from a known value are involved.Consider the square root function, for...