Related Experiment Video
Updated: Dec 8, 2025

07:34
A Simple Stimulatory Device for Evoking Point-like Tactile Stimuli: A Searchlight for LFP to Spike Transitions
Published on: March 25, 2014
10.2K
Sparse Sliced Inverse Regression Via Lasso
Qian Lin1,2,3, Zhigen Zhao1,2,3, Jun S Liu1,2,3
1Center of Statistical Science, Tsinghua University.
Journal of the American Statistical Association
|September 21, 2020
Summary
A new Lasso-SIR method estimates the sufficient dimension reduction (SDR) space consistently, even when data dimensions exceed sample size. This approach offers optimal convergence rates under sparsity conditions.
Area of Science:
- Statistics
- Machine Learning
- Dimensionality Reduction
Background:
- Sliced inverse regression (SIR) consistency for sufficient dimension reduction (SDR) estimation requires p << n.
- High-dimensional data (p >= n) necessitates additional assumptions like sparsity for SIR consistency.
Purpose of the Study:
- To develop a consistent SDR estimation method for high-dimensional data.
- To introduce a novel algorithm, Lasso-SIR, that addresses the limitations of traditional SIR.
Main Methods:
- Constructing artificial response variables from eigenvectors of the conditional covariance matrix.
- Applying Lasso regression to estimate the SDR space using these artificial responses.
Main Results:
- Lasso-SIR achieves consistency and optimal convergence rates under sparsity.
- Performance is validated through simulations and real-world data analysis, outperforming existing methods.
Conclusions:
- Lasso-SIR provides a robust solution for SDR estimation in high-dimensional settings.
- The method enhances the applicability of SDR techniques when dimensionality is a challenge.
More Related Videos
Related Concept Videos
Residuals and Least-Squares Property
8.7K
The vertical distance between the actual value of y and the estimated value of y. In other words, it measures the vertical distance between the actual data point and the predicted point on the line
If the observed data point lies above the line, the residual is positive, and the line underestimates the actual data value for y. If the observed data point lies below the line, the residual is negative, and the line overestimates the actual data value for y.
The process of fitting the best-fit...
If the observed data point lies above the line, the residual is positive, and the line underestimates the actual data value for y. If the observed data point lies below the line, the residual is negative, and the line overestimates the actual data value for y.
The process of fitting the best-fit...
8.7K
Regression Toward the Mean
6.7K
Regression toward the mean (“RTM”) is a phenomenon in which extremely high or low values—for example, and individual’s blood pressure at a particular moment—appear closer to a group’s average upon remeasuring. Although this statistical peculiarity is the result of random error and chance, it has been problematic across various medical, scientific, financial and psychological applications. In particular, RTM, if not taken into account, can interfere when...
6.7K
Calibration Curves: Linear Least Squares
3.9K
A calibration curve is a plot of the instrument's response against a series of known concentrations of a substance. This curve is used to set the instrument response levels, using the substance and its concentrations as standards. Alternatively, or additionally, an equation is fitted to the calibration curve plot and subsequently used to calculate the unknown concentrations of other samples reliably.
For data that follow a straight line, the standard method for fitting is the linear...
For data that follow a straight line, the standard method for fitting is the linear...
3.9K
Regression Analysis
7.4K
Regression analysis is a statistical tool that describes a mathematical relationship between a dependent variable and one or more independent variables.
In regression analysis, a regression equation is determined based on the line of best fit– a line that best fits the data points plotted in a graph. This line is also called the regression line. The algebraic equation for the regression line is called the regression equation. It is represented as:
In regression analysis, a regression equation is determined based on the line of best fit– a line that best fits the data points plotted in a graph. This line is also called the regression line. The algebraic equation for the regression line is called the regression equation. It is represented as:
7.4K
Multiple Regression
3.6K
Multiple regression assesses a linear relationship between one response or dependent variable and two or more independent variables. It has many practical applications.
Farmers can use multiple regression to determine the crop yield based on more than one factor, such as water availability, fertilizer, soil properties, etc. Here, the crop yield is the response or dependent variable as it depends on the other independent variables. The analysis requires the construction of a scatter plot...
Farmers can use multiple regression to determine the crop yield based on more than one factor, such as water availability, fertilizer, soil properties, etc. Here, the crop yield is the response or dependent variable as it depends on the other independent variables. The analysis requires the construction of a scatter plot...
3.6K
Quadratic Models
95
Quadratic models are mathematical representations used to describe relationships in which the rate of change changes at a constant rate. These models appear in a wide variety of natural and engineered systems, especially those involving motion, forces, and optimization. One common application is analyzing the vertical motion of objects influenced by gravity, such as a ball thrown into the air.In such scenarios, the object's height changes over time in a curved pattern, rising to a maximum point...
95

