Related Experiment Video
Updated: Mar 27, 2026

A Psychophysics Paradigm for the Collection and Analysis of Similarity Judgments
Published on: March 1, 2022
Self-weighted low-rank representation for multivariate compositional data
Zhengyan Liu1, Huiwen Wang2, Qing Zhao3
1School of Economics and Management,Beihang University, Beijing, 100191, China; Beijing Key Laboratory of Emergency Support Simulation Technologies for City Operations, Beijing, 100191, China.
Abstract:
Compositional data can effectively capture the relative information among different parts of a whole, which is frequently utilized in practical applications in recent years. Unfortunately, few works are yet available for clustering multivariate compositional data, due to the potential challenges created by the complex grouping structure and the existence of uninformative variables. In this paper, we propose a self-weighted low-rank representation (SWLRR) method to cluster multivariate compositional data. Specifically, a variable weighting strategy is introduced to learn appropriate weights for different compositional data variables, which can highlight the informative variables and resist the uninformative ones. Meanwhile, for the need of exploring the grouping structure, the global and local structures of data are simultaneously captured within the weighted data space, which is realized through generalizing the self-expressive property and adding a graph constraint term to multivariate compositional data, respectively. Moreover, the low-rank constraint is imposed on the representation for robustness. On this basis, we construct a unified optimization framework and present the solving algorithm by means of the alternating direction method of multipliers (ADMM). Experimental results on the synthetic and practical datasets show the advantages of the proposed clustering method compared with other competitors, and demonstrate the effectiveness of the proposed method to recognize the contributions of different compositional data variables to the clustering process.
Related Concept Videos
Weighted Mean
For example, consider the number of goals scored in the matches of a tournament. While computing the average number of goals scored in the tournament, it may be more important to...
Vector Algebra: Method of Components
In many applications, the magnitudes and directions of...
Reduced Mass Coordinates: Isolated Two-body Problem
Friedman Two-way Analysis of Variance by Ranks
Regression Toward the Mean
Residuals and Least-Squares Property
If the observed data point lies above the line, the residual is positive, and the line underestimates the actual data value for y. If the observed data point lies below the line, the residual is negative, and the line overestimates the actual data value for y.
The process of fitting the best-fit...

