Related Experiment Video
Updated: Oct 22, 2025

ExCYT: A Graphical User Interface for Streamlining Analysis of High-Dimensional Cytometry Data
Published on: January 16, 2019
Unsupervised Clustering of Hyperspectral Paper Data Using t-SNE
Binu Melit Devassy1, Sony George1, Peter Nussbaum1
1Department of Computer Science, Norwegian University of Science and Technology, 2802 Gjøvik, Norway.
Abstract:
For a suspected forgery that involves the falsification of a document or its contents, the investigator will primarily analyze the document's paper and ink in order to establish the authenticity of the subject under investigation. As a non-destructive and contactless technique, Hyperspectral Imaging (HSI) is gaining popularity in the field of forensic document analysis. HSI returns more information compared to conventional three channel imaging systems due to the vast number of narrowband images recorded across the electromagnetic spectrum. As a result, HSI can provide better classification results. In this publication, we present results of an approach known as the t-Distributed Stochastic Neighbor Embedding (t-SNE) algorithm, which we have applied to HSI paper data analysis. Even though t-SNE has been widely accepted as a method for dimensionality reduction and visualization of high dimensional data, its usefulness has not yet been evaluated for the classification of paper data. In this research, we present a hyperspectral dataset of paper samples, and evaluate the clustering quality of the proposed method both visually and quantitatively. The t-SNE algorithm shows exceptional discrimination power when compared to traditional PCA with k-means clustering, in both visual and quantitative evaluations.
Related Concept Videos
Attenuated Total Reflectance (ATR) Infrared Spectroscopy: Overview
The ATR process begins by directing a beam...
Light Acquisition
Distance Measurements by Taping
Ultraviolet and Visible (UV–Vis) Spectroscopy: Overview
Cluster Sampling Method
To choose a cluster sample, divide the population into clusters (groups) and then randomly select some of the clusters. All the members from these clusters are in the cluster sample. For example, if you randomly sample four departments from your...

