Related Experiment Video
Updated: Sep 14, 2026

Removal of Exogenous Materials from the Outer Portion of Frozen Cores to Investigate the Ancient Biological Communities Harbored Inside
Published on: July 3, 2016
Automated detection of freeze-thaw signatures in archaeological sediments using deep learning
Sofia Kouki1, Li Li1,2,3, Vera Aldeias1
1Universidade do Algarve, The Interdisciplinary Center for Archaeology and the Evolution of Human Behaviour ICArHEB, Faro, Campus de Gambelas, 8005-139, Portugal.
Background:
Freeze-thaw processes leave diagnostic traces in archaeological soils and sediments that are central to reconstructing past climates and understanding hominin adaptations to glacial environments. Frost features can be defined through soil micromorphology techniques using well-defined diagnostic criteria, but these traces can be subtle, variably expressed, and overlap with other pedogenic processes, making their identification time-consuming, expert-dependent, and subject to high inter-observer variability.
Methods:
We trained five convolutional neural network architectures on photomicrographs from eleven Plio-Pleistocene archaeological sites, implementing a two-stage classification approach by detecting first the presence of those features (binary classification) and then the feature type (multiclass classification), and validated model outputs against both interpretability analysis and a blind survey of practicing micromorphologists.
Results:
Results reveal a performance paradox: models achieve high performance but rely on spurious correlations rather than diagnostic criteria, while models that focus on micromorphologically relevant features show lower overall performance. Expert agreement on the same task is low, with uncertainty concentrated in feature detection rather than classification. Crucially, model and expert errors are largely independent, and each captures different aspects of frost feature recognition, establishing a basis for complementarity.
Conclusions:
These findings demonstrate that effective computational integration in micromorphology requires not only accurate classification but interpretability validation ensuring that model's reason from the same diagnostic criteria as experts. We propose a human-in-the-loop approach where models provide consistent first screening, while experts offer contextual interpretation and diagnostic validation. Additionally, we present an interactive open-access tool that implements this pipeline to facilitate adoption and repeatability.

