Related Experiment Video
Updated: Sep 27, 2026

Quantifying Pain Location and Intensity with Multimodal Pain Body Diagrams
Published on: July 7, 2023
Democritus: Homotopy-Localized Causal Discourse Extraction from Language
1Adobe Research, 345 Park Avenue, San Jose, CA 95110, USA.
Abstract:
Natural-language documents contain many causal claims, but those claims are unstable under paraphrase, granularity shifts, and contextual drift. A collection may express one mechanism in many surface forms, while neighboring studies may agree locally yet fail to glue globally because relation families, polarities, temporal scopes, or regimes differ. This paper studies that post-extraction problem through Democritus, an implemented system for extracting, normalizing, localizing, and diagnosing causal discourse. We do not claim to identify ground-truth causal structure from text alone. Instead, we formulate the post-extraction layer as a homotopical repair problem for a non-compositional causal-discourse sketch. A normalization functor induces a class of weak equivalences on textual mentions; we prove that these data form a relative category with two-out-of-three and that normalization factors through its localization. Strict repair requires a corpus diagram to satisfy the sketch exactly, whereas homotopy repair asks for a weakly equivalent diagram that factors coherently after localization. We also define the conditions under which grounded discourse localization maps functorially into the observational-equivalence spaces of causal models studied by higher algebraic K-theory. The implementation approximates these formal constructions through normalized claim classes, an observed compatibility complex, regime-sensitive gluing, and provenance-preserving database artifacts. On a frozen 404-passage AltLex/UniCausal comparison, a restricted verbatim-span Democritus extractor attains 57.1% causal-classification F1 and 39.1% macro span F1, compared with 59.2% and 43.6% for the public UniCausal baseline. This similar classification F1 reflects a clear precision-recall tradeoff: Democritus achieves higher recall (80.9% versus 68.7%) but lower precision (44.1% versus 52.0%). Case studies on Emperor-penguin climate discourse, Mediterranean-diet studies, red-wine cardiovascular studies, and rising-ocean-temperature corpora expose both stable repeated mechanisms and cases where local causal claims cannot be assembled into one coherent corpus-level account.
Related Concept Videos
Language and Cognition
Nociception
Local Anesthetics: Differential Sensitivity of Nerve Fibers
Empathy
Agonism and Antagonism: Quantification
To quantify these effects, researchers use a dose-response curve, which provides valuable information about the potency and efficacy of a drug. Potency refers to...
Somatosensation