Related Experiment Video
Updated: Sep 13, 2025

Evidence-based Knowledge Synthesis and Hypothesis Validation: Navigating Biomedical Knowledge Bases via Explainable AI and Agentic Systems
Published on: June 13, 2025
Hierarchical knowledge-guided reasoning for text-based person re-identification
Ruigeng Zeng1, Wentao Ma2, Tongqing Zhou3
1Laboratory of Digitizing Software for Frontier Equipment, National University of Defence Technology, Changsha, 410073, Hunan, China; National Key Laboratory of Parallel and Distributed Computing, National University of Defense Technology, Changsha, 410073, Hunan, China; School of Information and Artificial Intelligence, Anhui Agricultural University, Hefei, 230036, Anhui, China.
None:
Masked language modeling (MLM) has expanded the exploration of text-image person re-identification (TIReID) tasks from coarse-granularity to fine-grained alignment. Whereas, we note that vanilla MLM picks random tokens for visual-to-token reasoning, which could fail the intention of semantic visual-textual alignment by indistinguishably focusing on all the sub-words. This work proposes to leverage the inherent hierarchical scene graph knowledge in each text for guiding token masking and enhancing cross-modal representation in TIReID, thus relieving the pitfall of blind visual-textual alignment. The proposed framework, Hierarchical Knowledge-Guided Reasoning (HKGR), parses object-level, attribute-level, and relation-level masking according to phrase knowledge constructions and explicitly lets the training of a dedicated encoder focus on the visual-to-token reasoning of these highlighted tokens. In addition, we propose a Multi-Grained Semantic Alignment (MGA) module, which leverages the token selection method and image-text similarity distribution constraint to further facilitate the semantic alignment between image and text at both coarse-grained and fine-grained levels. Experimental results demonstrate that our HKGR framework achieves state-of-the-art (SoTA) performance on three public benchmark datasets at all evaluation metrics. We believe that the knowledge-guided idea is beneficial to other multi-modal research communities, including cross-modal retrieval and visual question answering. Code is available at https://github.com/Ray-Zhen/HKGR.git.
More Related Videos
08:25Combining Eye-tracking Data with an Analysis of Video Content from Free-viewing a Video of a Walk in an Urban Park Environment
Published on: May 7, 2019
07:35A Knowledge Graph Approach to Elucidate the Role of Organellar Pathways in Disease via Biomedical Reports
Published on: October 13, 2023
Related Concept Videos
Deductive Reasoning
For example, a researcher can deduce specific predictions...
Inductive Reasoning
Inductive reasoning is common in descriptive science. A life scientist makes observations and records them. This data can be qualitative or...
Reasoning
Inductive reasoning involves deriving generalizations from specific observations. This type of reasoning helps form beliefs about the world. For example,...
The Representativeness Heuristic
Reason and Intuition
Methods of Classification and Identification