Revisiting Text-Based Person Retrieval: Mitigating Annotation-Induced Mismatches with Multimodal Large Language

Zihang Han1, Chao Zhu1, Mengyin Liu1

  • 1School of Computer and Communication Engineering, University of Science and Technology Beijing, Beijing 100083, China.

PubMed
Summary

Existing text-based person retrieval benchmarks contain ambiguous descriptions, leading to inaccurate model evaluations. This study introduces an annotation refinement framework using multimodal large language models to generate distinctive descriptions, improving benchmark quality and TBPR model performance.

Related Concept Videos

Improving Translational Accuracy02:07

Improving Translational Accuracy

Base complementarity between the three base pairs of mRNA codon and the tRNA anticodon is not a failsafe mechanism. Inaccuracies can range from a single mismatch to no correct base pairing at all. The free energy difference between the correct and nearly correct base pairs can be as small as 3 kcal/ mol. With complementarity being the only proofreading step, the estimated error frequency would be one wrong amino acid in every 100 amino acids incorporated. However, error frequencies observed in...
15.4K
Improving Translational Accuracy02:07

Improving Translational Accuracy

3.7K
Mismatch Repair01:36

Mismatch Repair

Overview
44.5K
Mismatch Repair01:20

Mismatch Repair

Organisms are capable of detecting and fixing nucleotide mismatches that occur during DNA replication. This sophisticated process requires identifying the new strand and replacing the erroneous bases with correct nucleotides. Mismatch repair is coordinated by many proteins in both prokaryotes and eukaryotes.
The Mutator Protein Family Plays a Key Role in DNA Mismatch Repair
The human genome has more than 3 billion base pairs of DNA per cell. Prior to cell division, that vast amount of genetic...
6.9K