TIGEr is a caption evaluation metric that uses a pre-trained text-to-image grounding model to compare candidate captions with human references via region rank and weight distribution similarity.
Title resolution pending
1 Pith paper cite this work. Polarity classification is still indexing.
1
Pith paper citing it
fields
cs.CL 1years
2019 1verdicts
CONDITIONAL 1representative citing papers
citing papers explorer
-
TIGEr: Text-to-Image Grounding for Image Caption Evaluation
TIGEr is a caption evaluation metric that uses a pre-trained text-to-image grounding model to compare candidate captions with human references via region rank and weight distribution similarity.