Pith. sign in

Paper Citation Record · LEDGER

Mitigating Modality and Language-Style Gaps for Zero-Shot Video Moment Retrieval

As of 18 August 2026, this Paper Citation Record lists 52 of 52 outbound references and 0 inbound Pith citation observations for arXiv:2607.19027.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2607.19027 v1

Coverage vector

measured 52 of 52 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-15T15:36:43.741522Z

measured 52 of 52 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-18T06:34:40.430872+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

52 of 52 outbound references displayed

  • verified exact2
  • verified fuzzy22
  • unresolved27
  • parse uncertain0
  • malformed identifier1
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 4b0d7f98-8450-448d-8bc1-d41791981ea2 · outbound

This paper cites GPT-4 Technical Report.

Mitigating Modality and Language-Style Gaps for Zero-Shot Video Moment Retrieval GPT-4 Technical Report

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-15T15:36:43.520012Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T15:36:43.520012Z digest=sha256:131fb8824900558583386b88ec1f57a10fb98935cbd79dda99706cbeaae22cfb

Observation 81d764f2-fc39-496a-9437-9e93a34ea719 · outbound

This paper cites IEEE transactions on pattern analysis and machine intelligence38(10), 2069–2081 (2015).

Mitigating Modality and Language-Style Gaps for Zero-Shot Video Moment Retrieval IEEE transactions on pattern analysis and machine intelligence38(10), 2069–2081 (2015)

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T15:36:44.543613Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T15:36:43.525234Z digest=sha256:9fcc77d8c826f3b264ef82d71d0ed0bb1cc03ecc803006735f1764c267564226

Observation 1165ce58-a597-4f99-8d30-4b62dd0c23b1 · outbound

This paper cites In: Proceedings of the ieee conference on computer vision and pattern recognition.

Mitigating Modality and Language-Style Gaps for Zero-Shot Video Moment Retrieval In: Proceedings of the ieee conference on computer vision and pattern recognition

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-15T15:36:43.529664Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T15:36:43.529664Z digest=sha256:c5ff56a63de2e54f4a257b8e66be95c126be037f4a5a6a8b8617eadc8bd463f4

Observation fb53f648-76a4-4195-ae7d-2e50191e7823 · outbound

This paper cites In: Proceedings of the IEEE/CVF International Conference on Computer Vision.

Mitigating Modality and Language-Style Gaps for Zero-Shot Video Moment Retrieval In: Proceedings of the IEEE/CVF International Conference on Computer Vision

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T15:36:44.520261Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T15:36:43.534180Z digest=sha256:e9e0749783fa0eb7defd3331165faa32bac0dd30ce595507dac42e642e37ea5e

Observation 17eaca40-aa40-4c5c-a70b-b4cd4a5171bd · outbound

This paper cites MiniGPT-v2: large language model as a unified interface for vision-language multi-task learning.

Mitigating Modality and Language-Style Gaps for Zero-Shot Video Moment Retrieval MiniGPT-v2: large language model as a unified interface for vision-language multi-task learning

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-15T15:36:43.538839Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T15:36:43.538839Z digest=sha256:be40790cf68e3bfbad79ed436e14e76ee344a0dd49667f6ded90debc57663105

Observation 01479933-8141-4996-a352-b434b846aa2e · outbound

This paper cites MLLM Is a Strong Reranker: Advancing Multimodal Retrieval-augmented Generation via Knowledge-enhanced Reranking and Noise-injected Training.

Mitigating Modality and Language-Style Gaps for Zero-Shot Video Moment Retrieval MLLM Is a Strong Reranker: Advancing Multimodal Retrieval-augmented Generation via Knowledge-enhanced Reranking and Noise-injected Training

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-15T15:36:43.543626Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T15:36:43.543626Z digest=sha256:fac2d68c43cee0ed54a493fdb3d742fe1377acd8dfb48f983f7e68397f5324ad

Observation 06583ffb-0d9d-46c3-b6ca-6fc17ed7c199 · outbound

This paper cites In: Transfer Learning for Natural Language Processing Workshop.

Mitigating Modality and Language-Style Gaps for Zero-Shot Video Moment Retrieval In: Transfer Learning for Natural Language Processing Workshop

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T15:36:44.505789Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T15:36:43.548627Z digest=sha256:2db2cf727b8226c745e7eb7af5159231704a6ad323d080f87dffc1c2ba986608

Observation 7c20b64c-2e53-4795-8888-782c48fb8ef1 · outbound

This paper cites In: Proceedings of the IEEE international conference on computer vision.

Mitigating Modality and Language-Style Gaps for Zero-Shot Video Moment Retrieval In: Proceedings of the IEEE international conference on computer vision

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-15T15:36:43.552914Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T15:36:43.552914Z digest=sha256:5d8613a859b0d11cb73ab5ab1b58d55ced82ea898806972c109bac6f1f8bbc51

Observation 4ea5e12b-9a68-412a-a4d5-fa8852ae87d6 · outbound

This paper cites The Llama 3 Herd of Models.

Mitigating Modality and Language-Style Gaps for Zero-Shot Video Moment Retrieval The Llama 3 Herd of Models

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-15T15:36:43.557069Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T15:36:43.557069Z digest=sha256:1cc8dcdde7e7e020f8f5f1e6cc7d792adbcadfdf431bf8af931123c0945125cf

Observation 6d9bfb05-b27a-4576-80e1-12285a40d541 · outbound

This paper cites In: Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition.

Mitigating Modality and Language-Style Gaps for Zero-Shot Video Moment Retrieval In: Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-15T15:36:43.561208Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T15:36:43.561208Z digest=sha256:542c370ed885b5820d72a4822d839afbd3d758e329595d715612fb3ec5812ad9

Observation 8909095e-4af8-442b-b03b-3997bf2ecfd6 · outbound

This paper cites In: Companion Proceedings of the ACM on Web Conference 2025.

Mitigating Modality and Language-Style Gaps for Zero-Shot Video Moment Retrieval In: Companion Proceedings of the ACM on Web Conference 2025

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T15:36:44.473518Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T15:36:43.565415Z digest=sha256:a05ac65a05d40c8cef5849c00dc2b8ad2c179b47a594d5c58d7a3c156ec31cea

Observation 48fae83c-08db-4daa-a523-e5237b1a5391 · outbound

This paper cites In: Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition.

Mitigating Modality and Language-Style Gaps for Zero-Shot Video Moment Retrieval In: Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T15:36:44.459558Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T15:36:43.569869Z digest=sha256:02d4d5d73ae3b4c99d987d896c108eaf695af9a91eb2fb72e006990150097e71

Observation 2c6dac03-e002-4d76-ab60-2f5cf91eb4c3 · outbound

This paper cites In: Proceedings of the IEEE/CVF Winter Conference on Applications of Computer Vision.

Mitigating Modality and Language-Style Gaps for Zero-Shot Video Moment Retrieval In: Proceedings of the IEEE/CVF Winter Conference on Applications of Computer Vision

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T15:36:44.445734Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T15:36:43.573962Z digest=sha256:0cd1a84dbc0f35ce00f8053a0b3cd817f61dbc8d601e1cd0ba5d896a936e1519

Observation ff9366fa-c590-43f4-834c-ce9051dfbd9f · outbound

This paper cites In: Proceedings of the 31st ACM International Conference on Multimedia.

Mitigating Modality and Language-Style Gaps for Zero-Shot Video Moment Retrieval In: Proceedings of the 31st ACM International Conference on Multimedia

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T15:36:44.431462Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T15:36:43.578073Z digest=sha256:44cf96f5fefea80263ebac153c3e0643a2876544feed9ea70f993eeb7209f65e

Observation ccd48965-8dc0-41c5-bb6e-1418f31100f6 · outbound

This paper cites In: Proceedings of the IEEE international conference on computer vision.

Mitigating Modality and Language-Style Gaps for Zero-Shot Video Moment Retrieval In: Proceedings of the IEEE international conference on computer vision

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T15:36:44.417487Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T15:36:43.582240Z digest=sha256:15ef6566a5bdcc485ed3a847a9841eb02219657f150fa2f2932f19f1b68f90d6

Observation 38bc2419-bad6-47e1-8742-c7bf21e14dd1 · outbound

This paper cites NV-Embed: Improved Techniques for Training LLMs as Generalist Embedding Models.

Mitigating Modality and Language-Style Gaps for Zero-Shot Video Moment Retrieval NV-Embed: Improved Techniques for Training LLMs as Generalist Embedding Models

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-15T15:36:43.586285Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T15:36:43.586285Z digest=sha256:e23294114b1eb3a7877e298883b8aa0214d5bd3c8bd9d76140751e8e74ece3e7

Observation c1961ef8-7a70-4129-a569-b51d1ca93455 · outbound

This paper cites Advances in Neural Information Processing Systems34, 11846–11858 (2021).

Mitigating Modality and Language-Style Gaps for Zero-Shot Video Moment Retrieval Advances in Neural Information Processing Systems34, 11846–11858 (2021)

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-15T15:36:43.590519Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T15:36:43.590519Z digest=sha256:46c30aa85d6222ef84646a8c586188cac542a75507d4c7e0b03bde89140d757a

Observation 917bc589-0ea3-4707-8b4f-caada90ec9ff · outbound

This paper cites In: Computer Vision–ECCV 2020: 16th European Con- ference, Glasgow, UK, August 23–28, 2020, Proceedings, Part XXI 16.

Mitigating Modality and Language-Style Gaps for Zero-Shot Video Moment Retrieval In: Computer Vision–ECCV 2020: 16th European Con- ference, Glasgow, UK, August 23–28, 2020, Proceedings, Part XXI 16

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T15:36:44.394828Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T15:36:43.594794Z digest=sha256:a0d664239bcabeb585ed30eef45eccee4cc5d89183ae2e68c87430c518ba5c51

Observation 52157333-27b5-4100-be62-70d72b316826 · outbound

This paper cites In: International conference on machine learning.

Mitigating Modality and Language-Style Gaps for Zero-Shot Video Moment Retrieval In: International conference on machine learning

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-15T15:36:43.599067Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T15:36:43.599067Z digest=sha256:f83f529ec7ab0d8ee161625bd2d1d961c4ef97bd0f962e41c24f56f875ab7452

Observation 491fceb6-8aaf-4055-b35d-df1c1065415e · outbound

This paper cites In: International confer- ence on machine learning.

Mitigating Modality and Language-Style Gaps for Zero-Shot Video Moment Retrieval In: International confer- ence on machine learning

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-15T15:36:43.603476Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T15:36:43.603476Z digest=sha256:fa11750e3b3f10cdf2fb20d55d4f39c611f39cfdff3172955a2c05c10a8c9a29

Observation bec05775-74d0-4105-bb14-853c3bbb394c · outbound

This paper cites MM-Embed: Universal Multimodal Retrieval with Multimodal LLMs.

Mitigating Modality and Language-Style Gaps for Zero-Shot Video Moment Retrieval MM-Embed: Universal Multimodal Retrieval with Multimodal LLMs

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-15T15:36:43.607636Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T15:36:43.607636Z digest=sha256:74b5069c929ce2d36da3ef9a50547da97767f99eb6641615fafde8f4430c3835

Observation ba80e7f2-3d93-43b6-8b94-6f54381d4e28 · outbound

This paper cites an unresolved cited work.

Mitigating Modality and Language-Style Gaps for Zero-Shot Video Moment Retrieval Unresolved cited work

Reference 22

Resolution
unresolved
raw_fallback, observed 2026-08-15T15:36:44.361394Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T15:36:43.612074Z digest=sha256:18b9093ab754247bcf3d127b0b0103f335bd88bab5d328902bd091c77843aeba

Observation c0883dcb-bc3c-4531-9527-f9c764c976b7 · outbound

This paper cites In: Proceed- ings of the IEEE/CVF conference on computer vision and pattern recognition.

Mitigating Modality and Language-Style Gaps for Zero-Shot Video Moment Retrieval In: Proceed- ings of the IEEE/CVF conference on computer vision and pattern recognition

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T15:36:44.347859Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T15:36:43.616116Z digest=sha256:57f8c10efe6b6f25f97eb8c9cb0c214611a2287c4ba70148b15607588c9f35d1

Observation 88b171f6-72a1-433b-8110-f4c40231fccf · outbound

This paper cites In: Proceedings of the AAAI conference on artificial intelligence.

Mitigating Modality and Language-Style Gaps for Zero-Shot Video Moment Retrieval In: Proceedings of the AAAI conference on artificial intelligence

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T15:36:44.333718Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T15:36:43.620223Z digest=sha256:fd1b52f20af97da67158296078a6ab9259c6fb3e40d5d5e26fec9684babe2128

Observation 5f6eed93-67b6-4468-8950-b9d26f0881a7 · outbound

This paper cites In: Proceedings of the IEEE/CVF winter con- ference on applications of computer vision.

Mitigating Modality and Language-Style Gaps for Zero-Shot Video Moment Retrieval In: Proceedings of the IEEE/CVF winter con- ference on applications of computer vision

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T15:36:44.319867Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T15:36:43.624351Z digest=sha256:ce10a87d69263e50bd4e7266e888af20340e6da8e8130d32d95b25b6e968b90e

Observation 9dce1a8f-0d99-460e-a489-8804b779e0c6 · outbound

This paper cites In: Proceedings of the IEEE/CVF International Conference on Computer Vision.

Mitigating Modality and Language-Style Gaps for Zero-Shot Video Moment Retrieval In: Proceedings of the IEEE/CVF International Conference on Computer Vision

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T15:36:44.305513Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T15:36:43.628620Z digest=sha256:d800ad49f74483a1c6fcd8cb1f89060a64adc3d5133b1df7033e9127979c021b

Observation e7f2cc77-82d9-4959-8be6-98fcefb04bad · outbound

This paper cites In: Proceedings of the 62nd Annual Meeting of the Association for Computational Linguistics (Volume 1: Long Papers).

Mitigating Modality and Language-Style Gaps for Zero-Shot Video Moment Retrieval In: Proceedings of the 62nd Annual Meeting of the Association for Computational Linguistics (Volume 1: Long Papers)

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-15T15:36:43.632812Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T15:36:43.632812Z digest=sha256:dc1e76bda66eaafede5f330fe6f62b93f4ef7b6a9dc87e7ecf27dbae450498f1

Observation 985bc98a-ec0f-4a5f-8329-40dbe8391d95 · outbound

This paper cites Video-ChatGPT: Towards Detailed Video Understanding via Large Vision and Language Models.

Mitigating Modality and Language-Style Gaps for Zero-Shot Video Moment Retrieval Video-ChatGPT: Towards Detailed Video Understanding via Large Vision and Language Models

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-15T15:36:43.637093Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T15:36:43.637093Z digest=sha256:83f07f0da898b116beb8ef68e5211f2614f6a4e06677cdf15930645ecb3ca5b8

Observation 817683a9-2c3c-4ec2-af6a-c7b381425e6f · outbound

This paper cites In: Proceedings of the IEEE/CVF international conference on computer vision.

Mitigating Modality and Language-Style Gaps for Zero-Shot Video Moment Retrieval In: Proceedings of the IEEE/CVF international conference on computer vision

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T15:36:44.282770Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T15:36:43.641675Z digest=sha256:2e14531978533ba7f1aa394a3340a377d582a6dd58d24fed1fe9dcf79b4f5bb2

Observation 1b1bc65b-2b77-422d-ab35-f184b960cc8c · outbound

This paper cites DINOv2: Learning Robust Visual Features without Supervision.

Mitigating Modality and Language-Style Gaps for Zero-Shot Video Moment Retrieval DINOv2: Learning Robust Visual Features without Supervision

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-15T15:36:43.645788Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T15:36:43.645788Z digest=sha256:ad30458a94b3d169da6c6e8ae61d975a1e2ed8c413ef2a2a560ef80abc0ce1b7

Observation 964a9f15-e6f8-40b0-a61b-fd35d164912d · outbound

This paper cites Learning Transferable Visual Models From Natural Language Supervision.

Mitigating Modality and Language-Style Gaps for Zero-Shot Video Moment Retrieval Learning Transferable Visual Models From Natural Language Supervision

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-15T15:36:43.650352Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T15:36:43.650352Z digest=sha256:18478bcb8d31481b30275bfbcfe90c69ad3f3b9888dca0e1df5c618afb722553

Observation ee81fede-8b39-4421-959c-a910c1ad24e9 · outbound

This paper cites In: International conference on machine learning.

Mitigating Modality and Language-Style Gaps for Zero-Shot Video Moment Retrieval In: International conference on machine learning

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-15T15:36:43.654562Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T15:36:43.654562Z digest=sha256:337952b4c9ff274e1867b2ab7d59b5fe55c3dba92898758702d8d908f6e297a2

Observation 4070d86d-ac9d-4bbe-98b5-df4f2961ad72 · outbound

This paper cites Sentence-BERT: Sentence Embeddings using Siamese BERT-Networks.

Mitigating Modality and Language-Style Gaps for Zero-Shot Video Moment Retrieval Sentence-BERT: Sentence Embeddings using Siamese BERT-Networks

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-15T15:36:43.658646Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T15:36:43.658646Z digest=sha256:61409d554d221b404bcea74c0082cf5be236161c97e7c944eb39cadab1f6a086

Observation 3885910a-9c35-4e85-ba94-1a58afb362de · outbound

This paper cites In: Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition.

Mitigating Modality and Language-Style Gaps for Zero-Shot Video Moment Retrieval In: Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-15T15:36:43.663075Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T15:36:43.663075Z digest=sha256:c4f2abd38d5f37850b2a07c61fd238246c65182bfc254511d68273f901416895

Observation fe63f298-62a3-4347-b642-84d9989c3092 · outbound

This paper cites In: 2007 IEEE conference on computer vision and pattern recognition.

Mitigating Modality and Language-Style Gaps for Zero-Shot Video Moment Retrieval In: 2007 IEEE conference on computer vision and pattern recognition

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T15:36:44.249828Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T15:36:43.667296Z digest=sha256:5b16ec7044bf6fc5fc1f6eb7e8152066b669497f8ea36a18bda2f6c96daba471

Observation aeef787e-e90c-4f17-ba08-f84ee41798f6 · outbound

This paper cites Is ChatGPT Good at Search? Investigating Large Language Models as Re-Ranking Agents.

Mitigating Modality and Language-Style Gaps for Zero-Shot Video Moment Retrieval Is ChatGPT Good at Search? Investigating Large Language Models as Re-Ranking Agents

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-15T15:36:43.671742Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T15:36:43.671742Z digest=sha256:40fbb28f96b41a8756c19ac3708c80ecf08fdc200337c34818d2ef99c5bced74

Observation e7a378b7-d7a4-46da-8100-79540e1d8f22 · outbound

This paper cites IEEE Signal Process- ing Letters31, 521–525 (2023).

Mitigating Modality and Language-Style Gaps for Zero-Shot Video Moment Retrieval IEEE Signal Process- ing Letters31, 521–525 (2023)

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T15:36:44.235972Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T15:36:43.676102Z digest=sha256:7c04134a9b1f7338c386986be76937c6c1de8eb795f2bac1aef8d6db6fcc6078

Observation ba3447de-773c-4e87-a323-48b9c103cb43 · outbound

This paper cites LLaMA: Open and Efficient Foundation Language Models.

Mitigating Modality and Language-Style Gaps for Zero-Shot Video Moment Retrieval LLaMA: Open and Efficient Foundation Language Models

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-15T15:36:43.680326Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T15:36:43.680326Z digest=sha256:a13a503db2cc2e4f43ab8140147abef8d8e282e710d3a0fe1a525aa019fc9383

Observation fc148b43-b398-4672-b504-3d49e315d160 · outbound

This paper cites Llama 2: Open Foundation and Fine-Tuned Chat Models.

Mitigating Modality and Language-Style Gaps for Zero-Shot Video Moment Retrieval Llama 2: Open Foundation and Fine-Tuned Chat Models

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-15T15:36:43.684741Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T15:36:43.684741Z digest=sha256:0ccd60e681ebb78f571314987332b97b27365a3fedcb4439d3758916351505a0

Observation dde679b1-a34e-43b6-9aab-ef76ab99cb14 · outbound

This paper cites In: Proceedings of the 30th ACM international conference on multimedia.

Mitigating Modality and Language-Style Gaps for Zero-Shot Video Moment Retrieval In: Proceedings of the 30th ACM international conference on multimedia

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T15:36:44.221162Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T15:36:43.689108Z digest=sha256:f2ed5c620c854326bc84a14fc52dd19edb8719d3ae69b4b49db5c68d7113a1cc

Observation a7e7a4fc-8a83-44fa-9ed2-321d659e88c4 · outbound

This paper cites an unresolved cited work.

Mitigating Modality and Language-Style Gaps for Zero-Shot Video Moment Retrieval Unresolved cited work

Reference 41

Resolution
unresolved
raw_fallback, observed 2026-08-15T15:36:44.207118Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T15:36:43.693394Z digest=sha256:3fc9adbc16f021861809939cde1bedd742fd01b29b16e4d8dc88b33a5adf78c3

Observation 8c9e8033-36f6-40d0-bfc4-b6ebeef793ed · outbound

This paper cites an unresolved cited work.

Mitigating Modality and Language-Style Gaps for Zero-Shot Video Moment Retrieval Unresolved cited work

Reference 42

Resolution
unresolved
raw_fallback, observed 2026-08-15T15:36:44.193397Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T15:36:43.697813Z digest=sha256:586550a4123b8e3b92e566bffbe8f60198136c0b7928d1ceaaaa98a48c7ed508

Observation 18060491-b26b-40be-a32d-3d3f151e0036 · outbound

This paper cites arXiv preprint arXiv:2505.12499 (2025).

Mitigating Modality and Language-Style Gaps for Zero-Shot Video Moment Retrieval arXiv preprint arXiv:2505.12499 (2025)

Reference 43

Resolution
verified exact
raw_fallback, observed 2026-08-15T15:36:43.889388Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T15:36:43.702089Z digest=sha256:dc2d37bf24ca2d9ef5fc7254e52a5f054e5a896521d87a94f26282da60154bdb

Observation 21f00db2-dcf1-4262-a5cb-35ce818a5835 · outbound

This paper cites Applied Sciences14(5), 1894 (Feb 2024).https: //doi.org/10.3390/app14051894,http://dx.doi.org/10.3390/app14051894.

Mitigating Modality and Language-Style Gaps for Zero-Shot Video Moment Retrieval Applied Sciences14(5), 1894 (Feb 2024).https: //doi.org/10.3390/app14051894,http://dx.doi.org/10.3390/app14051894

Reference 44

Resolution
verified exact
doi, observed 2026-08-15T15:36:43.779138Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T15:36:43.706580Z digest=sha256:18a08ec948339e2d5f8d5beba87f23198aacac19700255d170165a63b43a9bd2

Observation e75cd50a-f651-475c-a2e1-5107e5c52926 · outbound

This paper cites In: 2024 International Joint Conference on Neural Networks (IJCNN).

Mitigating Modality and Language-Style Gaps for Zero-Shot Video Moment Retrieval In: 2024 International Joint Conference on Neural Networks (IJCNN)

Reference 45

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T15:36:44.179586Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T15:36:43.711058Z digest=sha256:1291afa310976e0559257084d476c40c823f66585584e55893bd97a32a6e6e4d

Observation 2da1ffdf-739f-4b8e-a8aa-7027768a9e1e · outbound

This paper cites In: Proceedings of the AAAI Conference on Artificial Intelligence.

Mitigating Modality and Language-Style Gaps for Zero-Shot Video Moment Retrieval In: Proceedings of the AAAI Conference on Artificial Intelligence

Reference 46

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T15:36:44.164682Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T15:36:43.715227Z digest=sha256:a4e79962e2f2f7a5b5e51f104df3cf486d1141dcf783ad96011edb251a79c3d5

Observation 3ef94fb4-f7ed-4b78-8323-5f9c59d4fde2 · outbound

This paper cites In: 2024 IEEE International Conference on Multimedia and Expo (ICME).

Mitigating Modality and Language-Style Gaps for Zero-Shot Video Moment Retrieval In: 2024 IEEE International Conference on Multimedia and Expo (ICME)

Reference 47

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T15:36:44.149098Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T15:36:43.719475Z digest=sha256:63c69ad9117fc484e405356f5b1c655b6286a0726463fd5ad79d416bbbe34554

Observation c16f0a09-0bf4-4e39-a3be-70e64d8afda2 · outbound

This paper cites an unresolved cited work.

Mitigating Modality and Language-Style Gaps for Zero-Shot Video Moment Retrieval Unresolved cited work

Reference 48

Resolution
unresolved
raw_fallback, observed 2026-08-15T15:36:44.134680Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T15:36:43.723769Z digest=sha256:333e19da11165dbe354be1587c58bad794c191aedce7ceb42c0b1b6f307a2ffd

Observation 186125b7-6443-4fcd-a5f3-ea2755705896 · outbound

This paper cites Video-LLaMA: An Instruction-tuned Audio-Visual Language Model for Video Understanding.

Mitigating Modality and Language-Style Gaps for Zero-Shot Video Moment Retrieval Video-LLaMA: An Instruction-tuned Audio-Visual Language Model for Video Understanding

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-15T15:36:43.728051Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T15:36:43.728051Z digest=sha256:965abe76a453d80ead1fd8b2d7efdd63b32a24f1b155c1dfc4dc678e2e5e1a14

Observation 898e4125-309c-427e-b042-bca895939769 · outbound

This paper cites In: European Conference on Com- puter Vision.

Mitigating Modality and Language-Style Gaps for Zero-Shot Video Moment Retrieval In: European Conference on Com- puter Vision

Reference 50

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T15:36:44.120202Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T15:36:43.732809Z digest=sha256:b9d1acde23a171b61ccfd68b928f125ea5de8eb2820850c513d599174ed2ebeb

Observation 64fdb272-ffe7-4fd1-a822-88bc0b59c2d4 · outbound

This paper cites In: Proceedings of the AAAI Conference on Artificial Intelligence.

Mitigating Modality and Language-Style Gaps for Zero-Shot Video Moment Retrieval In: Proceedings of the AAAI Conference on Artificial Intelligence

Reference 51

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T15:36:44.105783Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T15:36:43.737013Z digest=sha256:cd17d185fa08e1b2e72dfafe48fd81a1b783eefcd546951d7d7acc1e0730fe04

Observation 4ad58414-8e77-4558-b72f-573a31fd3bfa · outbound

This paper cites Yes” and “No.

Mitigating Modality and Language-Style Gaps for Zero-Shot Video Moment Retrieval Yes” and “No

Reference 52

Resolution
malformed identifier
raw_fallback, observed 2026-08-15T15:36:44.090595Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T15:36:43.741522Z digest=sha256:49f6f2ec4c17480ff94cc6f6af8b99dea6797b537dc0dfd721642a24e9bb834a

Pith citing papers

No inbound Pith citation observations are available.