Pith. sign in

Paper Citation Record · LEDGER

Domain Adaptation of VLM for Soccer Video Understanding

As of 21 August 2026, this Paper Citation Record lists 46 of 46 outbound references and 0 inbound Pith citation observations for arXiv:2505.13860.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2505.13860 v2

Coverage vector

measured 46 of 46 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-15T20:13:16.791596Z

measured 46 of 46 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-21T06:32:19.484+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

46 of 46 outbound references displayed

  • verified exact0
  • verified fuzzy34
  • unresolved10
  • parse uncertain1
  • malformed identifier0
  • metadata mismatch1

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 9beeeae2-54db-4532-b015-0e8d46aeac8f · outbound

This paper cites Claude 3.5 sonnet: Our best model yet,.

Domain Adaptation of VLM for Soccer Video Understanding Claude 3.5 sonnet: Our best model yet,

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T20:13:18.296907Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-15T20:13:16.433605Z digest=sha256:2dfaa0ae966577ed20ec8a8cda5f6211f37b5074348da29144754f176ddd71c6

Observation a79133e7-3081-4dd6-b1ab-b20c286532e9 · outbound

This paper cites MiniGPT4-Video: Advancing Multimodal LLMs for Video Understanding with Interleaved Visual-Textual Tokens.

Domain Adaptation of VLM for Soccer Video Understanding MiniGPT4-Video: Advancing Multimodal LLMs for Video Understanding with Interleaved Visual-Textual Tokens

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-15T20:13:16.445974Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T20:13:16.445974Z digest=sha256:fa4765f281455f7c99da572d5a7b4d95f649dc42dedd21305f63227e209c34ce

Observation e6aee8ba-e7d1-444e-87d7-bd7d0a56d8a0 · outbound

This paper cites An Introduction to Vision-Language Modeling.

Domain Adaptation of VLM for Soccer Video Understanding An Introduction to Vision-Language Modeling

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-15T20:13:16.452421Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T20:13:16.452421Z digest=sha256:25bf38dcd49707d6ae6ccca872474596ffc8c692e13af8da360945cee3eef5cc

Observation 4f3c502f-5eee-4485-9cc3-a292443794fa · outbound

This paper cites RT-2: Vision-Language-Action Models Transfer Web Knowledge to Robotic Control.

Domain Adaptation of VLM for Soccer Video Understanding RT-2: Vision-Language-Action Models Transfer Web Knowledge to Robotic Control

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-15T20:13:16.458159Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T20:13:16.458159Z digest=sha256:1822b0aef032aa0968b838349d87fdc6d6bd0fea0853928dcdb967bbd9d6542e

Observation 6d2e9682-404e-4cc0-a359-8aa7d9ff51b4 · outbound

This paper cites Soccernet-tracking: Multiple object tracking dataset and benchmark in soccer videos.

Domain Adaptation of VLM for Soccer Video Understanding Soccernet-tracking: Multiple object tracking dataset and benchmark in soccer videos

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T20:13:18.190912Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-15T20:13:16.464277Z digest=sha256:db818fde6a31f06bdc5392c2264861c4ffb2224ca22e680f5be283eaf3b3feb2

Observation a466e25a-479c-4a46-a3a4-ef9759b38bd5 · outbound

This paper cites Soccernet- v2: A dataset and benchmarks for holistic under- standing of broadcast soccer videos.

Domain Adaptation of VLM for Soccer Video Understanding Soccernet- v2: A dataset and benchmarks for holistic under- standing of broadcast soccer videos

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T20:13:18.075354Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-15T20:13:16.469431Z digest=sha256:f96ae8b0c7d1cb2b68403ccda41cc5eb122f5a7b8f469a5e047698e443b32c88

Observation ffa40242-5515-4de4-b52c-fb5a00dca55c · outbound

This paper cites Soccernet- v2: A dataset and benchmarks for holistic under- standing of broadcast soccer videos.

Domain Adaptation of VLM for Soccer Video Understanding Soccernet- v2: A dataset and benchmarks for holistic under- standing of broadcast soccer videos

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T20:13:18.055462Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-15T20:13:16.474940Z digest=sha256:6856e462f857ffdef3adadac8c4810a71dd91b527728f6deba536efb4481e22f

Observation d1175e54-adc5-4c04-a51c-796700b18d7e · outbound

This paper cites Mmbench-video: A long-form multi-shot bench- mark for holistic video understanding.

Domain Adaptation of VLM for Soccer Video Understanding Mmbench-video: A long-form multi-shot bench- mark for holistic video understanding

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T20:13:18.036643Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-15T20:13:16.479825Z digest=sha256:6421586bada820ba796365832ef4e42df151201895deae42566f4d4505440788

Observation 757d4228-ba92-4c9b-9f61-164983b90624 · outbound

This paper cites Soccernet- echoes: A soccer game audio commentary dataset.

Domain Adaptation of VLM for Soccer Video Understanding Soccernet- echoes: A soccer game audio commentary dataset

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T20:13:18.015999Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-15T20:13:16.485065Z digest=sha256:3d7595d6da0ac6dd23602f797994859245be636b632d3fe50cb557823e5fdceb

Observation 46bcd736-1893-41c4-ada6-11086348fca7 · outbound

This paper cites Soccernet: A scal- able dataset for action spotting in soccer videos.

Domain Adaptation of VLM for Soccer Video Understanding Soccernet: A scal- able dataset for action spotting in soccer videos

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T20:13:17.974672Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-15T20:13:16.490041Z digest=sha256:15edc273f72c97390f8c9d8915793c2540a81c3ee5bf3def26f18d174b0c68c7

Observation e656c400-8a5c-4697-8dca-129d52a5711b · outbound

This paper cites Fine- grained action recognition on a novel basketball dataset.

Domain Adaptation of VLM for Soccer Video Understanding Fine- grained action recognition on a novel basketball dataset

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T20:13:17.954687Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-15T20:13:16.495666Z digest=sha256:c36d7ec4d389986b34fd5fdbf9c7d6ed30efd28ae8509613597cbb6b60d4d47f

Observation 895f31fa-1f6a-4bb9-9ff6-4b44108dd8cb · outbound

This paper cites V ars: Video assistant referee sys- tem for automated soccer decision making from multiple views.

Domain Adaptation of VLM for Soccer Video Understanding V ars: Video assistant referee sys- tem for automated soccer decision making from multiple views

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T20:13:17.872670Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-15T20:13:16.500600Z digest=sha256:f15162b8bdef123974a281a0d7dba5f061d5256dee1f85ceaeb846e3e4c60c1d

Observation 63d296d0-fa26-4348-9bac-ca4b7580fc7f · outbound

This paper cites X-vars: Introducing explainability in foot- ball refereeing with multi-modal large language 9 models.

Domain Adaptation of VLM for Soccer Video Understanding X-vars: Introducing explainability in foot- ball refereeing with multi-modal large language 9 models

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T20:13:17.850989Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-15T20:13:16.507167Z digest=sha256:cb7af4d8fe63bbe3355513ae5dace9abfc263302986198ee69fce9d9d96f3cc6

Observation 8fd5e0aa-a4bd-47f5-88b7-cff797dc110d · outbound

This paper cites LoRA: Low-rank adaptation of large language models.

Domain Adaptation of VLM for Soccer Video Understanding LoRA: Low-rank adaptation of large language models

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T20:13:17.833018Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-15T20:13:16.513578Z digest=sha256:a35aef2e7e3c561a65f2099a3592360d1f4c97d81e500310134e7db5d8f58a8f

Observation f058eff5-c6d4-46c4-bc15-17cab4fb388f · outbound

This paper cites Vtimellm: Empower llm to grasp video moments.

Domain Adaptation of VLM for Soccer Video Understanding Vtimellm: Empower llm to grasp video moments

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T20:13:17.811858Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-15T20:13:16.518977Z digest=sha256:405348eb6c886c162780833bba566e47ef5805f1f48f9b05b4ba64a8bdccfafd

Observation 14299915-6d3b-47e4-95a7-57bbee4cd1c0 · outbound

This paper cites Wyscout: Football data and analytics plat- form, 2024.

Domain Adaptation of VLM for Soccer Video Understanding Wyscout: Football data and analytics plat- form, 2024

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T20:13:17.791261Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-15T20:13:16.524882Z digest=sha256:428707cfa00524f6b24a4d9479e0b89c5e301870a120945f3922f75015890327

Observation 8fe5d139-5613-4cfe-acf1-e50166895364 · outbound

This paper cites Quantifying the hierarchical scales of scientists'mobility.

Domain Adaptation of VLM for Soccer Video Understanding Quantifying the hierarchical scales of scientists'mobility

Reference 17

Resolution
metadata mismatch
local_arxiv, observed 2026-08-15T20:13:17.060941Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-15T20:13:16.531239Z digest=sha256:206546781e65bcb77d96a94f9462500c977a96b351f787f174a809e2fead8a1e

Observation d4e751e3-8eab-43d7-990b-5562ede86dbf · outbound

This paper cites Llava-med: Training a large language-and-vision assistant for biomedicine in one day.

Domain Adaptation of VLM for Soccer Video Understanding Llava-med: Training a large language-and-vision assistant for biomedicine in one day

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T20:13:17.772524Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-15T20:13:16.536940Z digest=sha256:6a3e062934ca776068f17e6bd73be2e83a2d4e956790230473c97749ca01f46d

Observation a8b6a607-722d-4ff6-af52-d1e5739be2e7 · outbound

This paper cites Sports-qa: A large-scale video ques- tion answering benchmark for complex and pro- fessional sports.

Domain Adaptation of VLM for Soccer Video Understanding Sports-qa: A large-scale video ques- tion answering benchmark for complex and pro- fessional sports

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-15T20:13:16.542060Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T20:13:16.542060Z digest=sha256:8aea9353c3625e76770274d9dc1ce5b3b60eec195bdb2668eee300834aafc39c

Observation 19adbdab-f186-4153-86e6-6b6d2454d6ae · outbound

This paper cites BLIP-2: Bootstrapping language-image pre- training with frozen image encoders and large lan- guage models.

Domain Adaptation of VLM for Soccer Video Understanding BLIP-2: Bootstrapping language-image pre- training with frozen image encoders and large lan- guage models

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T20:13:17.751145Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-15T20:13:16.547033Z digest=sha256:bc99760220dcdfbd52b1258484fabe808d2ec5442ee4bea060420b0c289ea0f4

Observation 2aa25aa6-a1ba-4d85-a56b-21536547c10a · outbound

This paper cites VisualBERT: A Simple and Performant Baseline for Vision and Language.

Domain Adaptation of VLM for Soccer Video Understanding VisualBERT: A Simple and Performant Baseline for Vision and Language

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-15T20:13:16.552939Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T20:13:16.552939Z digest=sha256:672d115845076b133f7559daa090ff63c0a4c9786e304821fa3dade9751369c2

Observation 0e950316-41db-41f2-abfc-dae8e1204151 · outbound

This paper cites Chain-of-region: Visual language models need details for diagram analysis.

Domain Adaptation of VLM for Soccer Video Understanding Chain-of-region: Visual language models need details for diagram analysis

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T20:13:17.678493Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-15T20:13:16.559862Z digest=sha256:3e5e57295bdc66528cd4eba02f59d93a781fc3102a58b46b20d361ef7c5c4804

Observation f8884262-657a-4ccc-b08e-7c1cd9e41d68 · outbound

This paper cites Video-LLaVA: Learning United Visual Representation by Alignment Before Projection.

Domain Adaptation of VLM for Soccer Video Understanding Video-LLaVA: Learning United Visual Representation by Alignment Before Projection

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-15T20:13:16.565851Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T20:13:16.565851Z digest=sha256:a970adf94d8ec119fc5ccbed3eb60d29e53b07159e047cfdf484289079945102

Observation abadc824-4261-438f-959a-551794d6132c · outbound

This paper cites Vila: On pre- training for visual language models.

Domain Adaptation of VLM for Soccer Video Understanding Vila: On pre- training for visual language models

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T20:13:17.638996Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-15T20:13:16.571679Z digest=sha256:41eca60819473e2cc48ef8ebf95a13a5c914d10c5c80f3e3a6e26e704932a4f1

Observation 62dcb978-69ef-453e-aef2-1ae79c680529 · outbound

This paper cites Llava-next: Improved reasoning, ocr, and world knowledge, 2024.

Domain Adaptation of VLM for Soccer Video Understanding Llava-next: Improved reasoning, ocr, and world knowledge, 2024

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T20:13:17.613751Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-15T20:13:16.577558Z digest=sha256:fba5d56555cdf43e1290aae50d76c40ef9ee1331d8eeaf767dec56e6a4784dfc

Observation 2c375a8c-2bc6-4001-9274-f209c5c26c4a · outbound

This paper cites Vilbert: Pretraining task-agnostic visiolin- guistic representations for vision-and-language tasks.

Domain Adaptation of VLM for Soccer Video Understanding Vilbert: Pretraining task-agnostic visiolin- guistic representations for vision-and-language tasks

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T20:13:17.591125Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-15T20:13:16.582754Z digest=sha256:ea3fd85ec7f4b48bf2806475c22c52409bfda1766fd23d7f8e592d8fc8bd36ad

Observation e8c114dd-c1a2-448a-afa3-e02652505516 · outbound

This paper cites Video-chatgpt: Towards detailed video understanding via large vi- sion and language models.

Domain Adaptation of VLM for Soccer Video Understanding Video-chatgpt: Towards detailed video understanding via large vi- sion and language models

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T20:13:17.570448Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-15T20:13:16.587829Z digest=sha256:ba25d4fd0f8218f2af8c8da9f6d6d25279df5072fc10050a25dc44472266e251

Observation bf9fe0f6-1560-45e8-b385-2959bca26098 · outbound

This paper cites Soccernet-caption: Dense video caption- ing for soccer broadcasts commentaries.

Domain Adaptation of VLM for Soccer Video Understanding Soccernet-caption: Dense video caption- ing for soccer broadcasts commentaries

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T20:13:17.539760Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-15T20:13:16.594577Z digest=sha256:2fa26e613a5b8376f7e46260ae8361506785b584f8897b0282b7b868f08c931b

Observation c76b2762-9fd5-496b-bc95-ad4e76a483f9 · outbound

This paper cites Med-flamingo: a multimodal medical few-shot learner.

Domain Adaptation of VLM for Soccer Video Understanding Med-flamingo: a multimodal medical few-shot learner

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T20:13:17.518346Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-15T20:13:16.600460Z digest=sha256:d5eb15f77f4f938c0d79e2f42de8c9154c224e958a57d14ff2154f0f8704d0ea

Observation 2595de97-26d0-4f5e-991b-3a2fb6d7e79b · outbound

This paper cites Sports video captioning via attentive mo- tion representation and group relationship model- ing.

Domain Adaptation of VLM for Soccer Video Understanding Sports video captioning via attentive mo- tion representation and group relationship model- ing

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T20:13:17.488273Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-15T20:13:16.606584Z digest=sha256:b04ede003304d341c789a32bcd14d270963fea134a6dbbbe2587fa6ea9c24151

Observation 8aa6558e-5230-433b-842f-0f1b39bd65c7 · outbound

This paper cites Learning transferable visual models from natural language supervision.

Domain Adaptation of VLM for Soccer Video Understanding Learning transferable visual models from natural language supervision

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-15T20:13:16.611650Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T20:13:16.611650Z digest=sha256:821057e097e7a5b0503e1ada9993ef259a3a87e40d88effb78a31834c39b272c

Observation e399b24e-c142-4b2d-a118-44307ab4719a · outbound

This paper cites Matchtime: Towards au- tomatic soccer game commentary generation.

Domain Adaptation of VLM for Soccer Video Understanding Matchtime: Towards au- tomatic soccer game commentary generation

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T20:13:17.441772Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-15T20:13:16.617004Z digest=sha256:25a9849eaef6044f0d4cf61e948b693e3f1d5602e1a98083c8fb31eb10d6d32b

Observation 87ebcc53-3fa0-4f62-9145-637873a6aef1 · outbound

This paper cites Multilingual vision-language pre-training for the remote sensing domain.

Domain Adaptation of VLM for Soccer Video Understanding Multilingual vision-language pre-training for the remote sensing domain

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T20:13:17.420149Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-15T20:13:16.626537Z digest=sha256:f4f8a44c8312b545f1bb5c29022733e070756cfdf5a739645aa88baf6dbab269

Observation 8a7a7515-572d-4359-9680-2033792993d8 · outbound

This paper cites Computer vi- sion for sports: Current applications and research topics.

Domain Adaptation of VLM for Soccer Video Understanding Computer vi- sion for sports: Current applications and research topics

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T20:13:17.389467Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-15T20:13:16.636461Z digest=sha256:49249af34cad2b5587325f7d44157e5380641f0350658ea152b1f4e202f171e2

Observation f375522b-fdcf-4b1d-9ca8-7d01ea1897c0 · outbound

This paper cites Semi-supervised training to improve player and ball detection in soccer.

Domain Adaptation of VLM for Soccer Video Understanding Semi-supervised training to improve player and ball detection in soccer

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T20:13:17.360718Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-15T20:13:16.650959Z digest=sha256:67dab9c4c21cc98eec80ab1bc7462ae7e38180880b1050785079a932cd048025

Observation 38235bfa-79fd-48f5-b046-42b0719e252b · outbound

This paper cites Knowledge Guided Entity-aware Video Captioning and A Basketball Benchmark.

Domain Adaptation of VLM for Soccer Video Understanding Knowledge Guided Entity-aware Video Captioning and A Basketball Benchmark

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-15T20:13:16.666462Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T20:13:16.666462Z digest=sha256:28da618be1a454745f98c1b562bdd7d3f4598e5da2449cc3e3b3ad757e1dfaae

Observation fa424fce-20e2-4778-b215-4bafeec169e0 · outbound

This paper cites Sportqa: A benchmark for sports under- standing in large language models.

Domain Adaptation of VLM for Soccer Video Understanding Sportqa: A benchmark for sports under- standing in large language models

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T20:13:17.338300Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-15T20:13:16.706521Z digest=sha256:b990db5c8d897bd45472d4b4a8a920f982febc5adaa2d53e914381ae43fc2220

Observation 214228e1-bdf0-4ff2-b81b-89415bee76f8 · outbound

This paper cites Fine- grained video captioning for sports narrative.

Domain Adaptation of VLM for Soccer Video Understanding Fine- grained video captioning for sports narrative

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T20:13:17.315919Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-15T20:13:16.747296Z digest=sha256:60ba4b559e136f4444e8e2ac20ea27ae7a571cda5345d09c9f25dbae4d85297d

Observation 57d4240b-068b-4eae-a248-8959c4a6e395 · outbound

This paper cites Merlot: Multimodal neural script knowledge models.

Domain Adaptation of VLM for Soccer Video Understanding Merlot: Multimodal neural script knowledge models

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T20:13:17.289548Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-15T20:13:16.753181Z digest=sha256:6919876c8ba1b3d935f41dde1e85771c23912e9cac41901e1cdd45190e8498e2

Observation 8f519505-f4dd-4a95-a329-dffe27421246 · outbound

This paper cites Video- llama: An instruction-tuned audio-visual language model for video understanding.

Domain Adaptation of VLM for Soccer Video Understanding Video- llama: An instruction-tuned audio-visual language model for video understanding

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T20:13:17.266310Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-15T20:13:16.759649Z digest=sha256:ac9d8bba57fd8da70b95f7ababcee00045650a658eb3ec4ce3135e6c3b5f6c70

Observation e3ef2d32-f2ef-492b-b66f-bb221bd539e6 · outbound

This paper cites Llava-next: A strong zero-shot video understanding model, 2024.

Domain Adaptation of VLM for Soccer Video Understanding Llava-next: A strong zero-shot video understanding model, 2024

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T20:13:17.240594Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-15T20:13:16.766670Z digest=sha256:3824457a8cb70fa76f50f74e53d5d98f60e06ea2e114079cdf7e0d2a2a9905be

Observation 69eaf598-9fb5-4281-849a-488602ec782a · outbound

This paper cites Feature Combination Meets Attention: Baidu Soccer Embeddings and Transformer based Temporal Detection.

Domain Adaptation of VLM for Soccer Video Understanding Feature Combination Meets Attention: Baidu Soccer Embeddings and Transformer based Temporal Detection

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-15T20:13:16.772602Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T20:13:16.772602Z digest=sha256:dd8df528cf017238548a6c16a2d3fadf8fd5458ceec62d11ac5433440c900ab1

Observation 2a997ea6-3cef-4f06-8d83-f5bd34a8de9f · outbound

This paper cites an unresolved cited work.

Domain Adaptation of VLM for Soccer Video Understanding Unresolved cited work

Reference 44

Resolution
unresolved
raw_fallback, observed 2026-08-15T20:13:17.218338Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-15T20:13:16.779796Z digest=sha256:f3d1c80b6cd1cb7529bcd9838a37b9ff06b487975556b8c77887e4268510acc7

Observation e5e475ba-101b-4999-adea-015722aaaa15 · outbound

This paper cites Dimension.

Domain Adaptation of VLM for Soccer Video Understanding Dimension

Reference 45

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T20:13:17.196398Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-15T20:13:16.785549Z digest=sha256:3ecb1c59ba20b7d57e33e91997ef7ddab243e2f49b23f5559b1172ca7c56cb16

Observation 5a6c6a42-64f5-48ec-96ac-aa4745fe7dd2 · outbound

This paper cites Given this, we only opted for Rank 64 when scaling up trainin g to the 20k datasets, as it provides greater capacity for handling complex patterns.

Domain Adaptation of VLM for Soccer Video Understanding Given this, we only opted for Rank 64 when scaling up trainin g to the 20k datasets, as it provides greater capacity for handling complex patterns

Reference 64

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T20:13:17.169882Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-15T20:13:16.791596Z digest=sha256:e03848223ef1dd749d52b709dac5624c00b79cc9caad54788890ab560f96bca0

Observation 96fc7ccc-8e83-4a1f-a7b2-446cc38166e7 · outbound

This paper cites an unresolved cited work.

Domain Adaptation of VLM for Soccer Video Understanding Unresolved cited work

Reference 2024

Resolution
parse uncertain
raw_fallback, observed 2026-08-15T20:13:18.262126Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-15T20:13:16.439669Z digest=sha256:3a2a5e75a105a847fc9b1afe6bb8ab22ebd66d72679823843933b9fbb83f7e4e

Pith citing papers

No inbound Pith citation observations are available.