Pith. sign in

Paper Citation Record · LEDGER

Player-Centric Multimodal Prompt Generation for Large Language Model Based Identity-Aware Basketball Video Captioning

As of 18 August 2026, this Paper Citation Record lists 58 of 58 outbound references and 1 inbound Pith citation observation for arXiv:2507.20163.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2507.20163 v1

Coverage vector

measured 58 of 58 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-06T13:47:53.840267Z

measured 59 of 59 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-18T06:34:40.430872+00:00

measured 1 of 1 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-03T15:55:21.629729Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

58 of 58 outbound references displayed

  • verified exact1
  • verified fuzzy49
  • unresolved8
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 615c7d59-3460-4dbd-ba72-8425a78f764c · outbound

This paper cites GPT-4 Technical Report.

Player-Centric Multimodal Prompt Generation for Large Language Model Based Identity-Aware Basketball Video Captioning GPT-4 Technical Report

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-06T13:47:48.338225Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T13:47:48.338225Z digest=sha256:a03c17930ca93c9b397d4301af11001b567f711b9f1fcc5e21355c44e9f78486

Observation c2c43baf-7352-4256-a8b6-28ee0082f827 · outbound

This paper cites Meteor: An automatic metric for mt evaluation with improved correlation with hu- man judgments.

Player-Centric Multimodal Prompt Generation for Large Language Model Based Identity-Aware Basketball Video Captioning Meteor: An automatic metric for mt evaluation with improved correlation with hu- man judgments

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:48:03.233657Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T13:47:48.387302Z digest=sha256:f39af64f6e022fd59fe27c9ca3a58d17d54d90c46e6323fbed2d40dd5fcc21a9

Observation fff18909-4ae3-4d86-8f4e-47e8ba2e2ad9 · outbound

This paper cites Is space-time attention all you need for video understanding? In ICML, page 4, 2021.

Player-Centric Multimodal Prompt Generation for Large Language Model Based Identity-Aware Basketball Video Captioning Is space-time attention all you need for video understanding? In ICML, page 4, 2021

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:48:03.098776Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T13:47:48.471255Z digest=sha256:34d9a893ef2656e9020a9d459cc55e5a1ae928d45f33eddff68681bf4f3d6418

Observation 411cb742-a6d4-41c2-babd-432823cd6856 · outbound

This paper cites Activitynet: A large-scale video benchmark for human activity understanding.

Player-Centric Multimodal Prompt Generation for Large Language Model Based Identity-Aware Basketball Video Captioning Activitynet: A large-scale video benchmark for human activity understanding

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:48:02.909167Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T13:47:48.557375Z digest=sha256:2fda52042959e7c989927984df39b8abe10c87c6f786624e04c4bb1523591726

Observation 0ec81a08-6cf2-43e0-8e5d-3dadb61c85c3 · outbound

This paper cites Quo vadis, action recognition? a new model and the kinetics dataset.

Player-Centric Multimodal Prompt Generation for Large Language Model Based Identity-Aware Basketball Video Captioning Quo vadis, action recognition? a new model and the kinetics dataset

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:48:02.760554Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T13:47:48.651033Z digest=sha256:5b1d468d18353e71d1646dba4ff526ecd97d98a19deb1ef8aafc3689610b22f6

Observation 66722b3d-12f4-41f1-9a81-a015d7c49748 · outbound

This paper cites Collecting highly paral- lel data for paraphrase evaluation.

Player-Centric Multimodal Prompt Generation for Large Language Model Based Identity-Aware Basketball Video Captioning Collecting highly paral- lel data for paraphrase evaluation

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:48:02.593360Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T13:47:48.689591Z digest=sha256:eb372fb63a14ed83021d6a3c3a8da1883f22cdee19a34b80ffea280dd3288f20

Observation 059320fb-9413-43d9-9774-722ab4638539 · outbound

This paper cites Learning Phrase Representations using RNN Encoder-Decoder for Statistical Machine Translation.

Player-Centric Multimodal Prompt Generation for Large Language Model Based Identity-Aware Basketball Video Captioning Learning Phrase Representations using RNN Encoder-Decoder for Statistical Machine Translation

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-06T13:47:48.721206Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T13:47:48.721206Z digest=sha256:e1ac8ab80d0e63090de2c1c7b80ebcc24ae50ba6f9847a018da8683b6605a0a8

Observation e749d5fc-3a1c-493e-90f6-d4cdb1a3f469 · outbound

This paper cites Sportsmot: A large multi-object tracking dataset in multiple sports scenes.

Player-Centric Multimodal Prompt Generation for Large Language Model Based Identity-Aware Basketball Video Captioning Sportsmot: A large multi-object tracking dataset in multiple sports scenes

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:48:02.444435Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T13:47:48.739135Z digest=sha256:660a022762c92e446fb82953ded2cf7f5585bc43a419a30e680fa46d46936910

Observation 3ddf3f16-8a17-4b3c-9d57-bac93c3d0a91 · outbound

This paper cites A thou- sand frames in just a few words: Lingual description of videos through latent topics and sparse object stitching.

Player-Centric Multimodal Prompt Generation for Large Language Model Based Identity-Aware Basketball Video Captioning A thou- sand frames in just a few words: Lingual description of videos through latent topics and sparse object stitching

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:48:02.255384Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T13:47:48.785487Z digest=sha256:d91c1c3f4a2aa6adae995780639d885938c32a4d549bb56caaca24cfabc79100

Observation 1ba5df05-4642-4508-a44d-c8cd9c866ffd · outbound

This paper cites An Image is Worth 16x16 Words: Transformers for Image Recognition at Scale.

Player-Centric Multimodal Prompt Generation for Large Language Model Based Identity-Aware Basketball Video Captioning An Image is Worth 16x16 Words: Transformers for Image Recognition at Scale

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-06T13:47:48.831059Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T13:47:48.831059Z digest=sha256:744d9d1a1adb0837f8e0ee2138a9690011ae849f9e23c32cecf3657934933452

Observation b0c30544-5258-410e-8fa8-503210534ff5 · outbound

This paper cites Soccer captioning: dataset, transformer-based model, and triple-level evaluation.

Player-Centric Multimodal Prompt Generation for Large Language Model Based Identity-Aware Basketball Video Captioning Soccer captioning: dataset, transformer-based model, and triple-level evaluation

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:48:02.121089Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T13:47:48.890391Z digest=sha256:2bf0dd6d53bb86336fe754fe3dc52bfe330d9b0b95ff4d38d82d8eb9be1938ba

Observation 28674be7-f982-4455-aefa-ff76dfeaade9 · outbound

This paper cites an unresolved cited work.

Player-Centric Multimodal Prompt Generation for Large Language Model Based Identity-Aware Basketball Video Captioning Unresolved cited work

Reference 12

Resolution
unresolved
raw_fallback, observed 2026-08-06T13:48:01.960339Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T13:47:48.942982Z digest=sha256:8b9e9e8b0bb8e1576a6bacc6411e7b91da2a2b6b57c334fedc71f7a73e9d474e

Observation ca1ec4e3-4f73-457f-bbb0-82bcc2a39c01 · outbound

This paper cites Deep residual learning for image recognition.

Player-Centric Multimodal Prompt Generation for Large Language Model Based Identity-Aware Basketball Video Captioning Deep residual learning for image recognition

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:48:01.789107Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T13:47:48.982219Z digest=sha256:114807d870268ef07ea4c6b98e3acbdf4fb147061b3c5491951df4f93300b935

Observation f9ab540d-86f4-4e49-82d5-b22887b02b66 · outbound

This paper cites Gaussian Error Linear Units (GELUs).

Player-Centric Multimodal Prompt Generation for Large Language Model Based Identity-Aware Basketball Video Captioning Gaussian Error Linear Units (GELUs)

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-06T13:47:49.051493Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T13:47:49.051493Z digest=sha256:e1fee6e9c70eec334039056414f2c38d99e021018bf9e94aafbcecfa8de55b87

Observation 830bdf3f-bf20-47c4-b106-f1922277feca · outbound

This paper cites Overview of temporal action detection based on deep learning.

Player-Centric Multimodal Prompt Generation for Large Language Model Based Identity-Aware Basketball Video Captioning Overview of temporal action detection based on deep learning

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:48:01.578230Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T13:47:49.111293Z digest=sha256:d398ee853c54dacf765dacaa0c8bcd1127af6e209bdb5d4a472be5f2cb7ee0cf

Observation 4d3642e8-9f19-4e66-afe3-f34de4e784d0 · outbound

This paper cites Learn- ing to generate move-by-move commentary for chess games from large-scale social forum data.

Player-Centric Multimodal Prompt Generation for Large Language Model Based Identity-Aware Basketball Video Captioning Learn- ing to generate move-by-move commentary for chess games from large-scale social forum data

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:48:01.371450Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T13:47:49.229047Z digest=sha256:e8f1c7582c105d4ec51693f1c962944be1cb6fa9f83b7c92f7832ed5bfc9cb46

Observation 9b0d2555-6dd6-4586-bb63-a0dd0deded6e · outbound

This paper cites Learning seg- ment similarity and alignment in large-scale content based video retrieval.

Player-Centric Multimodal Prompt Generation for Large Language Model Based Identity-Aware Basketball Video Captioning Learning seg- ment similarity and alignment in large-scale content based video retrieval

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:48:01.166288Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T13:47:49.383887Z digest=sha256:f4c208e8a25cdda724346594db2dceea4cfb4daa08328ae18dc870abeb175bba

Observation 9163806d-f645-4003-a9b9-4116062688c1 · outbound

This paper cites Automatic baseball commentary generation using deep learning.

Player-Centric Multimodal Prompt Generation for Large Language Model Based Identity-Aware Basketball Video Captioning Automatic baseball commentary generation using deep learning

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:48:01.005079Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T13:47:49.465481Z digest=sha256:50174c957e8d82f5e6dbd493abdaf61ee55c32283a369550bc4504bb123d5af4

Observation 8d357874-4819-4675-b8b5-8f6966fb7123 · outbound

This paper cites Adam: A Method for Stochastic Optimization.

Player-Centric Multimodal Prompt Generation for Large Language Model Based Identity-Aware Basketball Video Captioning Adam: A Method for Stochastic Optimization

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-06T13:47:49.583990Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T13:47:49.583990Z digest=sha256:6256799ede79949d8c681156d192429fdc8071d09ee79e5c7b07099e57a62e27

Observation 639bf675-0666-4cec-92d7-0ac52890f53c · outbound

This paper cites Video story- telling: Textual summaries for events.

Player-Centric Multimodal Prompt Generation for Large Language Model Based Identity-Aware Basketball Video Captioning Video story- telling: Textual summaries for events

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:48:00.869574Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T13:47:49.829969Z digest=sha256:a7ee35aba8acfa3b3a1347e724552c9115064490c430d204baacba82cf0fa7a8

Observation 3550829e-2298-4d99-8af8-0e6f3c1021da · outbound

This paper cites Rouge: A package for automatic evaluation of summaries.

Player-Centric Multimodal Prompt Generation for Large Language Model Based Identity-Aware Basketball Video Captioning Rouge: A package for automatic evaluation of summaries

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:48:00.680689Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T13:47:49.952635Z digest=sha256:276a2ba909f498285da2e1eee389ef6ec0413b1ffcfcbf88650b1a273d3c4611

Observation 717f7a0e-aa5a-4888-8221-67085a8a4267 · outbound

This paper cites Swinbert: End- to-end transformers with sparse attention for video caption- ing.

Player-Centric Multimodal Prompt Generation for Large Language Model Based Identity-Aware Basketball Video Captioning Swinbert: End- to-end transformers with sparse attention for video caption- ing

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:48:00.501771Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T13:47:50.083025Z digest=sha256:f61267b87eabb775316d8976151a4b21c75f029a1c8871d62d7bd064d35a8e01

Observation fd23068d-2a81-4dd3-a882-e03f26b28b21 · outbound

This paper cites Video swin transformer.

Player-Centric Multimodal Prompt Generation for Large Language Model Based Identity-Aware Basketball Video Captioning Video swin transformer

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:48:00.299165Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T13:47:50.245337Z digest=sha256:5996caa3a978fe778dc2f54b29aad57d3ad0a5269f9056ee9cf6ea83b6ed499a

Observation bce3b90b-dbca-4d5a-91c8-665ec89a251a · outbound

This paper cites Llama 3.2 quantized models, 2024.

Player-Centric Multimodal Prompt Generation for Large Language Model Based Identity-Aware Basketball Video Captioning Llama 3.2 quantized models, 2024

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:48:00.097823Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T13:47:50.372060Z digest=sha256:f8a5f9bd343a1ed2c681937141b08e7ee3446e84b972e417ebf38e0ed9277e50

Observation ccd780b4-4273-4abd-a016-7267ffcf0f44 · outbound

This paper cites Soccernet- caption: Dense video captioning for soccer broadcasts com- mentaries.

Player-Centric Multimodal Prompt Generation for Large Language Model Based Identity-Aware Basketball Video Captioning Soccernet- caption: Dense video captioning for soccer broadcasts com- mentaries

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:47:59.854049Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T13:47:50.522585Z digest=sha256:955f11bf1d2aedb45229e83d170e14b309615f76878e9f6052e3f1beda21f4fc

Observation 0fe48295-4fb5-4b2d-add4-c27f80b8e0b0 · outbound

This paper cites Search-oriented micro-video captioning.

Player-Centric Multimodal Prompt Generation for Large Language Model Based Identity-Aware Basketball Video Captioning Search-oriented micro-video captioning

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:47:59.618980Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T13:47:50.650288Z digest=sha256:b2384e94c0df6a23527b1e7800279b6018b4572e4b27345802a410b3307d9b76

Observation b68e4fb3-067f-4dec-a121-f022be7d66ce · outbound

This paper cites Enhancing visual question answering through question-driven image captions as prompts.

Player-Centric Multimodal Prompt Generation for Large Language Model Based Identity-Aware Basketball Video Captioning Enhancing visual question answering through question-driven image captions as prompts

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:47:59.373041Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T13:47:50.783054Z digest=sha256:58262d5c4c6de652c5e1ce6acd42eab84034261b370eb5fe12d283239cfb6d8d

Observation 4a3e2810-37bb-4324-8303-65485346603d · outbound

This paper cites Bleu: a method for automatic evaluation of machine translation.

Player-Centric Multimodal Prompt Generation for Large Language Model Based Identity-Aware Basketball Video Captioning Bleu: a method for automatic evaluation of machine translation

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:47:59.134945Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T13:47:51.004697Z digest=sha256:6fbe7bab7d16cf5ebf2f2160bffb026bcfd61de7c047582f9a1482105cb98dd9

Observation b53fca66-fa12-430d-8a6e-7451aac9892a · outbound

This paper cites Identity- aware multi-sentence video description.

Player-Centric Multimodal Prompt Generation for Large Language Model Based Identity-Aware Basketball Video Captioning Identity- aware multi-sentence video description

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:47:58.894291Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T13:47:51.226089Z digest=sha256:1e4eebc65a130c049d610833dd2d67e8a03893fc05abd74bb1a20fe85e2e7116

Observation 37534dc4-5bcd-4462-9b01-fe8af62f296d · outbound

This paper cites Goal: A challenging knowledge-grounded video captioning benchmark for real- time soccer commentary generation.

Player-Centric Multimodal Prompt Generation for Large Language Model Based Identity-Aware Basketball Video Captioning Goal: A challenging knowledge-grounded video captioning benchmark for real- time soccer commentary generation

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:47:58.654739Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T13:47:51.400669Z digest=sha256:bb58a064cc738fa7f8b94173b1956d975c48a07d03386b8bbca674c36c65562b

Observation 77af1dcb-5943-4b4e-96b4-73de32742b74 · outbound

This paper cites Sports video captioning via attentive motion representation and group re- lationship modeling.

Player-Centric Multimodal Prompt Generation for Large Language Model Based Identity-Aware Basketball Video Captioning Sports video captioning via attentive motion representation and group re- lationship modeling

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:47:58.457349Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T13:47:51.533250Z digest=sha256:8ff63925caf5ac985e48b3f9c18c18e55cd561f9ea1995c25894073f226cdf81

Observation f7630cbe-3fa3-4c7e-b011-dcb167702c73 · outbound

This paper cites Language models are unsupervised multitask learners.

Player-Centric Multimodal Prompt Generation for Large Language Model Based Identity-Aware Basketball Video Captioning Language models are unsupervised multitask learners

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:47:58.256585Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T13:47:51.721508Z digest=sha256:15d7c0623d26f1926268a6dd002320b71b77eaba98561dcf3c1e8034f760572b

Observation 5fbd5acf-adb7-4f10-9ac1-637ecc7db084 · outbound

This paper cites Learn- ing transferable visual models from natural language super- vision.

Player-Centric Multimodal Prompt Generation for Large Language Model Based Identity-Aware Basketball Video Captioning Learn- ing transferable visual models from natural language super- vision

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:47:58.105707Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T13:47:51.901836Z digest=sha256:50321d12ee24ef655f51a7d9b7816c46d66d184d7037b5674b6e89b61fa1b5bd

Observation 17fcd46c-ce79-4d66-9dc7-ee5dda44c2ee · outbound

This paper cites MatchTime: Towards Automatic Soccer Game Commentary Generation.

Player-Centric Multimodal Prompt Generation for Large Language Model Based Identity-Aware Basketball Video Captioning MatchTime: Towards Automatic Soccer Game Commentary Generation

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-06T13:47:52.050904Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T13:47:52.050904Z digest=sha256:7be4cbe20f4c9f72b3a373a8973b0e0ace741a7a010b7dfd172810d08b154e4b

Observation 9ed940e0-1380-4904-adcc-1aaf416794ad · outbound

This paper cites Towards universal soccer video under- standing.

Player-Centric Multimodal Prompt Generation for Large Language Model Based Identity-Aware Basketball Video Captioning Towards universal soccer video under- standing

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:47:57.935512Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T13:47:52.195571Z digest=sha256:40ce1c89a1a50d737ff4046b416082a545ea7960935112ff9a09dbbcfc61acc8

Observation c026b22e-20c4-4f7e-a4c7-94e13b78a84a · outbound

This paper cites Grounding action descriptions in videos.

Player-Centric Multimodal Prompt Generation for Large Language Model Based Identity-Aware Basketball Video Captioning Grounding action descriptions in videos

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:47:57.685758Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T13:47:52.288758Z digest=sha256:827865077a7a7422667c859d63c8ec36e6e994e53f5eba65ac6740316dde5b65

Observation dd8cd761-f1a1-428d-ac77-a9f617db746e · outbound

This paper cites Timechat: A time-sensitive multimodal large language model for long video understanding.

Player-Centric Multimodal Prompt Generation for Large Language Model Based Identity-Aware Basketball Video Captioning Timechat: A time-sensitive multimodal large language model for long video understanding

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:47:57.497245Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T13:47:52.384320Z digest=sha256:af43670a6163507287447df79f3b3cec728f76bca278a69dc159d7572e00afe3

Observation cccc58cf-1f89-4681-9e80-55edcc540cee · outbound

This paper cites A dataset for movie description.

Player-Centric Multimodal Prompt Generation for Large Language Model Based Identity-Aware Basketball Video Captioning A dataset for movie description

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:47:57.261849Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T13:47:52.476419Z digest=sha256:df95a19407c0177d9c6acf8b4b15dea589b932ed92f1d4ead4a236d948aa1c7d

Observation b1a26b85-86e0-4762-87ba-0b6a957a1e6c · outbound

This paper cites Accurate and fast compressed video captioning.

Player-Centric Multimodal Prompt Generation for Large Language Model Based Identity-Aware Basketball Video Captioning Accurate and fast compressed video captioning

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:47:57.090318Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T13:47:52.570908Z digest=sha256:20d16614d147de62736ac961b9f0d1de38932b463fbf3488428256db436d9e71

Observation 358d8369-dea1-4109-938e-9adea259b196 · outbound

This paper cites An overview of the tesseract ocr engine.

Player-Centric Multimodal Prompt Generation for Large Language Model Based Identity-Aware Basketball Video Captioning An overview of the tesseract ocr engine

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:47:56.929106Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T13:47:52.629207Z digest=sha256:72484dc48e63698b1dbb18011e2734fa983047c42bcf9f733d82c905ba6c8087

Observation f79916d6-9ca0-4e4f-abeb-d43dac9275c1 · outbound

This paper cites Clip4caption: Clip for video caption.

Player-Centric Multimodal Prompt Generation for Large Language Model Based Identity-Aware Basketball Video Captioning Clip4caption: Clip for video caption

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:47:56.762986Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T13:47:52.681290Z digest=sha256:1cbd4896d61235aa7f59e6696a901e98c844bbae768fd047a8dfd5a2a3f6274c

Observation 37b257fc-62a8-4905-8fe4-d1b9b93c9c43 · outbound

This paper cites Qwen2.5: A party of foundation models, 2024.

Player-Centric Multimodal Prompt Generation for Large Language Model Based Identity-Aware Basketball Video Captioning Qwen2.5: A party of foundation models, 2024

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:47:56.634122Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T13:47:52.753771Z digest=sha256:47d5a6b8f1cd699cbe375e360d39a6872a802222189a8c9c16d28e99d1fa1a6b

Observation 6ce10f36-56a2-4295-816e-cb5b157b3675 · outbound

This paper cites Llama 2: Open Foundation and Fine-Tuned Chat Models.

Player-Centric Multimodal Prompt Generation for Large Language Model Based Identity-Aware Basketball Video Captioning Llama 2: Open Foundation and Fine-Tuned Chat Models

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-06T13:47:52.792329Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T13:47:52.792329Z digest=sha256:d29bbf2868ae83466cc837edff7148c1c1ddc6aaa3f88234c6ebac4d0a5de3fc

Observation f50975b2-6527-41d1-b5bb-58210b190ba9 · outbound

This paper cites Atten- tion is all you need.

Player-Centric Multimodal Prompt Generation for Large Language Model Based Identity-Aware Basketball Video Captioning Atten- tion is all you need

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:47:56.470095Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T13:47:52.859030Z digest=sha256:348387f58f2b5485fe8b60473c4cdfc3ce497b68399269957ecf08dc8fc2b8f2

Observation 1c967b64-34d4-4209-935d-c7c83bc22406 · outbound

This paper cites Player tracking and identification in ice hockey.Expert systems with applications, 213:119250, 2023.

Player-Centric Multimodal Prompt Generation for Large Language Model Based Identity-Aware Basketball Video Captioning Player tracking and identification in ice hockey.Expert systems with applications, 213:119250, 2023

Reference 45

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:47:56.298252Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T13:47:52.919724Z digest=sha256:460392ef4c6648f945056e6a3f9ab2b37e1f86287ad7b1414cfd45d86d6cfdfd

Observation 98ad022a-2a58-4206-80ae-5b22fdb2ef3a · outbound

This paper cites Cider: Consensus-based image description evalua- tion.

Player-Centric Multimodal Prompt Generation for Large Language Model Based Identity-Aware Basketball Video Captioning Cider: Consensus-based image description evalua- tion

Reference 46

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:47:56.092821Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T13:47:52.975748Z digest=sha256:8e5d5ca386fc23b9b58b0f59fee5e5ba4e25f4b89f3b54239e883c3138b5fc86

Observation 265d17ab-2956-4a69-85f1-c0a1c0511237 · outbound

This paper cites Omnivid: A generative framework for universal video understanding.

Player-Centric Multimodal Prompt Generation for Large Language Model Based Identity-Aware Basketball Video Captioning Omnivid: A generative framework for universal video understanding

Reference 47

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:47:55.936975Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T13:47:53.023439Z digest=sha256:4c2df6d7d113aedfd586e77ea8f6a8a7d3d907d0a2d80d59dbf5772ef950074e

Observation 371e9520-823e-41f1-aa81-b12c4ac8fd27 · outbound

This paper cites Sports video anal- ysis on large-scale data.

Player-Centric Multimodal Prompt Generation for Large Language Model Based Identity-Aware Basketball Video Captioning Sports video anal- ysis on large-scale data

Reference 48

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:47:55.808364Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T13:47:53.059771Z digest=sha256:7e08b4a9f23550fa0500cc6c4e5d6a32220b8002c0e7e94f1b31a8af67f3f64a

Observation 4e6ee851-9abe-4eab-8126-257f9d148816 · outbound

This paper cites Learning label semantics for weakly supervised group activity recognition.

Player-Centric Multimodal Prompt Generation for Large Language Model Based Identity-Aware Basketball Video Captioning Learning label semantics for weakly supervised group activity recognition

Reference 49

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:47:55.644024Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T13:47:53.106462Z digest=sha256:a3afdbd0652fc4e7f3370e0f96d20076d9c0175b70a6c822dbd93f0cbe9c86f6

Observation d8c8a3ac-9f04-42ef-b5ff-18212b579d07 · outbound

This paper cites A simple yet effective knowledge guided method for entity-aware video captioning on a basketball benchmark.

Player-Centric Multimodal Prompt Generation for Large Language Model Based Identity-Aware Basketball Video Captioning A simple yet effective knowledge guided method for entity-aware video captioning on a basketball benchmark

Reference 50

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:47:55.465434Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T13:47:53.202516Z digest=sha256:c6ebcfb5775d2c33761f958d241a63088e6fba95273b2bacdd6955865a58ad71

Observation c9368bf4-c517-4bea-9e43-b0b69e46abe0 · outbound

This paper cites Eika: Explicit & im- plicit knowledge-augmented network for entity-aware sports video captioning.

Player-Centric Multimodal Prompt Generation for Large Language Model Based Identity-Aware Basketball Video Captioning Eika: Explicit & im- plicit knowledge-augmented network for entity-aware sports video captioning

Reference 51

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:47:55.304502Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T13:47:53.279524Z digest=sha256:f8ca53dfdc3981f7f87f7be275acc26bf1f410afb29f9fb996c5d7db1a63f286

Observation 93235d0d-b85a-4a05-b083-928108e89e78 · outbound

This paper cites Msr-vtt: A large video description dataset for bridging video and language.

Player-Centric Multimodal Prompt Generation for Large Language Model Based Identity-Aware Basketball Video Captioning Msr-vtt: A large video description dataset for bridging video and language

Reference 52

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:47:55.163118Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T13:47:53.363030Z digest=sha256:43bb171eae9e98e4f53513933bb269db7f412819a33c4968c74ef88d6899f54b

Observation 66f835c6-e6b0-4a1e-9771-1efe8ff04925 · outbound

This paper cites Hierarchical modular network for video captioning.

Player-Centric Multimodal Prompt Generation for Large Language Model Based Identity-Aware Basketball Video Captioning Hierarchical modular network for video captioning

Reference 53

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:47:54.964288Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T13:47:53.424019Z digest=sha256:e97cdaf895830fc94f887555e7127556f48488696d22eb503525fb043cbab551

Observation 4357c369-9a65-48c0-9f0f-736c24d4eb78 · outbound

This paper cites Fine-grained video captioning for sports narrative.

Player-Centric Multimodal Prompt Generation for Large Language Model Based Identity-Aware Basketball Video Captioning Fine-grained video captioning for sports narrative

Reference 54

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:47:54.829046Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T13:47:53.530146Z digest=sha256:856b5a4414e837ce27da80fc2113a2b25296d0c89766b028c4a42ce8366b734b

Observation 0fe3ff17-e83f-4ba0-9c7d-568a37f848fb · outbound

This paper cites Movie101: A New Movie Understanding Benchmark.

Player-Centric Multimodal Prompt Generation for Large Language Model Based Identity-Aware Basketball Video Captioning Movie101: A New Movie Understanding Benchmark

Reference 55

Resolution
verified exact
local_arxiv, observed 2026-08-06T13:47:54.028783Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T13:47:53.617274Z digest=sha256:29627655e2a79f57602e3399cd1f2db111043f746d4d78e6781b2308be78ed03

Observation c7834a93-67fc-4937-83c0-de9ac951366d · outbound

This paper cites Harnessing large language models for training-free video anomaly detection.

Player-Centric Multimodal Prompt Generation for Large Language Model Based Identity-Aware Basketball Video Captioning Harnessing large language models for training-free video anomaly detection

Reference 56

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:47:54.711821Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T13:47:53.702448Z digest=sha256:355dfd583b309fe7f33b42d6382feae9e0409c83eb687222981f5b11f0c3f851

Observation ac9f3192-e1d3-4827-9cf4-099253389e48 · outbound

This paper cites A descriptive basketball highlight dataset for automatic commentary gen- eration.

Player-Centric Multimodal Prompt Generation for Large Language Model Based Identity-Aware Basketball Video Captioning A descriptive basketball highlight dataset for automatic commentary gen- eration

Reference 57

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:47:54.589144Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T13:47:53.781975Z digest=sha256:aebad78cb2671be218b8c3f9130b312c350643c894e45fb3a8a181ac06206a6e

Observation 3d044e4e-301e-400c-90f2-2c0b96895b92 · outbound

This paper cites jump ball.

Player-Centric Multimodal Prompt Generation for Large Language Model Based Identity-Aware Basketball Video Captioning jump ball

Reference 58

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:47:54.449519Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T13:47:53.840267Z digest=sha256:95f473a53f9a65800aa5be7a2ce486f9fa07484734988a5530744266472f4e52

Pith citing papers

Observation 950a8896-23ab-4e88-9383-3528f1d53a28 · inbound

Explainable Action Form Assessment by Exploiting Multimodal Chain-of-Thoughts Reasoning cites this paper.

Explainable Action Form Assessment by Exploiting Multimodal Chain-of-Thoughts Reasoning Player-Centric Multimodal Prompt Generation for Large Language Model Based Identity-Aware Basketball Video Captioning

Reference 62

Resolution
unresolved
no resolver link, observed 2026-08-03T15:55:21.629729Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T15:55:21.629729Z digest=sha256:f601ca10e90d4a0317d2d223c451510005526550d0e3cb4bdc4e5a78c337f347