Pith. sign in

Paper Citation Record · LEDGER

SCBench: A Sports Commentary Benchmark for Video LLMs

As of 15 August 2026, this Paper Citation Record lists 47 of 47 outbound references and 2 inbound Pith citation observations for arXiv:2412.17637.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2412.17637 v1

Coverage vector

measured 47 of 47 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-11T05:24:17.842065Z

measured 49 of 49 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-14T06:32:32.682623+00:00

measured 2 of 2 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-05-10T19:10:07.699225Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-10T23:25:51.674329Z

Reference resolution

47 of 47 outbound references displayed

  • verified exact0
  • verified fuzzy29
  • unresolved18
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 8b6535f5-61bb-4049-acd1-2fe75a8907d3 · outbound

This paper cites write newline.

SCBench: A Sports Commentary Benchmark for Video LLMs write newline

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-11T05:24:17.683409Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T05:24:17.683409Z digest=sha256:1c184d4a4e56b51be31499e55fee06d231bd6d44bb53a740bdff5a5a32903fce

Observation e55d5ec1-caed-4f9d-bfec-fc3518af17e9 · outbound

This paper cites Spice: Semantic propositional image caption evaluation, 2016.

SCBench: A Sports Commentary Benchmark for Video LLMs Spice: Semantic propositional image caption evaluation, 2016

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T05:24:18.277345Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=arxiv_source observed=2026-08-11T05:24:17.687975Z digest=sha256:66e8ef208860e52d7ab2a895bfc26bfe0f80becdb332dd125bf301d9a0e9e873

Observation 92643bb3-53a7-4370-82e1-8eb769cc1569 · outbound

This paper cites METEOR : An automatic metric for MT evaluation with improved correlation with human judgments.

SCBench: A Sports Commentary Benchmark for Video LLMs METEOR : An automatic metric for MT evaluation with improved correlation with human judgments

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T05:24:18.267206Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=arxiv_source observed=2026-08-11T05:24:17.691725Z digest=sha256:a32e725cbc0e21f1caf9aa36a547b9d49f96a46575660d5daebc2d9ea7e82517

Observation 81aa4ebc-e894-4060-853e-1f87748c46ba · outbound

This paper cites P2anet: A dataset and benchmark for dense action detection from table tennis match broadcasting videos, 2024.

SCBench: A Sports Commentary Benchmark for Video LLMs P2anet: A dataset and benchmark for dense action detection from table tennis match broadcasting videos, 2024

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T05:24:18.257020Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=arxiv_source observed=2026-08-11T05:24:17.695503Z digest=sha256:4fdf8bd150b6e4e1b606114b3a91ebfe20aabf9982e63123ec6c9067cab3a90c

Observation d38a36f9-cdc7-4668-81bc-d2bff5b5d90e · outbound

This paper cites an unresolved cited work.

SCBench: A Sports Commentary Benchmark for Video LLMs Unresolved cited work

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-11T05:24:17.699465Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T05:24:17.699465Z digest=sha256:f93d9a3960586158fb54eed1ed7e6b2223ad433ac4317da737cca3efd3ad21a0

Observation 01fb1590-e1ab-4220-ae56-24e768fe32bf · outbound

This paper cites Autoeval-video: An automatic benchmark for assessing large vision language models in open-ended video question answering, 2024 a.

SCBench: A Sports Commentary Benchmark for Video LLMs Autoeval-video: An automatic benchmark for assessing large vision language models in open-ended video question answering, 2024 a

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T05:24:18.240253Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=arxiv_source observed=2026-08-11T05:24:17.703060Z digest=sha256:c7fc890691957530fe0c11cb062d55c036e493c9fdd87035f208b22e4918ceda

Observation 4885a10a-2b21-4003-bcdf-004717ab064e · outbound

This paper cites Internvl: Scaling up vision foundation models and aligning for generic visual-linguistic tasks, 2024 b.

SCBench: A Sports Commentary Benchmark for Video LLMs Internvl: Scaling up vision foundation models and aligning for generic visual-linguistic tasks, 2024 b

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T05:24:18.229784Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=arxiv_source observed=2026-08-11T05:24:17.706887Z digest=sha256:30bd2582f61754be467f6fb49da7088c872a70af9ab2b954b1039fc0f71a2d86

Observation 49590249-4420-459c-86fb-112a787c439a · outbound

This paper cites Sports re-id: Improving re-identification of players in broadcast videos of team sports, 2022.

SCBench: A Sports Commentary Benchmark for Video LLMs Sports re-id: Improving re-identification of players in broadcast videos of team sports, 2022

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T05:24:18.218808Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=arxiv_source observed=2026-08-11T05:24:17.710525Z digest=sha256:ddc9a59558aba1ef9209e01f8fb10a474c5019110339b0fc0ed9de506bc39207

Observation a09997db-526f-4884-9426-88cfca800994 · outbound

This paper cites Seikavandi, Jacob V.

SCBench: A Sports Commentary Benchmark for Video LLMs Seikavandi, Jacob V

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T05:24:18.208594Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=arxiv_source observed=2026-08-11T05:24:17.713845Z digest=sha256:6dbb97119d2d5d5e7fe8175737c2dde6c58a0e10db5d33dc825f1d266df1a977

Observation bd052060-cac3-4270-9fd0-2fc4fa6f6edc · outbound

This paper cites Mmbench-video: A long-form multi-shot benchmark for holistic video understanding, 2024.

SCBench: A Sports Commentary Benchmark for Video LLMs Mmbench-video: A long-form multi-shot benchmark for holistic video understanding, 2024

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T05:24:18.199517Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=arxiv_source observed=2026-08-11T05:24:17.717112Z digest=sha256:379491087bc65907b6fe657595e94207c7fc347adef54a68973101537d3c3b6f

Observation b6b813a9-e419-4a29-be55-548f911c8fe4 · outbound

This paper cites an unresolved cited work.

SCBench: A Sports Commentary Benchmark for Video LLMs Unresolved cited work

Reference 11

Resolution
unresolved
raw_fallback, observed 2026-08-11T05:24:18.190434Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=arxiv_source observed=2026-08-11T05:24:17.720502Z digest=sha256:56bdcc7f9a81df7f634d436d5de3620584039c4548cb2667549095f7e754b804

Observation 201e4a8c-2feb-4498-8750-0a5980363c45 · outbound

This paper cites Chatpose: Chatting about 3d human pose.

SCBench: A Sports Commentary Benchmark for Video LLMs Chatpose: Chatting about 3d human pose

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T05:24:18.180323Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=arxiv_source observed=2026-08-11T05:24:17.723725Z digest=sha256:4c55b5ae1e9d6d281f32bf9ede65f83884b7ca8f68af1e04b2043ab047300b91

Observation ee94795f-eb39-4741-ba3c-83ad9bbe4e93 · outbound

This paper cites Video-mme: The first-ever comprehensive evaluation benchmark of multi-modal llms in video analysis, 2024.

SCBench: A Sports Commentary Benchmark for Video LLMs Video-mme: The first-ever comprehensive evaluation benchmark of multi-modal llms in video analysis, 2024

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-11T05:24:17.731047Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T05:24:17.731047Z digest=sha256:4592d22f1e545f6fe1feb1daa85616d4dd33a49baff8ab3e0a20232ea351a5a2

Observation f753cd98-bff0-460d-ac1f-22c9961d08f5 · outbound

This paper cites Mini-internvl: A flexible-transfer pocket multimodal model with 5.

SCBench: A Sports Commentary Benchmark for Video LLMs Mini-internvl: A flexible-transfer pocket multimodal model with 5

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T05:24:18.163828Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=arxiv_source observed=2026-08-11T05:24:17.734832Z digest=sha256:a5868be786ce1c0a62c52e55ac28f8bf9a21dd820f5d83f3193070bec1593b2e

Observation 5dab9f72-2add-4922-81fc-e92616989e7d · outbound

This paper cites Tgif-qa: Toward spatio-temporal reasoning in visual question answering, 2017.

SCBench: A Sports Commentary Benchmark for Video LLMs Tgif-qa: Toward spatio-temporal reasoning in visual question answering, 2017

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T05:24:18.153707Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=arxiv_source observed=2026-08-11T05:24:17.738244Z digest=sha256:a20558bc5a1056217b723a5a3ad490f69d3261a3c62e404cab14eabfced09e19

Observation f367ceff-082e-4348-8159-16ea306ededa · outbound

This paper cites Chat-univi: Unified visual representation empowers large language models with image and video understanding, 2024.

SCBench: A Sports Commentary Benchmark for Video LLMs Chat-univi: Unified visual representation empowers large language models with image and video understanding, 2024

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T05:24:18.143230Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=arxiv_source observed=2026-08-11T05:24:17.741346Z digest=sha256:f33afb27241a2627a14c6338af02aaf213ea0276fe940f6c5319e8e327cf4eee

Observation 7fa9b383-63be-4c0d-96d7-e83b3e6a0fb5 · outbound

This paper cites Retrieval-augmented generation for knowledge-intensive nlp tasks, 2021.

SCBench: A Sports Commentary Benchmark for Video LLMs Retrieval-augmented generation for knowledge-intensive nlp tasks, 2021

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-11T05:24:17.744141Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T05:24:17.744141Z digest=sha256:c1c2b48a364c09e9e2a133be1052721857ef9619d5b52039e062f33f771444e2

Observation d0c9d076-ab13-4cf7-ba77-886c8f7d102d · outbound

This paper cites Mvbench: A comprehensive multi-modal video understanding benchmark, 2024.

SCBench: A Sports Commentary Benchmark for Video LLMs Mvbench: A comprehensive multi-modal video understanding benchmark, 2024

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T05:24:18.127754Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=arxiv_source observed=2026-08-11T05:24:17.746975Z digest=sha256:89784455c74b52690b44382621a1a41276888721e0b7ae739779d3b917afa618

Observation 63c2705e-2a9b-4b8f-94ca-53e371c0145a · outbound

This paper cites Video-llava: Learning united visual representation by alignment before projection, 2023.

SCBench: A Sports Commentary Benchmark for Video LLMs Video-llava: Learning united visual representation by alignment before projection, 2023

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-11T05:24:17.749606Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T05:24:17.749606Z digest=sha256:37415c883203217b064ee4ef8c7f67959ab303c5e405e126609fad32b880882b

Observation a8cb6f5d-a8c1-4014-abe8-39cf5a8b1c5f · outbound

This paper cites ROUGE : A package for automatic evaluation of summaries.

SCBench: A Sports Commentary Benchmark for Video LLMs ROUGE : A package for automatic evaluation of summaries

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-11T05:24:17.752192Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T05:24:17.752192Z digest=sha256:d0fd0de6f7a2a67ca4d0408cd41bc7c0445483d16dd01478b85bdc4198eeb4b2

Observation 2dde1edb-3b92-4512-b7ae-b2c24263de32 · outbound

This paper cites Vila: On pre-training for visual language models, 2024.

SCBench: A Sports Commentary Benchmark for Video LLMs Vila: On pre-training for visual language models, 2024

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T05:24:18.105842Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=arxiv_source observed=2026-08-11T05:24:17.755116Z digest=sha256:8faffd19e6cfaa3103ae2d425d1fce18622a973ec46149e6e4df4df6780332fd

Observation 272a74c3-c035-45fa-b6d7-4b5bcc9a23e1 · outbound

This paper cites Llava-next: Improved reasoning, ocr, and world knowledge, 2024 a.

SCBench: A Sports Commentary Benchmark for Video LLMs Llava-next: Improved reasoning, ocr, and world knowledge, 2024 a

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T05:24:18.094806Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=arxiv_source observed=2026-08-11T05:24:17.757693Z digest=sha256:6ead08cc4a70ccda3f83379f29564c8570d424d335743e54b4b4920c009b0b25

Observation 5f312c6c-c424-4e8b-9f60-e60a9de35532 · outbound

This paper cites Kangaroo: A powerful video-language model supporting long-context video input, 2024 b.

SCBench: A Sports Commentary Benchmark for Video LLMs Kangaroo: A powerful video-language model supporting long-context video input, 2024 b

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T05:24:18.083902Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=arxiv_source observed=2026-08-11T05:24:17.760420Z digest=sha256:124ae63bf29d29a8788371a9b10751d535e5f1a95a5d8c4e3b34bd9e4393adea

Observation d441747c-01e2-428a-a79d-1b96fcec39e9 · outbound

This paper cites Fineaction: A fine-grained video dataset for temporal action localization.

SCBench: A Sports Commentary Benchmark for Video LLMs Fineaction: A fine-grained video dataset for temporal action localization

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T05:24:18.073076Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=arxiv_source observed=2026-08-11T05:24:17.763437Z digest=sha256:19375c728481c36cd859c3da988f0a13cb8693fcef1a674eb8ba9c676609f3d7

Observation a5b2fb07-83a1-4be2-8113-58add5c6f016 · outbound

This paper cites Tempcompass: Do video llms really understand videos?, 2024 c.

SCBench: A Sports Commentary Benchmark for Video LLMs Tempcompass: Do video llms really understand videos?, 2024 c

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T05:24:18.063005Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=arxiv_source observed=2026-08-11T05:24:17.767152Z digest=sha256:923fd1d6524986ea140383536474e0bf3d4036c51552092463a1ac58a8ff4e0d

Observation 495a771b-3799-45de-8bc8-c6f62180a8e1 · outbound

This paper cites Cross-block fine-grained semantic cascade for skeleton-based sports action recognition, 2024 d.

SCBench: A Sports Commentary Benchmark for Video LLMs Cross-block fine-grained semantic cascade for skeleton-based sports action recognition, 2024 d

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T05:24:18.053087Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=arxiv_source observed=2026-08-11T05:24:17.770554Z digest=sha256:3bcc97e919a5db772ed4e3c2919d7c42b6ad385d937deb0f7881c80007be9d72

Observation 1bf485d7-49e8-4bf9-87c5-ea611345cbe1 · outbound

This paper cites Egoschema: A diagnostic benchmark for very long-form video language understanding, 2023.

SCBench: A Sports Commentary Benchmark for Video LLMs Egoschema: A diagnostic benchmark for very long-form video language understanding, 2023

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-11T05:24:17.773955Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T05:24:17.773955Z digest=sha256:cb85dc380fdaa6034db339302b638fbd6c63ac48ebef0df6a0a87711bf4507c9

Observation 914eb97b-14ed-45ae-b0f7-c9c1f81aa64d · outbound

This paper cites Howto100m: Learning a text-video embedding by watching hundred million narrated video clips.

SCBench: A Sports Commentary Benchmark for Video LLMs Howto100m: Learning a text-video embedding by watching hundred million narrated video clips

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-11T05:24:17.777547Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T05:24:17.777547Z digest=sha256:880e93f3a4f306f4ec60420eefe1ff145991da77ef941dcb552d5f457822a43b

Observation 790bf743-95bc-4de3-be39-e8ff36a81b80 · outbound

This paper cites Video-Bench: A Comprehensive Benchmark and Toolkit for Evaluating Video-based Large Language Models.

SCBench: A Sports Commentary Benchmark for Video LLMs Video-Bench: A Comprehensive Benchmark and Toolkit for Evaluating Video-based Large Language Models

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-11T05:24:17.780895Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T05:24:17.780895Z digest=sha256:f106fc9a73fed89f2f378b03cd01c73399e3eda61db79ebbb051a15127fd1a65

Observation 00e92d4e-b813-4d32-9c87-1de79dd35e53 · outbound

This paper cites Video-bench: A comprehensive benchmark and toolkit for evaluating video-based large language models, 2023 b.

SCBench: A Sports Commentary Benchmark for Video LLMs Video-bench: A comprehensive benchmark and toolkit for evaluating video-based large language models, 2023 b

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T05:24:18.031890Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=arxiv_source observed=2026-08-11T05:24:17.784738Z digest=sha256:36eae62f1bdbb1b21e3eeda34efef743100e25581b09ff7f704d7ddef4b22133

Observation 318f1244-9d23-4391-96f8-c4276da2d98e · outbound

This paper cites B leu: a method for automatic evaluation of machine translation.

SCBench: A Sports Commentary Benchmark for Video LLMs B leu: a method for automatic evaluation of machine translation

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T05:24:18.020253Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=arxiv_source observed=2026-08-11T05:24:17.787974Z digest=sha256:52b1db3bb6ddd2e0824119cdfa03c91cbfdd99aaec90014d5638aaca488f5739

Observation 1e62d750-37c1-48f2-a52e-ff97f2b59ddd · outbound

This paper cites Perception test: A diagnostic benchmark for multimodal video models, 2023.

SCBench: A Sports Commentary Benchmark for Video LLMs Perception test: A diagnostic benchmark for multimodal video models, 2023

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T05:24:18.009190Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=arxiv_source observed=2026-08-11T05:24:17.791466Z digest=sha256:eb850263f8c7c07477236e4104f259eabecb20b33e9324c68cf048993c59c7d3

Observation 7ec357c2-f255-4e71-b76b-c16c38beb508 · outbound

This paper cites A survey of video datasets for grounded event understanding.

SCBench: A Sports Commentary Benchmark for Video LLMs A survey of video datasets for grounded event understanding

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T05:24:17.996400Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=arxiv_source observed=2026-08-11T05:24:17.794736Z digest=sha256:f00f0f8d681676dc5f387dc846455ad11848688097014913d58cc30a55d0c730

Observation 194d3318-23d9-447d-b2c2-4ee768d1fea6 · outbound

This paper cites Finegym: A hierarchical video dataset for fine-grained action understanding, 2020.

SCBench: A Sports Commentary Benchmark for Video LLMs Finegym: A hierarchical video dataset for fine-grained action understanding, 2020

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-11T05:24:17.797956Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T05:24:17.797956Z digest=sha256:d1e96a6d66e336ed7d632034eaf5601ac9fa3c0cf48f1994a94e76732994962a

Observation 1ad71774-beb3-43d1-9c46-8f1b44452bf0 · outbound

This paper cites Visual cot: Advancing multi-modal language models with a comprehensive dataset and benchmark for chain-of-thought reasoning, 2024.

SCBench: A Sports Commentary Benchmark for Video LLMs Visual cot: Advancing multi-modal language models with a comprehensive dataset and benchmark for chain-of-thought reasoning, 2024

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-11T05:24:17.801306Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T05:24:17.801306Z digest=sha256:b16dbcd23eeba57ae22755b07414da1a573c58e78ae47128d30aa851be498520

Observation 22ceb1ea-6e79-4b8f-a840-4c7cec682770 · outbound

This paper cites Playertv: Advanced player tracking and identification for automatic soccer highlight clips, 2024.

SCBench: A Sports Commentary Benchmark for Video LLMs Playertv: Advanced player tracking and identification for automatic soccer highlight clips, 2024

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T05:24:17.973119Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=arxiv_source observed=2026-08-11T05:24:17.804672Z digest=sha256:5a547b364ec48992ab0685b66ac74a9adcf812055ea89f27eb076b8807e0fbea

Observation 0e4d8d06-c469-4c5d-b539-8f68e0a4bde8 · outbound

This paper cites Internvl2: Better than the best—expanding performance boundaries of open-source multimodal models with the progressive scaling strategy.

SCBench: A Sports Commentary Benchmark for Video LLMs Internvl2: Better than the best—expanding performance boundaries of open-source multimodal models with the progressive scaling strategy

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T05:24:17.962914Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=arxiv_source observed=2026-08-11T05:24:17.808227Z digest=sha256:40699a6c8f484b71658ec33ae5367ed21defd6f04986b7751a193736c65e27b5

Observation 712f6b79-4947-472c-a454-94d31e375393 · outbound

This paper cites CIDEr: Consensus-based Image Description Evaluation.

SCBench: A Sports Commentary Benchmark for Video LLMs CIDEr: Consensus-based Image Description Evaluation

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-11T05:24:17.811515Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T05:24:17.811515Z digest=sha256:92abe0b2b1dec86fc27ad7c3262d94fb7f6725d3a05676355d72d34cb0c2b2df

Observation 48281dc9-9464-4065-9549-cac8cc4408b2 · outbound

This paper cites Lvbench: An extreme long video understanding benchmark, 2024.

SCBench: A Sports Commentary Benchmark for Video LLMs Lvbench: An extreme long video understanding benchmark, 2024

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-11T05:24:17.815018Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T05:24:17.815018Z digest=sha256:f5da96822399791f847194082f0752ec66a15b8cf92558d7c98281b4afc06cbc

Observation 9138a27a-d2d9-4035-8d99-52e1913656b4 · outbound

This paper cites Chain-of-thought prompting elicits reasoning in large language models, 2023.

SCBench: A Sports Commentary Benchmark for Video LLMs Chain-of-thought prompting elicits reasoning in large language models, 2023

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-11T05:24:17.818259Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T05:24:17.818259Z digest=sha256:2993e972ca5d17f194741bebd7876ffc079b79f11e7c8c59c51a1461351df0f7

Observation 2b9fd42c-5dab-4c68-852f-ececbfabeb6a · outbound

This paper cites Sportshhi: A dataset for human-human interaction detection in sports videos, 2024.

SCBench: A Sports Commentary Benchmark for Video LLMs Sportshhi: A dataset for human-human interaction detection in sports videos, 2024

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T05:24:17.941986Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=arxiv_source observed=2026-08-11T05:24:17.821236Z digest=sha256:3b60421ea450cbe3683739291ceff4bf38be42c3e62b4dff43ac6fddf1af0dbe

Observation 191afc41-22d3-4cb1-8203-7215a9c1bc65 · outbound

This paper cites Next-qa:next phase of question-answering to explaining temporal actions.

SCBench: A Sports Commentary Benchmark for Video LLMs Next-qa:next phase of question-answering to explaining temporal actions

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T05:24:17.932170Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=arxiv_source observed=2026-08-11T05:24:17.824496Z digest=sha256:77f235a80fb3d35351bd661e1eb3f5be5bc4d0c0fa2646dd8bc7ea2541162ebb

Observation 5a12612e-03b0-4cd0-92ba-11cf5040ecb1 · outbound

This paper cites Video question answering via gradually refined attention over appearance and motion.

SCBench: A Sports Commentary Benchmark for Video LLMs Video question answering via gradually refined attention over appearance and motion

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T05:24:17.923406Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=arxiv_source observed=2026-08-11T05:24:17.827835Z digest=sha256:111f359a4817e6f657c73ed11936d338ce6e673c5dfc388aef36fcf81f8d7cb0

Observation 6ece9f76-4af6-42f9-8fe9-ec73996239ee · outbound

This paper cites Youku-mplug: A 10 million large-scale chinese video-language dataset for pre-training and benchmarks, 2023.

SCBench: A Sports Commentary Benchmark for Video LLMs Youku-mplug: A 10 million large-scale chinese video-language dataset for pre-training and benchmarks, 2023

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T05:24:17.913531Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=arxiv_source observed=2026-08-11T05:24:17.830956Z digest=sha256:115bda6d22b2174b2d81e6c32a58b47619c0ac2fd632d24fac138e2413bddb58

Observation 2f72312a-1c0c-4b73-93f4-b98ac8a319bb · outbound

This paper cites Activitynet-qa: A dataset for understanding complex web videos via question answering, 2019.

SCBench: A Sports Commentary Benchmark for Video LLMs Activitynet-qa: A dataset for understanding complex web videos via question answering, 2019

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-11T05:24:17.834360Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T05:24:17.834360Z digest=sha256:73c910a9185c6e5b0783e1e5f304e5b8fa80fc50620fba1cd0653454c587f9b3

Observation ec67df3c-0e60-44f4-b323-567b1480ea59 · outbound

This paper cites Long context transfer from language to vision, 2024.

SCBench: A Sports Commentary Benchmark for Video LLMs Long context transfer from language to vision, 2024

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-11T05:24:17.837559Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T05:24:17.837559Z digest=sha256:d47db6ad4fb870edf761b7d5f73f8bf917098fe14a4fa56743d0dbb6f7ac9c9a

Observation f966328a-0266-435b-8b64-9392d51e1834 · outbound

This paper cites A Survey on Generative AI and LLM for Video Generation, Understanding, and Streaming.

SCBench: A Sports Commentary Benchmark for Video LLMs A Survey on Generative AI and LLM for Video Generation, Understanding, and Streaming

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-11T05:24:17.842065Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T05:24:17.842065Z digest=sha256:194fb2bbce0d8cc5757d2fb575a4ee73e7c2a3b57e6f023ef0014cfc96137555

Pith citing papers

Observation 9412f7ea-5c9d-4aa8-b13c-2020277ff3bd · inbound

BoxComm: Benchmarking Category-Aware Commentary Generation and Narration Rhythm in Boxing cites this paper.

BoxComm: Benchmarking Category-Aware Commentary Generation and Narration Rhythm in Boxing SCBench: A Sports Commentary Benchmark for Video LLMs

Reference 11

Resolution
verified exact
arxiv_id, observed 2026-05-10T23:25:51.682028Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-05-10T19:10:07.699225Z digest=sha256:c1265de59c58e45df2afdb13615627584ba6cb13777b5b576f0390d7e7cdaf9f

Observation 5ba8e445-dd51-4bbc-964e-2b2d7c78c243 · inbound

RefereeBench: Are Video MLLMs Ready to be Multi-Sport Referees cites this paper.

RefereeBench: Are Video MLLMs Ready to be Multi-Sport Referees SCBench: A Sports Commentary Benchmark for Video LLMs

Reference 17

Resolution
verified exact
arxiv_id, observed 2026-05-10T08:48:02.282471Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-05-10T08:38:27.081358Z digest=sha256:0454256d19ceb6406fe471532d981b5d1d73cb2b3b8fee6bc637417de8b456b4