Pith. sign in

Paper Citation Record · LEDGER

V-STaR: Benchmarking Video-LLMs on Video Spatio-Temporal Reasoning

As of 9 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 34 inbound Pith citation observations for arXiv:2503.11495.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2503.11495 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 34 of 34 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00

measured 34 of 34 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-06T23:13:11.605015Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-04T03:29:31.466663Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 70e69085-bb97-406a-9398-6610082473d3 · inbound

MUSEG: Reinforcing Video Temporal Understanding via Timestamp-Aware Multi-Segment Grounding cites this paper.

MUSEG: Reinforcing Video Temporal Understanding via Timestamp-Aware Multi-Segment Grounding V-STaR: Benchmarking Video-LLMs on Video Spatio-Temporal Reasoning

Reference 3

Resolution
verified exact
arxiv_id, observed 2026-05-19T13:17:18.593246Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-19T13:13:40.485342Z digest=sha256:70bfced29de1a3a2b887c013e611c85aa4d7977fe8c74403c079966c762613f0

Observation e4c8650b-988e-4782-8feb-b1e7b9e6369e · inbound

Video-Holmes: Can MLLM Think Like Holmes for Complex Video Reasoning? cites this paper.

Video-Holmes: Can MLLM Think Like Holmes for Complex Video Reasoning? V-STaR: Benchmarking Video-LLMs on Video Spatio-Temporal Reasoning

Reference 17

Resolution
verified exact
arxiv_id, observed 2026-05-17T05:40:56.110232Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-17T05:40:55.944288Z digest=sha256:95ddd1b305c2503bb2521b944e9a24212cf127dbb66debdd52edda19e1d65081

Observation 0c4697dc-4d72-452b-95f8-863bae8de435 · inbound

Video-XL-2: Towards Very Long-Video Understanding Through Task-Aware KV Sparsification cites this paper.

Video-XL-2: Towards Very Long-Video Understanding Through Task-Aware KV Sparsification V-STaR: Benchmarking Video-LLMs on Video Spatio-Temporal Reasoning

Reference 55

Resolution
unresolved
no resolver link, observed 2026-08-06T23:13:11.605015Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:13:11.605015Z digest=sha256:4151b9baea8b7f7c64e8aa68b12e037661793b3230df9d13d81466f572e7f113

Observation 14647125-7f8d-47c6-b29e-99a62ffd6386 · inbound

Position: Reasoning After Perception Means Reasoning Without Vision cites this paper.

Position: Reasoning After Perception Means Reasoning Without Vision V-STaR: Benchmarking Video-LLMs on Video Spatio-Temporal Reasoning

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-06T15:26:01.737634Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T15:26:01.737634Z digest=sha256:6d295a3c80b8ea63c5dfdcc292acf6016b1efc71daa2fc92a2b1ee404a084cbe

Observation ff4940c9-173f-439d-af2e-64e36c8d9b52 · inbound

Video Reasoning without Training cites this paper.

Video Reasoning without Training V-STaR: Benchmarking Video-LLMs on Video Spatio-Temporal Reasoning

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-04T09:12:07.014228Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T09:12:07.014228Z digest=sha256:266cffda9c65aba1f379479383c2955f1b8cd6e33c1d32341eb799f0d822270f

Observation 0dab8555-7807-4152-81bb-771f59931837 · inbound

SPHINX: A Synthetic Environment for Visual Perception and Reasoning cites this paper.

SPHINX: A Synthetic Environment for Visual Perception and Reasoning V-STaR: Benchmarking Video-LLMs on Video Spatio-Temporal Reasoning

Reference 10

Resolution
verified exact
arxiv_id, observed 2026-05-17T04:21:30.760430Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-17T04:19:26.808804Z digest=sha256:c9daa4a67cb5778aed4f0f5796cf2fc20e73060e7f69e96627edad7c04d5a85c

Observation 49fdc14f-3a78-4c35-8d12-ed88dc970b80 · inbound

ToG-Bench: Task-Oriented Spatio-Temporal Grounding in Egocentric Videos cites this paper.

ToG-Bench: Task-Oriented Spatio-Temporal Grounding in Egocentric Videos V-STaR: Benchmarking Video-LLMs on Video Spatio-Temporal Reasoning

Reference 5

Resolution
verified exact
arxiv_id, observed 2026-05-17T02:51:28.025965Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-17T02:50:17.955287Z digest=sha256:d1afd5fba45b1054d33859a01448608494d0f2f7a776d3e7dd20fd03abf30422

Observation 05787ab9-47af-4202-a225-2e5012a2b0f3 · inbound

Detector-Empowered Video Large Language Model for Efficient Spatio-Temporal Grounding cites this paper.

Detector-Empowered Video Large Language Model for Efficient Spatio-Temporal Grounding V-STaR: Benchmarking Video-LLMs on Video Spatio-Temporal Reasoning

Reference 11

Resolution
verified exact
arxiv_id, observed 2026-05-17T00:58:46.542264Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-17T00:54:53.789523Z digest=sha256:144bfd9eb5bb65033ce83b4e0b1abea145c9dfe95b8128d95355f25473226e3a

Observation 34752852-56d0-4893-b0d6-2cc488e0785a · inbound

$M^3-Verse$: A "Spot the Difference" Challenge for Large Multimodal Models cites this paper.

$M^3-Verse$: A "Spot the Difference" Challenge for Large Multimodal Models V-STaR: Benchmarking Video-LLMs on Video Spatio-Temporal Reasoning

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-03T14:57:29.101267Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T14:57:29.101267Z digest=sha256:f83d168b869628f200d58af3156a219ca6e76a64496cfcab2befbcebff4f71c0

Observation 4434c415-058f-4bc4-b1d9-d3381d95eb76 · inbound

GraphThinker: Reinforcing Temporally Grounded Video Reasoning with Event Graph Thinking cites this paper.

GraphThinker: Reinforcing Temporally Grounded Video Reasoning with Event Graph Thinking V-STaR: Benchmarking Video-LLMs on Video Spatio-Temporal Reasoning

Reference 8

Resolution
verified exact
arxiv_id, observed 2026-05-15T20:50:17.357781Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-15T20:48:44.933542Z digest=sha256:5459d354778069f1fa8548c7a6b3d3f44ca76f6a73b79ae57dd9ab47f08d0d52

Observation 01b27391-7998-452c-82c8-15a0639236b9 · inbound

Reasoning over Video: Evaluating How MLLMs Extract, Integrate, and Reconstruct Spatiotemporal Evidence cites this paper.

Reasoning over Video: Evaluating How MLLMs Extract, Integrate, and Reconstruct Spatiotemporal Evidence V-STaR: Benchmarking Video-LLMs on Video Spatio-Temporal Reasoning

Reference 5

Resolution
verified exact
arxiv_id, observed 2026-05-15T11:29:58.756762Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-15T11:28:29.772341Z digest=sha256:89788242b0283e0a6f17440ea516350f33ae0e1160b9541baf1b25ae8a842291

Observation 39ea0772-cfc9-4060-adbe-f6b28a0d1bf0 · inbound

LanteRn: Latent Visual Structured Reasoning cites this paper.

LanteRn: Latent Visual Structured Reasoning V-STaR: Benchmarking Video-LLMs on Video Spatio-Temporal Reasoning

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-02T17:26:40.211189Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T17:26:40.211189Z digest=sha256:f9019aa5695a00288d203ccd376ec8cc4b1e95b18ffc4a4b5238d8fb8d333a93

Observation d296c926-fb7b-4b53-bc23-7760c7688fd7 · inbound

Thinking Diffusion: Penalize and Guide Visual-Grounded Reasoning in Diffusion Multimodal Language Models cites this paper.

Thinking Diffusion: Penalize and Guide Visual-Grounded Reasoning in Diffusion Multimodal Language Models V-STaR: Benchmarking Video-LLMs on Video Spatio-Temporal Reasoning

Reference 4

Resolution
verified exact
arxiv_id, observed 2026-05-10T23:30:50.483615Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-10T19:06:45.417233Z digest=sha256:fe3e0a5ddf5cd030020c2359b195b9216222c13dae1d48eeff2cff94b19b7038

Observation 5fe25213-2759-40e5-916b-3de23524f770 · inbound

PinpointQA: A Dataset and Benchmark for Small Object-Centric Spatial Understanding in Indoor Videos cites this paper.

PinpointQA: A Dataset and Benchmark for Small Object-Centric Spatial Understanding in Indoor Videos V-STaR: Benchmarking Video-LLMs on Video Spatio-Temporal Reasoning

Reference 5

Resolution
verified exact
arxiv_id, observed 2026-05-11T08:21:01.691072Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-10T16:40:39.180502Z digest=sha256:22b08ac3bcc4c304e37efcdf6be2f69303746b1c9e1c536c65502d776a10088d

Observation fdd9b21a-10ae-4891-874d-3e9774f2f93b · inbound

PinpointQA: A Dataset and Benchmark for Small Object-Centric Spatial Understanding in Indoor Videos cites this paper.

PinpointQA: A Dataset and Benchmark for Small Object-Centric Spatial Understanding in Indoor Videos V-STaR: Benchmarking Video-LLMs on Video Spatio-Temporal Reasoning

Reference 5

Resolution
verified exact
arxiv_id, observed 2026-05-15T06:39:48.776931Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-15T06:36:57.544235Z digest=sha256:ebfb640fa8e876de49606ba47a257a6a34336be51f9f7bed8d5fd2ee717d661d

Observation bfed0c39-e49c-4fa0-8d6a-bc365b8e3c52 · inbound

Grounding Video Reasoning in Physical Signals cites this paper.

Grounding Video Reasoning in Physical Signals V-STaR: Benchmarking Video-LLMs on Video Spatio-Temporal Reasoning

Reference 5

Resolution
verified exact
arxiv_id, observed 2026-05-09T22:34:07.276611Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-09T22:29:21.240010Z digest=sha256:888d915df7d4f9ef1c15ba75bb5daf10056d29109fa7c5a9553914b869ec95e3

Observation 4f64092f-2de4-4f72-848b-535b6c76bdb1 · inbound

Lost in the Vibrations: Vision Language Models Fail the Dynamic Gauges Test cites this paper.

Lost in the Vibrations: Vision Language Models Fail the Dynamic Gauges Test V-STaR: Benchmarking Video-LLMs on Video Spatio-Temporal Reasoning

Reference 5

Resolution
verified exact
arxiv_id, observed 2026-05-10T06:21:26.652635Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-10T06:21:17.509796Z digest=sha256:3a65f6ad73f2e9c5cc6941d2d6885b592174627f8a04aad5a2c95ad4e1bd35af

Observation 580cd081-71b7-489a-8ca3-6df290398c05 · inbound

State Beyond Appearance: Diagnosing and Improving State Consistency in Dial-Based Measurement Reading cites this paper.

State Beyond Appearance: Diagnosing and Improving State Consistency in Dial-Based Measurement Reading V-STaR: Benchmarking Video-LLMs on Video Spatio-Temporal Reasoning

Reference 5

Resolution
verified exact
arxiv_id, observed 2026-05-12T09:16:26.443266Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-07T11:45:57.291112Z digest=sha256:c7ab0e8c1547db1a570b4f3fd2e22fa66d331c0af83525340b6d8f235c2778d4

Observation 40c04b4a-def2-4a06-94e8-64e9fb4882ac · inbound

VISD: Enhancing Video Reasoning via Structured Self-Distillation cites this paper.

VISD: Enhancing Video Reasoning via Structured Self-Distillation V-STaR: Benchmarking Video-LLMs on Video Spatio-Temporal Reasoning

Reference 7

Resolution
verified exact
arxiv_id, observed 2026-05-11T18:46:08.696801Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-08T14:06:27.953376Z digest=sha256:a2b7e7ecece9614cc410343f98aee58156d221df322f7053a7813d365de730f5

Observation 3139deae-cdd6-464c-9dac-4bae4faeca9b · inbound

VISD: Enhancing Video Reasoning via Structured Self-Distillation cites this paper.

VISD: Enhancing Video Reasoning via Structured Self-Distillation V-STaR: Benchmarking Video-LLMs on Video Spatio-Temporal Reasoning

Reference 7

Resolution
verified exact
arxiv_id, observed 2026-05-11T01:50:51.219025Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-11T01:49:41.654207Z digest=sha256:2e3b1176a0a5e6ba521dcae8f4a99bf0ba640b900e9e7f457f426ed2dcd1cae8

Observation e03cac00-ffab-4db3-8d13-f374ff4d5d1c · inbound

VISD: Enhancing Video Reasoning via Structured Self-Distillation cites this paper.

VISD: Enhancing Video Reasoning via Structured Self-Distillation V-STaR: Benchmarking Video-LLMs on Video Spatio-Temporal Reasoning

Reference 7

Resolution
verified exact
arxiv_id, observed 2026-05-12T07:11:27.939351Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-12T03:35:59.553683Z digest=sha256:7051a197e24dc22b69259cde5b1f42c68d07c90068a299daf4dfcacc25679262

Observation 4d0633f3-9ddf-4e90-90e3-aadf43a69681 · inbound

VISD: Enhancing Video Reasoning via Structured Self-Distillation cites this paper.

VISD: Enhancing Video Reasoning via Structured Self-Distillation V-STaR: Benchmarking Video-LLMs on Video Spatio-Temporal Reasoning

Reference 7

Resolution
verified exact
arxiv_id, observed 2026-05-25T06:10:23.976068Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-25T06:08:19.956833Z digest=sha256:68ca03349406f3ae70b79add54bbc564beb0acba3d81c1f4250cbe2cd0a1c126

Observation 575950d7-67f5-4551-84ea-80792507ff47 · inbound

Tracking the Truth: Object-Centric Spatio-Temporal Monitoring for Video Large Language Models cites this paper.

Tracking the Truth: Object-Centric Spatio-Temporal Monitoring for Video Large Language Models V-STaR: Benchmarking Video-LLMs on Video Spatio-Temporal Reasoning

Reference 4

Resolution
verified exact
arxiv_id, observed 2026-05-12T07:36:32.249808Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-12T02:31:40.463891Z digest=sha256:abf843c146c76b82cb9bcba5a76dd0d93d5589eb6030dc5470de937f1504a246

Observation 99feebf8-132f-4a55-8ed9-5c6316bba8ce · inbound

STAR: Failure-Aware Markovian Routing for Multi-Agent Spatiotemporal Reasoning cites this paper.

STAR: Failure-Aware Markovian Routing for Multi-Agent Spatiotemporal Reasoning V-STaR: Benchmarking Video-LLMs on Video Spatio-Temporal Reasoning

Reference 3

Resolution
verified exact
arxiv_id, observed 2026-05-12T07:31:25.499559Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-12T02:42:42.419975Z digest=sha256:3303e23f72a4312eed790c7c00a360a6426013a570dd1fadc10a9048d0e80058

Observation 60f3c841-2f41-4531-bd2f-6af2879bfb36 · inbound

STAR: Failure-Aware Markovian Routing for Multi-Agent Spatiotemporal Reasoning cites this paper.

STAR: Failure-Aware Markovian Routing for Multi-Agent Spatiotemporal Reasoning V-STaR: Benchmarking Video-LLMs on Video Spatio-Temporal Reasoning

Reference 3

Resolution
verified exact
arxiv_id, observed 2026-05-13T03:52:12.584242Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-13T03:50:25.089976Z digest=sha256:f3bb10b8b6249fe44e7d4b9ba1ac820e09ba17c2c046c507ad1e48a86b89185c

Observation bdfd50cc-c3a1-45ab-a806-613502762a66 · inbound

STAR: Failure-Aware Markovian Routing for Multi-Agent Spatiotemporal Reasoning cites this paper.

STAR: Failure-Aware Markovian Routing for Multi-Agent Spatiotemporal Reasoning V-STaR: Benchmarking Video-LLMs on Video Spatio-Temporal Reasoning

Reference 3

Resolution
verified exact
arxiv_id, observed 2026-05-19T14:47:36.112253Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-19T14:46:40.165981Z digest=sha256:458fcb82818abe2e48603b6517e1323217ca826218224413d78b4e718e95e30b

Observation 43114469-3aed-4f9b-93cc-492c154b77da · inbound

Leveraging Latent Visual Reasoning in Silence cites this paper.

Leveraging Latent Visual Reasoning in Silence V-STaR: Benchmarking Video-LLMs on Video Spatio-Temporal Reasoning

Reference 4

Resolution
verified exact
arxiv_id, observed 2026-05-20T10:28:11.881438Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-20T10:27:13.859216Z digest=sha256:a6cb5b6f95dec364943388f189277efbc426f71f9c03f0332324d279fa8e5c03

Observation b161ec6d-aff5-4fb8-a9d6-aba3e7bfb6e1 · inbound

EgoCoT-Bench: Benchmarking Grounded and Verifiable Operation-Centric Chain of Thought Reasoning for MLLMs cites this paper.

EgoCoT-Bench: Benchmarking Grounded and Verifiable Operation-Centric Chain of Thought Reasoning for MLLMs V-STaR: Benchmarking Video-LLMs on Video Spatio-Temporal Reasoning

Reference 8

Resolution
metadata mismatch
arxiv_id, observed 2026-05-20T05:53:22.276707Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-20T05:53:05.450946Z digest=sha256:6cc9c86a3bc01d3aa24313b4acc50ec6b43bbf9877d3bda006fbd62c5d333463

Observation acc17bc1-d23a-4fb0-bffd-c0c25bd9709d · inbound

Maestro: Reinforcement Learning to Orchestrate Hierarchical Model-Skill Ensembles cites this paper.

Maestro: Reinforcement Learning to Orchestrate Hierarchical Model-Skill Ensembles V-STaR: Benchmarking Video-LLMs on Video Spatio-Temporal Reasoning

Reference 9

Resolution
verified exact
arxiv_id, observed 2026-05-22T07:51:15.464644Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-22T07:51:13.362986Z digest=sha256:9657a444c81ead9446135a89ee71f5f6c2ea34797371c6c22e473076b4a3c758

Observation 4118160a-482e-4743-9a85-123aa026a9e0 · inbound

X-Stream: Exploring MLLMs as Multiplexers for Multi-Stream Understanding cites this paper.

X-Stream: Exploring MLLMs as Multiplexers for Multi-Stream Understanding V-STaR: Benchmarking Video-LLMs on Video Spatio-Temporal Reasoning

Reference 13

Resolution
metadata mismatch
arxiv_id, observed 2026-07-01T22:46:18.941027Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-06-28T15:06:22.102725Z digest=sha256:0583a78d2a1c5c78096b00dd71a4caef3ab852cb525fcc45c27509db931a89e1

Observation f42b710a-b94a-43a5-a40c-43ae22d8d8c4 · inbound

X-Stream: Exploring MLLMs as Multiplexers for Multi-Stream Understanding cites this paper.

X-Stream: Exploring MLLMs as Multiplexers for Multi-Stream Understanding V-STaR: Benchmarking Video-LLMs on Video Spatio-Temporal Reasoning

Reference 13

Resolution
metadata mismatch
arxiv_id, observed 2026-06-30T10:44:36.412491Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-06-30T10:42:37.401221Z digest=sha256:bc0bf9c0325d4f0e321ef4bbcbbc73639370e0cfffe6fce68b74431cfe496e65

Observation 3d6a04cc-ad82-4abd-948a-7b359b204343 · inbound

CARE: Competence-Aware Reward Shaping for Adaptive Reasoning Length in Video-MLLMs cites this paper.

CARE: Competence-Aware Reward Shaping for Adaptive Reasoning Length in Video-MLLMs V-STaR: Benchmarking Video-LLMs on Video Spatio-Temporal Reasoning

Reference 16

Resolution
verified exact
arxiv_id, observed 2026-07-04T03:29:31.469215Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-06-26T17:56:34.203135Z digest=sha256:ba5f5a3a7dc0946924894776ccba5fb49e14f7b755f71a2bf55dd2ba78b4df7d

Observation 8187d767-d0f0-47db-a4b8-d683cccb1b57 · inbound

Video-MME-Logical: A Controlled Diagnostic Benchmark for Video Temporal-Logical Reasoning cites this paper.

Video-MME-Logical: A Controlled Diagnostic Benchmark for Video Temporal-Logical Reasoning V-STaR: Benchmarking Video-LLMs on Video Spatio-Temporal Reasoning

Reference 28

Resolution
metadata mismatch
arxiv_id, observed 2026-06-29T19:13:52.945728Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-06-29T04:53:25.259840Z digest=sha256:f3e5a909c35b20b2c7a7d3e03453b879a772da6b2c775b49184c39f9c1de4edf

Observation 30754d7d-7d56-47de-8646-0b001dca03f1 · inbound

AnyGroundBench: A Specialized-Domain Benchmark for Video Grounding in Vision-Language Models cites this paper.

AnyGroundBench: A Specialized-Domain Benchmark for Video Grounding in Vision-Language Models V-STaR: Benchmarking Video-LLMs on Video Spatio-Temporal Reasoning

Reference 7

Resolution
verified exact
arxiv_id, observed 2026-07-03T15:48:34.882701Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-07-03T15:46:26.401577Z digest=sha256:eb3297bb1d53a64ca407a1a52287c8ae8f2a47d6087142ebbae1cffc309ebb23