Pith. sign in

Paper Citation Record · LEDGER

V-STaR: Benchmarking Video-LLMs on Video Spatio-Temporal Reasoning

As of 10 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 34 inbound Pith citation observations for arXiv:2503.11495.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2503.11495 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 34 of 34 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-10T06:31:04.303077+00:00

measured 34 of 34 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-06T23:13:11.605015Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-04T03:29:31.466663Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 70e69085-bb97-406a-9398-6610082473d3 · inbound

MUSEG: Reinforcing Video Temporal Understanding via Timestamp-Aware Multi-Segment Grounding cites this paper.

MUSEG: Reinforcing Video Temporal Understanding via Timestamp-Aware Multi-Segment Grounding V-STaR: Benchmarking Video-LLMs on Video Spatio-Temporal Reasoning

Reference 3

Resolution
verified exact
arxiv_id, observed 2026-05-19T13:17:18.593246Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-19T13:13:40.485342Z digest=sha256:45656485de37a33ae6bcc8a7b7262e005a0249cd387e129ce65eec92d8112a5a

Observation e4c8650b-988e-4782-8feb-b1e7b9e6369e · inbound

Video-Holmes: Can MLLM Think Like Holmes for Complex Video Reasoning? cites this paper.

Video-Holmes: Can MLLM Think Like Holmes for Complex Video Reasoning? V-STaR: Benchmarking Video-LLMs on Video Spatio-Temporal Reasoning

Reference 17

Resolution
verified exact
arxiv_id, observed 2026-05-17T05:40:56.110232Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-17T05:40:55.944288Z digest=sha256:2ff8824552dd3045fe8aa9017fd26895b529b9aa9783dc5c2d2470e52234ee42

Observation 0c4697dc-4d72-452b-95f8-863bae8de435 · inbound

Video-XL-2: Towards Very Long-Video Understanding Through Task-Aware KV Sparsification cites this paper.

Video-XL-2: Towards Very Long-Video Understanding Through Task-Aware KV Sparsification V-STaR: Benchmarking Video-LLMs on Video Spatio-Temporal Reasoning

Reference 55

Resolution
unresolved
no resolver link, observed 2026-08-06T23:13:11.605015Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:13:11.605015Z digest=sha256:a9f64f9e25edce92c7738ee7266803e8233ff2770d1e6e2aa99e1e7b04f449ee

Observation 14647125-7f8d-47c6-b29e-99a62ffd6386 · inbound

Position: Reasoning After Perception Means Reasoning Without Vision cites this paper.

Position: Reasoning After Perception Means Reasoning Without Vision V-STaR: Benchmarking Video-LLMs on Video Spatio-Temporal Reasoning

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-06T15:26:01.737634Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T15:26:01.737634Z digest=sha256:0010d75235400d6453118aadf65358fe3741f5f8cb620f6e1e7091d54f89c42a

Observation ff4940c9-173f-439d-af2e-64e36c8d9b52 · inbound

Video Reasoning without Training cites this paper.

Video Reasoning without Training V-STaR: Benchmarking Video-LLMs on Video Spatio-Temporal Reasoning

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-04T09:12:07.014228Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T09:12:07.014228Z digest=sha256:4fc6fd206aa74a8610d641f7b1fc63f76b3eec638e937f55f0dc375f64fa9b21

Observation 0dab8555-7807-4152-81bb-771f59931837 · inbound

SPHINX: A Synthetic Environment for Visual Perception and Reasoning cites this paper.

SPHINX: A Synthetic Environment for Visual Perception and Reasoning V-STaR: Benchmarking Video-LLMs on Video Spatio-Temporal Reasoning

Reference 10

Resolution
verified exact
arxiv_id, observed 2026-05-17T04:21:30.760430Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-17T04:19:26.808804Z digest=sha256:f7e339ff8e9e8c766a04bf70a9e1d5151bfaef5cc3711c383cb7d81df53a565c

Observation 49fdc14f-3a78-4c35-8d12-ed88dc970b80 · inbound

ToG-Bench: Task-Oriented Spatio-Temporal Grounding in Egocentric Videos cites this paper.

ToG-Bench: Task-Oriented Spatio-Temporal Grounding in Egocentric Videos V-STaR: Benchmarking Video-LLMs on Video Spatio-Temporal Reasoning

Reference 5

Resolution
verified exact
arxiv_id, observed 2026-05-17T02:51:28.025965Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-17T02:50:17.955287Z digest=sha256:246d64afded746a537f5ba09f471c07bb44cf87139f80d88342ca4c0ee4eb28e

Observation 05787ab9-47af-4202-a225-2e5012a2b0f3 · inbound

Detector-Empowered Video Large Language Model for Efficient Spatio-Temporal Grounding cites this paper.

Detector-Empowered Video Large Language Model for Efficient Spatio-Temporal Grounding V-STaR: Benchmarking Video-LLMs on Video Spatio-Temporal Reasoning

Reference 11

Resolution
verified exact
arxiv_id, observed 2026-05-17T00:58:46.542264Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-17T00:54:53.789523Z digest=sha256:249e323867998f6ac48ee99becd3a49f5c5e17e0eec09af97d648c59a590a981

Observation 34752852-56d0-4893-b0d6-2cc488e0785a · inbound

$M^3-Verse$: A "Spot the Difference" Challenge for Large Multimodal Models cites this paper.

$M^3-Verse$: A "Spot the Difference" Challenge for Large Multimodal Models V-STaR: Benchmarking Video-LLMs on Video Spatio-Temporal Reasoning

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-03T14:57:29.101267Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T14:57:29.101267Z digest=sha256:47663fae825eb801decf55ff75c489be6cbfee0b4803190c8fd54e7a63a758a5

Observation 4434c415-058f-4bc4-b1d9-d3381d95eb76 · inbound

GraphThinker: Reinforcing Temporally Grounded Video Reasoning with Event Graph Thinking cites this paper.

GraphThinker: Reinforcing Temporally Grounded Video Reasoning with Event Graph Thinking V-STaR: Benchmarking Video-LLMs on Video Spatio-Temporal Reasoning

Reference 8

Resolution
verified exact
arxiv_id, observed 2026-05-15T20:50:17.357781Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-15T20:48:44.933542Z digest=sha256:808b8096a29b7fda8aa3ae150d83b1c76fa52ead93a8f4f2e4839096ebaa8ba7

Observation 01b27391-7998-452c-82c8-15a0639236b9 · inbound

Reasoning over Video: Evaluating How MLLMs Extract, Integrate, and Reconstruct Spatiotemporal Evidence cites this paper.

Reasoning over Video: Evaluating How MLLMs Extract, Integrate, and Reconstruct Spatiotemporal Evidence V-STaR: Benchmarking Video-LLMs on Video Spatio-Temporal Reasoning

Reference 5

Resolution
verified exact
arxiv_id, observed 2026-05-15T11:29:58.756762Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-15T11:28:29.772341Z digest=sha256:512ecbd5a53a99017eb3c7a150192fe110cec7ac500d92751659126a5f5541d9

Observation 39ea0772-cfc9-4060-adbe-f6b28a0d1bf0 · inbound

LanteRn: Latent Visual Structured Reasoning cites this paper.

LanteRn: Latent Visual Structured Reasoning V-STaR: Benchmarking Video-LLMs on Video Spatio-Temporal Reasoning

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-02T17:26:40.211189Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T17:26:40.211189Z digest=sha256:c8ea3f994792e4dadbdff63a8b76f7da52dda93c1429e9dc012a2aa4be5e4de9

Observation d296c926-fb7b-4b53-bc23-7760c7688fd7 · inbound

Thinking Diffusion: Penalize and Guide Visual-Grounded Reasoning in Diffusion Multimodal Language Models cites this paper.

Thinking Diffusion: Penalize and Guide Visual-Grounded Reasoning in Diffusion Multimodal Language Models V-STaR: Benchmarking Video-LLMs on Video Spatio-Temporal Reasoning

Reference 4

Resolution
verified exact
arxiv_id, observed 2026-05-10T23:30:50.483615Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-10T19:06:45.417233Z digest=sha256:905cb9bf6e6eb9cf1242c52d01a5517faa6e15c4589730e1d86def46202ce8f4

Observation 5fe25213-2759-40e5-916b-3de23524f770 · inbound

PinpointQA: A Dataset and Benchmark for Small Object-Centric Spatial Understanding in Indoor Videos cites this paper.

PinpointQA: A Dataset and Benchmark for Small Object-Centric Spatial Understanding in Indoor Videos V-STaR: Benchmarking Video-LLMs on Video Spatio-Temporal Reasoning

Reference 5

Resolution
verified exact
arxiv_id, observed 2026-05-11T08:21:01.691072Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-10T16:40:39.180502Z digest=sha256:37d8aa8bdaa50a31c0002f5410cd67331191a1e95aded3312941724aface5556

Observation fdd9b21a-10ae-4891-874d-3e9774f2f93b · inbound

PinpointQA: A Dataset and Benchmark for Small Object-Centric Spatial Understanding in Indoor Videos cites this paper.

PinpointQA: A Dataset and Benchmark for Small Object-Centric Spatial Understanding in Indoor Videos V-STaR: Benchmarking Video-LLMs on Video Spatio-Temporal Reasoning

Reference 5

Resolution
verified exact
arxiv_id, observed 2026-05-15T06:39:48.776931Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-15T06:36:57.544235Z digest=sha256:0ec3c9b5c534655ce8b5a944cd41d4f5d448e90f1b238f74be0f64eae8cf6a8c

Observation bfed0c39-e49c-4fa0-8d6a-bc365b8e3c52 · inbound

Grounding Video Reasoning in Physical Signals cites this paper.

Grounding Video Reasoning in Physical Signals V-STaR: Benchmarking Video-LLMs on Video Spatio-Temporal Reasoning

Reference 5

Resolution
verified exact
arxiv_id, observed 2026-05-09T22:34:07.276611Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-09T22:29:21.240010Z digest=sha256:de46e05ea90b513931fa282f532607c0288c6d4f404339ea7c3b08a665168abc

Observation 4f64092f-2de4-4f72-848b-535b6c76bdb1 · inbound

Lost in the Vibrations: Vision Language Models Fail the Dynamic Gauges Test cites this paper.

Lost in the Vibrations: Vision Language Models Fail the Dynamic Gauges Test V-STaR: Benchmarking Video-LLMs on Video Spatio-Temporal Reasoning

Reference 5

Resolution
verified exact
arxiv_id, observed 2026-05-10T06:21:26.652635Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-10T06:21:17.509796Z digest=sha256:e87642e92f7a5373c81f88bfdc2a932e8a9d802e5702b7c6de489b253b412664

Observation 580cd081-71b7-489a-8ca3-6df290398c05 · inbound

State Beyond Appearance: Diagnosing and Improving State Consistency in Dial-Based Measurement Reading cites this paper.

State Beyond Appearance: Diagnosing and Improving State Consistency in Dial-Based Measurement Reading V-STaR: Benchmarking Video-LLMs on Video Spatio-Temporal Reasoning

Reference 5

Resolution
verified exact
arxiv_id, observed 2026-05-12T09:16:26.443266Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-07T11:45:57.291112Z digest=sha256:999f7df82dee5d414cb26e9fffe71cc22b51418bbe19c5aaa683c5e84bbf6939

Observation 40c04b4a-def2-4a06-94e8-64e9fb4882ac · inbound

VISD: Enhancing Video Reasoning via Structured Self-Distillation cites this paper.

VISD: Enhancing Video Reasoning via Structured Self-Distillation V-STaR: Benchmarking Video-LLMs on Video Spatio-Temporal Reasoning

Reference 7

Resolution
verified exact
arxiv_id, observed 2026-05-11T18:46:08.696801Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-08T14:06:27.953376Z digest=sha256:27e333d81407edcefae5be1b5d7744ce8572431efc9c6434080a9304334f63c0

Observation 3139deae-cdd6-464c-9dac-4bae4faeca9b · inbound

VISD: Enhancing Video Reasoning via Structured Self-Distillation cites this paper.

VISD: Enhancing Video Reasoning via Structured Self-Distillation V-STaR: Benchmarking Video-LLMs on Video Spatio-Temporal Reasoning

Reference 7

Resolution
verified exact
arxiv_id, observed 2026-05-11T01:50:51.219025Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-11T01:49:41.654207Z digest=sha256:4fddf0e3c15a20dc61e2399fdd00220d51d18b4fab531749cac5949f2fa35a10

Observation e03cac00-ffab-4db3-8d13-f374ff4d5d1c · inbound

VISD: Enhancing Video Reasoning via Structured Self-Distillation cites this paper.

VISD: Enhancing Video Reasoning via Structured Self-Distillation V-STaR: Benchmarking Video-LLMs on Video Spatio-Temporal Reasoning

Reference 7

Resolution
verified exact
arxiv_id, observed 2026-05-12T07:11:27.939351Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-12T03:35:59.553683Z digest=sha256:945fbf4301574de2990c7e488768ee0fbce66c30c022c85d8b093f968f0dcbf9

Observation 4d0633f3-9ddf-4e90-90e3-aadf43a69681 · inbound

VISD: Enhancing Video Reasoning via Structured Self-Distillation cites this paper.

VISD: Enhancing Video Reasoning via Structured Self-Distillation V-STaR: Benchmarking Video-LLMs on Video Spatio-Temporal Reasoning

Reference 7

Resolution
verified exact
arxiv_id, observed 2026-05-25T06:10:23.976068Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-25T06:08:19.956833Z digest=sha256:7fe03eeafd928bf4c787949426178fe37771a3a055b432937f3281350259cf50

Observation 575950d7-67f5-4551-84ea-80792507ff47 · inbound

Tracking the Truth: Object-Centric Spatio-Temporal Monitoring for Video Large Language Models cites this paper.

Tracking the Truth: Object-Centric Spatio-Temporal Monitoring for Video Large Language Models V-STaR: Benchmarking Video-LLMs on Video Spatio-Temporal Reasoning

Reference 4

Resolution
verified exact
arxiv_id, observed 2026-05-12T07:36:32.249808Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-12T02:31:40.463891Z digest=sha256:61a2e033023d5d57939af01acb1c572a53d520b52e737d5c9ef51e4657dfd4e5

Observation 99feebf8-132f-4a55-8ed9-5c6316bba8ce · inbound

STAR: Failure-Aware Markovian Routing for Multi-Agent Spatiotemporal Reasoning cites this paper.

STAR: Failure-Aware Markovian Routing for Multi-Agent Spatiotemporal Reasoning V-STaR: Benchmarking Video-LLMs on Video Spatio-Temporal Reasoning

Reference 3

Resolution
verified exact
arxiv_id, observed 2026-05-12T07:31:25.499559Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-12T02:42:42.419975Z digest=sha256:ba2b8960ef629901d3ce8a772beaf8066a379c742f767604d6ec924fdde8564e

Observation 60f3c841-2f41-4531-bd2f-6af2879bfb36 · inbound

STAR: Failure-Aware Markovian Routing for Multi-Agent Spatiotemporal Reasoning cites this paper.

STAR: Failure-Aware Markovian Routing for Multi-Agent Spatiotemporal Reasoning V-STaR: Benchmarking Video-LLMs on Video Spatio-Temporal Reasoning

Reference 3

Resolution
verified exact
arxiv_id, observed 2026-05-13T03:52:12.584242Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-13T03:50:25.089976Z digest=sha256:cf49fc20804302032d7b4b5e54a92da520a13dfb732e1936b639ff301dd532df

Observation bdfd50cc-c3a1-45ab-a806-613502762a66 · inbound

STAR: Failure-Aware Markovian Routing for Multi-Agent Spatiotemporal Reasoning cites this paper.

STAR: Failure-Aware Markovian Routing for Multi-Agent Spatiotemporal Reasoning V-STaR: Benchmarking Video-LLMs on Video Spatio-Temporal Reasoning

Reference 3

Resolution
verified exact
arxiv_id, observed 2026-05-19T14:47:36.112253Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-19T14:46:40.165981Z digest=sha256:46add2d5ab5877bcefdf9d7ab057d14821d29865cb0904f1d5070bda73de9e69

Observation 43114469-3aed-4f9b-93cc-492c154b77da · inbound

Leveraging Latent Visual Reasoning in Silence cites this paper.

Leveraging Latent Visual Reasoning in Silence V-STaR: Benchmarking Video-LLMs on Video Spatio-Temporal Reasoning

Reference 4

Resolution
verified exact
arxiv_id, observed 2026-05-20T10:28:11.881438Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-20T10:27:13.859216Z digest=sha256:2b10d7569d30eaae0aa582fe16eb65c0c59561d107785b28688b9fc488f839ce

Observation b161ec6d-aff5-4fb8-a9d6-aba3e7bfb6e1 · inbound

EgoCoT-Bench: Benchmarking Grounded and Verifiable Operation-Centric Chain of Thought Reasoning for MLLMs cites this paper.

EgoCoT-Bench: Benchmarking Grounded and Verifiable Operation-Centric Chain of Thought Reasoning for MLLMs V-STaR: Benchmarking Video-LLMs on Video Spatio-Temporal Reasoning

Reference 8

Resolution
metadata mismatch
arxiv_id, observed 2026-05-20T05:53:22.276707Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-20T05:53:05.450946Z digest=sha256:607e7a04b7d487ef4a68f93064eb171b0fb4d0d61b6f41346142f083ada478c6

Observation acc17bc1-d23a-4fb0-bffd-c0c25bd9709d · inbound

Maestro: Reinforcement Learning to Orchestrate Hierarchical Model-Skill Ensembles cites this paper.

Maestro: Reinforcement Learning to Orchestrate Hierarchical Model-Skill Ensembles V-STaR: Benchmarking Video-LLMs on Video Spatio-Temporal Reasoning

Reference 9

Resolution
verified exact
arxiv_id, observed 2026-05-22T07:51:15.464644Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-22T07:51:13.362986Z digest=sha256:75170bf5bb7ab50568b620b46b5591fd438596395f3929a11a6d25e39e4aa13b

Observation 4118160a-482e-4743-9a85-123aa026a9e0 · inbound

X-Stream: Exploring MLLMs as Multiplexers for Multi-Stream Understanding cites this paper.

X-Stream: Exploring MLLMs as Multiplexers for Multi-Stream Understanding V-STaR: Benchmarking Video-LLMs on Video Spatio-Temporal Reasoning

Reference 13

Resolution
metadata mismatch
arxiv_id, observed 2026-07-01T22:46:18.941027Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-06-28T15:06:22.102725Z digest=sha256:05ea1d58f281980303a081274dd8ff6721d44aabf76bb94b3f06089eb39b8fb2

Observation f42b710a-b94a-43a5-a40c-43ae22d8d8c4 · inbound

X-Stream: Exploring MLLMs as Multiplexers for Multi-Stream Understanding cites this paper.

X-Stream: Exploring MLLMs as Multiplexers for Multi-Stream Understanding V-STaR: Benchmarking Video-LLMs on Video Spatio-Temporal Reasoning

Reference 13

Resolution
metadata mismatch
arxiv_id, observed 2026-06-30T10:44:36.412491Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-06-30T10:42:37.401221Z digest=sha256:b43d00bb37c57494fb97a4f13e7fb6ba2655b95ceabf5fc5f2dd4bc1ecc03ea7

Observation 3d6a04cc-ad82-4abd-948a-7b359b204343 · inbound

CARE: Competence-Aware Reward Shaping for Adaptive Reasoning Length in Video-MLLMs cites this paper.

CARE: Competence-Aware Reward Shaping for Adaptive Reasoning Length in Video-MLLMs V-STaR: Benchmarking Video-LLMs on Video Spatio-Temporal Reasoning

Reference 16

Resolution
verified exact
arxiv_id, observed 2026-07-04T03:29:31.469215Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-06-26T17:56:34.203135Z digest=sha256:c06b07d2a1aaeb81cf80da34dc2681f6eca5faa1c89c412b8e97d4bb2f1e221c

Observation 8187d767-d0f0-47db-a4b8-d683cccb1b57 · inbound

Video-MME-Logical: A Controlled Diagnostic Benchmark for Video Temporal-Logical Reasoning cites this paper.

Video-MME-Logical: A Controlled Diagnostic Benchmark for Video Temporal-Logical Reasoning V-STaR: Benchmarking Video-LLMs on Video Spatio-Temporal Reasoning

Reference 28

Resolution
metadata mismatch
arxiv_id, observed 2026-06-29T19:13:52.945728Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-06-29T04:53:25.259840Z digest=sha256:6a7f845e81188c893935fcfa739f1a554753d18d1e960fb91f3afcb11d5074c4

Observation 30754d7d-7d56-47de-8646-0b001dca03f1 · inbound

AnyGroundBench: A Specialized-Domain Benchmark for Video Grounding in Vision-Language Models cites this paper.

AnyGroundBench: A Specialized-Domain Benchmark for Video Grounding in Vision-Language Models V-STaR: Benchmarking Video-LLMs on Video Spatio-Temporal Reasoning

Reference 7

Resolution
verified exact
arxiv_id, observed 2026-07-03T15:48:34.882701Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-07-03T15:46:26.401577Z digest=sha256:ca8e9dbf4c5182fa67a22defc3b5265b2865b2c3938640779a8a0db835df8c9b