Pith. sign in

Paper Citation Record · LEDGER

Vad-R1: Towards Video Anomaly Reasoning via Perception-to-Cognition Chain-of-Thought

As of 9 August 2026, this Paper Citation Record lists 94 of 94 outbound references and 5 inbound Pith citation observations for arXiv:2505.19877.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2505.19877 v1

Coverage vector

measured 94 of 94 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T14:09:16.669323Z

measured 99 of 99 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00

measured 5 of 5 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-06T13:30:09.804400Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-22T07:36:13.968526Z

Reference resolution

94 of 94 outbound references displayed

  • verified exact0
  • verified fuzzy40
  • unresolved53
  • parse uncertain0
  • malformed identifier1
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 74b36822-cd20-42bb-ad5f-b0b31c146b13 · outbound

This paper cites Ubnormal: New benchmark for supervised open-set video anomaly detection.

Vad-R1: Towards Video Anomaly Reasoning via Perception-to-Cognition Chain-of-Thought Ubnormal: New benchmark for supervised open-set video anomaly detection

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-07T14:09:05.474749Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:09:05.474749Z digest=sha256:7b92c0df3fd4816d3145ec3326f864cfef0747544073de37e076512e0acbb19c

Observation 0a812610-d509-40fb-af09-728b338fb3e6 · outbound

This paper cites Claude 3.5 haiku, 2024.

Vad-R1: Towards Video Anomaly Reasoning via Perception-to-Cognition Chain-of-Thought Claude 3.5 haiku, 2024

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-07T14:09:05.615261Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:09:05.615261Z digest=sha256:51a93867b4e3530130750874ba3ec8a5093de33f5604ce77d8eee9c751940584

Observation 26d5540a-20ea-4024-88c9-e95a8ebe8c35 · outbound

This paper cites Meteor: An automatic metric for mt evaluation with improved correlation with human judgments.

Vad-R1: Towards Video Anomaly Reasoning via Perception-to-Cognition Chain-of-Thought Meteor: An automatic metric for mt evaluation with improved correlation with human judgments

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-07T14:09:05.755439Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:09:05.755439Z digest=sha256:9cb95e6e32b3b9d8a09e8ef19e29dc36dc295f4edf23ca1adcf361a24cbb498a

Observation a97f2be7-6769-4435-b10e-77d3507e57ca · outbound

This paper cites Video generation models as world simulators.

Vad-R1: Towards Video Anomaly Reasoning via Perception-to-Cognition Chain-of-Thought Video generation models as world simulators

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-07T14:09:05.848111Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:09:05.848111Z digest=sha256:a3a2aa9abb2ac143e46688615c7a9c6437c75cc4f2208f31e46295899c5f6a57

Observation 007a6b95-d44e-496d-bac1-f5ee3bd80d15 · outbound

This paper cites A new comprehensive benchmark for semi-supervised video anomaly detection and anticipation.

Vad-R1: Towards Video Anomaly Reasoning via Perception-to-Cognition Chain-of-Thought A new comprehensive benchmark for semi-supervised video anomaly detection and anticipation

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-07T14:09:05.965114Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:09:05.965114Z digest=sha256:318064e88a29e1c0285ac51250b9e1a912df8cdb952de75ce2edd07628cfe6e6

Observation 75f469a1-55fb-4e12-bfdd-de9b17badfb2 · outbound

This paper cites Videollm-online: Online video large language model for streaming video.

Vad-R1: Towards Video Anomaly Reasoning via Perception-to-Cognition Chain-of-Thought Videollm-online: Online video large language model for streaming video

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-07T14:09:06.064765Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:09:06.064765Z digest=sha256:8e7ed21fcf95ad4c00eea850c2f6bebccc344c7c301d30b57924d4b52ed73720

Observation 4c726091-ed65-48b8-bb34-4b095c0e1dfe · outbound

This paper cites Prompt-enhanced multiple instance learning for weakly supervised video anomaly detection.

Vad-R1: Towards Video Anomaly Reasoning via Perception-to-Cognition Chain-of-Thought Prompt-enhanced multiple instance learning for weakly supervised video anomaly detection

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-07T14:09:06.185678Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:09:06.185678Z digest=sha256:ff7b6e77a48bd41d9349f67aa3ae4f30f2ccd9f43d82eef8da2a4a7b2e76b141

Observation e1e0d1f6-85ae-4ca4-9fe3-18bed1997561 · outbound

This paper cites Tevad: Improved video anomaly detection with captions.

Vad-R1: Towards Video Anomaly Reasoning via Perception-to-Cognition Chain-of-Thought Tevad: Improved video anomaly detection with captions

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-07T14:09:06.335558Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:09:06.335558Z digest=sha256:264f10bb529ca1daf1bee33248e518cb39f7ab1db4d4c811487f3524d84098c0

Observation 2bea092f-ab99-4ae8-8334-e2ad17b0e165 · outbound

This paper cites DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning.

Vad-R1: Towards Video Anomaly Reasoning via Perception-to-Cognition Chain-of-Thought DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-07T14:09:06.452918Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:09:06.452918Z digest=sha256:2c8d4a22e25fd4a8da754fcb8e162f30cdb7731ce0c502a5484c50fb0faf9c26

Observation 6fd8b443-55a2-4713-a11c-1352c1a8819a · outbound

This paper cites Streaming Video Question-Answering with In-context Video KV-Cache Retrieval.

Vad-R1: Towards Video Anomaly Reasoning via Perception-to-Cognition Chain-of-Thought Streaming Video Question-Answering with In-context Video KV-Cache Retrieval

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-07T14:09:06.602730Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:09:06.602730Z digest=sha256:8fb3ec30f013118ac08c0c87802fbd412fa62d08f3b483b7d6ce067d85bb6e67

Observation 6f6e1000-48b4-49d2-9084-00d401747976 · outbound

This paper cites SlowFastVAD: Video Anomaly Detection via Integrating Simple Detector and RAG-Enhanced Vision-Language Model.

Vad-R1: Towards Video Anomaly Reasoning via Perception-to-Cognition Chain-of-Thought SlowFastVAD: Video Anomaly Detection via Integrating Simple Detector and RAG-Enhanced Vision-Language Model

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-07T14:09:06.744774Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:09:06.744774Z digest=sha256:a8dc5b0fc52ccc41dead612c91f3decc98fb98a7a22fc4a01b6ef026f5910135

Observation d9af6d20-f371-4888-a015-7a7e9556afd8 · outbound

This paper cites Exploring What Why and How: A Multifaceted Benchmark for Causation Understanding of Video Anomaly.

Vad-R1: Towards Video Anomaly Reasoning via Perception-to-Cognition Chain-of-Thought Exploring What Why and How: A Multifaceted Benchmark for Causation Understanding of Video Anomaly

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-07T14:09:06.865811Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:09:06.865811Z digest=sha256:69db0423af83938167c83c7455e90712a7e862ae8cfae031aefa7f4210dd6893

Observation ae150774-d5e4-47b9-ad27-e5ab8f701bcb · outbound

This paper cites Uncovering what why and how: A comprehensive benchmark for causation understanding of video anomaly.

Vad-R1: Towards Video Anomaly Reasoning via Perception-to-Cognition Chain-of-Thought Uncovering what why and how: A comprehensive benchmark for causation understanding of video anomaly

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-07T14:09:06.941062Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:09:06.941062Z digest=sha256:ba142098451311efd82589aca0a8f5e777e48e8209bc3b37be087188a96fbeff

Observation c756657f-8bd2-4e55-a4e4-61d6835b3b62 · outbound

This paper cites Video-R1: Reinforcing Video Reasoning in MLLMs.

Vad-R1: Towards Video Anomaly Reasoning via Perception-to-Cognition Chain-of-Thought Video-R1: Reinforcing Video Reasoning in MLLMs

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-07T14:09:07.074909Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:09:07.074909Z digest=sha256:1130326e24dac253237cee85dc00d15667f66bd1356d95e34284f68039f3ec2f

Observation a502e150-fc80-453c-bf54-11fb568ebb92 · outbound

This paper cites Vane-bench: Video anomaly evaluation benchmark for conversational lmms.

Vad-R1: Towards Video Anomaly Reasoning via Perception-to-Cognition Chain-of-Thought Vane-bench: Video anomaly evaluation benchmark for conversational lmms

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-07T14:09:07.204754Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:09:07.204754Z digest=sha256:a3c2fa76c490814655bfc4a9b6afd919f4f1007c6a3f10475d503898ecdbe967

Observation 488b2ac5-11a5-4341-8c92-923f5ce21e6e · outbound

This paper cites Open-sora: Democratizing efficient video production for all.

Vad-R1: Towards Video Anomaly Reasoning via Perception-to-Cognition Chain-of-Thought Open-sora: Democratizing efficient video production for all

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-07T14:09:07.324835Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:09:07.324835Z digest=sha256:f58761757d1eb3ebc95d58114c580f1f2a40e5b4d5e28bfe94b34cef8d49840d

Observation 0c1d6677-8766-4f32-9967-07573de7014b · outbound

This paper cites Abnormal event detection using deep contrastive learning for intelligent video surveillance system.

Vad-R1: Towards Video Anomaly Reasoning via Perception-to-Cognition Chain-of-Thought Abnormal event detection using deep contrastive learning for intelligent video surveillance system

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-07T14:09:07.457830Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:09:07.457830Z digest=sha256:370d8f3ee8d294bdda29fc9e526671feba6abd8baecb376d304939657c57b676

Observation a6fdc99f-7990-47d5-9de8-e07f8c1467a5 · outbound

This paper cites Weakly supervised video anomaly detection via self-guided temporal discriminative transformer.

Vad-R1: Towards Video Anomaly Reasoning via Perception-to-Cognition Chain-of-Thought Weakly supervised video anomaly detection via self-guided temporal discriminative transformer

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-07T14:09:07.584828Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:09:07.584828Z digest=sha256:a51fb9cf4b79e6f317089dd1872e908b7840a813a5c22459d556e1af8138283b

Observation e92d8aac-4946-45a9-a887-1d6e19654b5b · outbound

This paper cites Self-supervised attentive generative adversarial networks for video anomaly detection.

Vad-R1: Towards Video Anomaly Reasoning via Perception-to-Cognition Chain-of-Thought Self-supervised attentive generative adversarial networks for video anomaly detection

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-07T14:09:07.694327Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:09:07.694327Z digest=sha256:94242b6af16f106bd380028a23ef3da010617a1852d266bb2fde4cf9df020e11

Observation 55e4a19f-f14a-4cd2-8baf-16877e66bd40 · outbound

This paper cites Long short-term dynamic prototype alignment learning for video anomaly detection.

Vad-R1: Towards Video Anomaly Reasoning via Perception-to-Cognition Chain-of-Thought Long short-term dynamic prototype alignment learning for video anomaly detection

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:09:26.839108Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T14:09:07.844750Z digest=sha256:b51d6a0bd2267e662c1e766677b667cfd2db80ddbe24115b0b3f927e06965a04

Observation df11acfd-9dd1-408b-8acd-3f5fef27bc4d · outbound

This paper cites Multi- modal evidential learning for open-world weakly-supervised video anomaly detection.

Vad-R1: Towards Video Anomaly Reasoning via Perception-to-Cognition Chain-of-Thought Multi- modal evidential learning for open-world weakly-supervised video anomaly detection

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:09:26.552206Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T14:09:07.974751Z digest=sha256:ef0ce54f3fda8bb184b5c35012059983055ae0967069b6f4da7db0dc5d5e8b33

Observation 5693514b-b470-4e39-8cc5-c8eb8e252e09 · outbound

This paper cites Vision-R1: Incentivizing Reasoning Capability in Multimodal Large Language Models.

Vad-R1: Towards Video Anomaly Reasoning via Perception-to-Cognition Chain-of-Thought Vision-R1: Incentivizing Reasoning Capability in Multimodal Large Language Models

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-07T14:09:08.062129Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:09:08.062129Z digest=sha256:bb103bb5b0f7d1f5e7c8bec56bcdc29c916de1487e631978a4ccdfdeaefbe19d

Observation ea05da91-16eb-45dc-9691-113ca936dd84 · outbound

This paper cites Chat-univi: Unified visual representation empowers large language models with image and video understanding.

Vad-R1: Towards Video Anomaly Reasoning via Perception-to-Cognition Chain-of-Thought Chat-univi: Unified visual representation empowers large language models with image and video understanding

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:09:26.217216Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T14:09:08.224753Z digest=sha256:aed18c9c47cb38e81d55e91f7c23e985b173418a1dfc17f2fda0e4d66733d32b

Observation bd4c24f3-0769-4fee-a6d4-4e3fc60cebda · outbound

This paper cites Clip-tsa: Clip-assisted temporal self-attention for weakly-supervised video anomaly detection.

Vad-R1: Towards Video Anomaly Reasoning via Perception-to-Cognition Chain-of-Thought Clip-tsa: Clip-assisted temporal self-attention for weakly-supervised video anomaly detection

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:09:25.929611Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T14:09:08.484751Z digest=sha256:acaeb886a7ae6d821a75eb8d941c8a93dc629217d6e4eae18691e749fc1209dd

Observation 055477c3-d512-49c2-84aa-4834b945b15b · outbound

This paper cites VideoChat: Chat-Centric Video Understanding.

Vad-R1: Towards Video Anomaly Reasoning via Perception-to-Cognition Chain-of-Thought VideoChat: Chat-Centric Video Understanding

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-07T14:09:08.616938Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:09:08.616938Z digest=sha256:052642137da28bb76a38c81738d7936b004be7889d64ba6bcabe728447bb76f8

Observation 43d8a0f8-915c-4c81-9e1c-28d63e786563 · outbound

This paper cites Anomaly detection and localization in crowded scenes.

Vad-R1: Towards Video Anomaly Reasoning via Perception-to-Cognition Chain-of-Thought Anomaly detection and localization in crowded scenes

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:09:25.694757Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T14:09:08.795213Z digest=sha256:a4f3edd3f83f3bd7795df3c715422d91c7b7cc540ded21b58d741089f92bb2b5

Observation 7cbb5882-cc3d-430f-b987-5b744e65f3ef · outbound

This paper cites VideoChat-Flash: Hierarchical Compression for Long-Context Video Modeling.

Vad-R1: Towards Video Anomaly Reasoning via Perception-to-Cognition Chain-of-Thought VideoChat-Flash: Hierarchical Compression for Long-Context Video Modeling

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-07T14:09:08.944838Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:09:08.944838Z digest=sha256:76d948daad115d23bd03a96a444c51a0e81cbab1eebab5828131778684a1b40c

Observation 553478bb-aace-4b35-baf5-b767d1820f7c · outbound

This paper cites VideoChat-R1: Enhancing Spatio-Temporal Perception via Reinforcement Fine-Tuning.

Vad-R1: Towards Video Anomaly Reasoning via Perception-to-Cognition Chain-of-Thought VideoChat-R1: Enhancing Spatio-Temporal Perception via Reinforcement Fine-Tuning

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-07T14:09:09.154751Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:09:09.154751Z digest=sha256:bf4baa6e622a25ded071a15ec0b27d41d6a88dda8b9258d0c02f3963fdeeca46

Observation 2d68afd5-ad4a-4810-ad0e-9ef78d163385 · outbound

This paper cites Llama-vid: An image is worth 2 tokens in large language models.

Vad-R1: Towards Video Anomaly Reasoning via Perception-to-Cognition Chain-of-Thought Llama-vid: An image is worth 2 tokens in large language models

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:09:25.440900Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T14:09:09.259223Z digest=sha256:52d230a49fb52d7c2cf319d0c93549bce1ae5242e5f5c901928a9390c2b71e99

Observation 691657a4-6481-4c8d-8176-03358e97511b · outbound

This paper cites Video-LLaVA: Learning United Visual Representation by Alignment Before Projection.

Vad-R1: Towards Video Anomaly Reasoning via Perception-to-Cognition Chain-of-Thought Video-LLaVA: Learning United Visual Representation by Alignment Before Projection

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-07T14:09:09.355173Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:09:09.355173Z digest=sha256:d8c7f2fb9e72f1092151d67d716d8e660c9a6868c863cc20ea55c568811d5e47

Observation 6d5d9df4-6d10-4750-9f61-435e7b11e8c2 · outbound

This paper cites Rouge: A package for automatic evaluation of summaries.

Vad-R1: Towards Video Anomaly Reasoning via Perception-to-Cognition Chain-of-Thought Rouge: A package for automatic evaluation of summaries

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-07T14:09:09.420025Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:09:09.420025Z digest=sha256:d0b62a26e9aa93edf44d53c9a40f15d0832726c86e03a9e5e00c9a5256bf1b48

Observation 4ebeb344-590d-49e2-9024-331040fdedd9 · outbound

This paper cites Future frame prediction for anomaly detection–a new baseline.

Vad-R1: Towards Video Anomaly Reasoning via Perception-to-Cognition Chain-of-Thought Future frame prediction for anomaly detection–a new baseline

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:09:25.114118Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T14:09:09.505566Z digest=sha256:1d20930320b0eaac7b3cb074d11ec77ed57f594b6dfef412b4d1a7225bf6a95a

Observation 05a3f08a-54d7-4589-8323-56bb18885531 · outbound

This paper cites Videomind: A chain-of-lora agent for long video reasoning.

Vad-R1: Towards Video Anomaly Reasoning via Perception-to-Cognition Chain-of-Thought Videomind: A chain-of-lora agent for long video reasoning

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-07T14:09:09.564739Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:09:09.564739Z digest=sha256:0407a412b4cf81088d4ea58e11a4402316c83d0ef73db06e8d69c8b9c2f192cd

Observation 7c41ad22-a1dd-4d6a-9c10-72f8d536cd85 · outbound

This paper cites A hybrid video anomaly detection framework via memory-augmented flow reconstruction and flow-guided frame prediction.

Vad-R1: Towards Video Anomaly Reasoning via Perception-to-Cognition Chain-of-Thought A hybrid video anomaly detection framework via memory-augmented flow reconstruction and flow-guided frame prediction

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:09:24.904990Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T14:09:09.627328Z digest=sha256:14b0bb07562a6b5507c9685f4e8822897dcc9920923c6e882f229ed7df4deef5

Observation fa7210e4-9495-44ff-8a9e-c199bdd8a1f7 · outbound

This paper cites Abnormal event detection at 150 fps in matlab.

Vad-R1: Towards Video Anomaly Reasoning via Perception-to-Cognition Chain-of-Thought Abnormal event detection at 150 fps in matlab

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:09:24.666630Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T14:09:09.713530Z digest=sha256:63acd2c47d49ec61ef5ce919031cd9ea7bde22a70e1e6a1661316aab1620b554

Observation e63d5cbe-9a66-4311-9a70-570cae6a6718 · outbound

This paper cites Video Anomaly Detection and Explanation via Large Language Models.

Vad-R1: Towards Video Anomaly Reasoning via Perception-to-Cognition Chain-of-Thought Video Anomaly Detection and Explanation via Large Language Models

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-07T14:09:09.799641Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:09:09.799641Z digest=sha256:8ff200a6ddca02e6b27a3b72162797c554ae1288cc80839d7e46e42e22ef769a

Observation b1165a0b-ab5b-460b-954b-803afd48c95a · outbound

This paper cites Localizing anomalies from weakly-labeled videos.

Vad-R1: Towards Video Anomaly Reasoning via Perception-to-Cognition Chain-of-Thought Localizing anomalies from weakly-labeled videos

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:09:24.434697Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T14:09:09.912993Z digest=sha256:dcf074835ed19dfa2be07ae18e7dc5f62f402f3fa9a466d27cc50aede47733c6

Observation b1f53c55-f8f8-4cc3-8378-0ffb6a2ddee9 · outbound

This paper cites Sherlock: Towards multi-scene video abnormal event extraction and localization via a global-local spatial-sensitive llm.

Vad-R1: Towards Video Anomaly Reasoning via Perception-to-Cognition Chain-of-Thought Sherlock: Towards multi-scene video abnormal event extraction and localization via a global-local spatial-sensitive llm

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:09:24.199170Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T14:09:10.058849Z digest=sha256:7fa877632e5a14ab0c37437e71362f5e3e6148cdd583a82a33bbbf0b1f5bdc48

Observation b176954c-530a-42a8-bfc5-798933f5caae · outbound

This paper cites Video-ChatGPT: Towards Detailed Video Understanding via Large Vision and Language Models.

Vad-R1: Towards Video Anomaly Reasoning via Perception-to-Cognition Chain-of-Thought Video-ChatGPT: Towards Detailed Video Understanding via Large Vision and Language Models

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-07T14:09:10.196311Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:09:10.196311Z digest=sha256:17c2c6219dfacd80920c9b60b5335653a1f322a306e07df2d694dd5ca2d2314f

Observation 07fa7e2a-920e-4c79-92d3-5dcd7075e8a5 · outbound

This paper cites GPT-4o System Card.

Vad-R1: Towards Video Anomaly Reasoning via Perception-to-Cognition Chain-of-Thought GPT-4o System Card

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-07T14:09:10.302054Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:09:10.302054Z digest=sha256:0cc1dd6b057299640aa6d9c254888aeb7a1506eac900b13049ad8445779b10ba

Observation a5b0c482-f59f-4766-b458-7d87296821ea · outbound

This paper cites OpenAI o1 System Card.

Vad-R1: Towards Video Anomaly Reasoning via Perception-to-Cognition Chain-of-Thought OpenAI o1 System Card

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-07T14:09:10.444749Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:09:10.444749Z digest=sha256:21aaf89c9a2adf97e6962f67d8b3d6c0287451c4cdc0f9c85b70706817bd3a62

Observation ff741184-1820-4969-8740-24d993a00f77 · outbound

This paper cites Openai o3 and o4-mini system card, 2025.

Vad-R1: Towards Video Anomaly Reasoning via Perception-to-Cognition Chain-of-Thought Openai o3 and o4-mini system card, 2025

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:09:23.939270Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T14:09:10.505981Z digest=sha256:a0d8688f45c566e7ba13452d29030194ed85b17ba0413efcab067ef512272be7

Observation c6805c43-c11e-466e-8f1b-c68ee9a2e2c8 · outbound

This paper cites Bleu: a method for automatic evaluation of machine translation.

Vad-R1: Towards Video Anomaly Reasoning via Perception-to-Cognition Chain-of-Thought Bleu: a method for automatic evaluation of machine translation

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:09:23.714214Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T14:09:10.667970Z digest=sha256:ecab4d62c567c2484f8657cf06fa94a2c2f39b81124bb37d3fd8e5f9de939e3d

Observation 97381c00-2406-43a0-a95b-e3928a48e55d · outbound

This paper cites Timechat: A time-sensitive multimodal large language model for long video understanding.

Vad-R1: Towards Video Anomaly Reasoning via Perception-to-Cognition Chain-of-Thought Timechat: A time-sensitive multimodal large language model for long video understanding

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:09:23.495945Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T14:09:10.756335Z digest=sha256:f8dd43d6aa93e7e77fa54ab1e4580ff870cc600a5097554c8ba3ecc6ac7fe2bd

Observation 5f14ff13-2a79-46b5-bef6-5d56de4466cf · outbound

This paper cites Self-distilled masked auto-encoders are efficient video anomaly detectors.

Vad-R1: Towards Video Anomaly Reasoning via Perception-to-Cognition Chain-of-Thought Self-distilled masked auto-encoders are efficient video anomaly detectors

Reference 45

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:09:23.327144Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T14:09:10.887212Z digest=sha256:dc888896f5bda9125b01d3bc504a0a1e16122452286e4184b886c225438d5543

Observation ff2d1d07-f497-4710-a4fb-0c452f49e8c7 · outbound

This paper cites Gen-2: The next step forward for generative ai.

Vad-R1: Towards Video Anomaly Reasoning via Perception-to-Cognition Chain-of-Thought Gen-2: The next step forward for generative ai

Reference 46

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:09:23.141205Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T14:09:11.035451Z digest=sha256:c31a8aa605a7af49e8e7ab599d6e34d798dda5a30bfa338d3a5c8479278010b2

Observation e02fa9d3-ca25-41f5-a676-0994b4c8d26d · outbound

This paper cites DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models.

Vad-R1: Towards Video Anomaly Reasoning via Perception-to-Cognition Chain-of-Thought DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-07T14:09:11.127199Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:09:11.127199Z digest=sha256:e3218829b4b1e470fb8823c3cc8c8bde19654960958d1fc9453e72b0d18dd148

Observation 2614eb40-de40-4d40-b6e9-943ee598bd43 · outbound

This paper cites Moviechat: From dense token to sparse memory for long video understanding.

Vad-R1: Towards Video Anomaly Reasoning via Perception-to-Cognition Chain-of-Thought Moviechat: From dense token to sparse memory for long video understanding

Reference 48

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:09:22.897638Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T14:09:11.208115Z digest=sha256:650e9e5423f5e5f3a033815e45cd610f68650c1cfbb3d3efd1f77017189f549a

Observation afb1d732-1205-4164-9414-c057c4dbcb29 · outbound

This paper cites Real-world anomaly detection in surveillance videos.

Vad-R1: Towards Video Anomaly Reasoning via Perception-to-Cognition Chain-of-Thought Real-world anomaly detection in surveillance videos

Reference 49

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:09:22.698560Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T14:09:11.296667Z digest=sha256:e59b15f1cfbea949f446f50c510c5f0773e0a48d61c5dac072f2aa08c3463c60

Observation 20eda1fa-7d2a-4aa5-b4cd-1d375a38df29 · outbound

This paper cites Hawk: Learning to understand open-world video anomalies.

Vad-R1: Towards Video Anomaly Reasoning via Perception-to-Cognition Chain-of-Thought Hawk: Learning to understand open-world video anomalies

Reference 50

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:09:22.477709Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T14:09:11.427614Z digest=sha256:34e6a172f926bad418470e6b7c959f054b8446f3893b374e18488ab77ef8e569

Observation 6f3f12c2-58de-4482-8856-24dfd35cde15 · outbound

This paper cites Gemini 2.5 flash preview model card, 2025.

Vad-R1: Towards Video Anomaly Reasoning via Perception-to-Cognition Chain-of-Thought Gemini 2.5 flash preview model card, 2025

Reference 51

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:09:22.248638Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T14:09:11.520760Z digest=sha256:06d91bfb9b5a919ef306411555ae7e6d9a692e7f3584db40e05a168115a67765

Observation 2744d918-66fa-439e-aaa5-e79ce0f16cc2 · outbound

This paper cites Gemini 2.5 pro preview model card, 2025.

Vad-R1: Towards Video Anomaly Reasoning via Perception-to-Cognition Chain-of-Thought Gemini 2.5 pro preview model card, 2025

Reference 52

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:09:22.011914Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T14:09:11.656655Z digest=sha256:5cb9d19611484d3ce226ca0f2e568ea44ff59801b6cf33b49d09fc8385853bc6

Observation 8ee43beb-a40e-4c8c-a2b2-37ea29aec736 · outbound

This paper cites Kimi k1.5: Scaling Reinforcement Learning with LLMs.

Vad-R1: Towards Video Anomaly Reasoning via Perception-to-Cognition Chain-of-Thought Kimi k1.5: Scaling Reinforcement Learning with LLMs

Reference 53

Resolution
unresolved
no resolver link, observed 2026-08-07T14:09:11.779386Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:09:11.779386Z digest=sha256:5885f6856e43c2ec146c6f9271ee7d8fdc1794ba0a0d7173ec01fa74c7946d2a

Observation 07226a3d-d355-4915-b1de-611b370b356f · outbound

This paper cites QwQ: Reflect deeply on the boundaries of the unknown, 2024.

Vad-R1: Towards Video Anomaly Reasoning via Perception-to-Cognition Chain-of-Thought QwQ: Reflect deeply on the boundaries of the unknown, 2024

Reference 54

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:09:21.768204Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T14:09:11.905318Z digest=sha256:83864dfc94b61eadbc693f2b2f2546995b81e9e38d62df7011e1052423110f4a

Observation 1f500267-5cd6-4dd0-8b16-da052f6e8663 · outbound

This paper cites Qwen2.5 Technical Report.

Vad-R1: Towards Video Anomaly Reasoning via Perception-to-Cognition Chain-of-Thought Qwen2.5 Technical Report

Reference 55

Resolution
unresolved
no resolver link, observed 2026-08-07T14:09:12.056747Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:09:12.056747Z digest=sha256:ce0a1fb1ca54401793fd4eab4dd0b45c8f07c9c7d136dc60fba787a737f2388d

Observation 1ab641df-1237-4bfe-a4bb-f070f355621d · outbound

This paper cites QVQ-Max: Think with evidence, 2025.

Vad-R1: Towards Video Anomaly Reasoning via Perception-to-Cognition Chain-of-Thought QVQ-Max: Think with evidence, 2025

Reference 56

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:09:21.561752Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T14:09:12.193934Z digest=sha256:4a0afb44e501bb9d2f196e4f9c3063e80ae6a9ebd5060b20357430e40f41313d

Observation 448e5500-10ea-4142-bb64-1724dfe07742 · outbound

This paper cites Qwen2.5-VL Technical Report.

Vad-R1: Towards Video Anomaly Reasoning via Perception-to-Cognition Chain-of-Thought Qwen2.5-VL Technical Report

Reference 57

Resolution
unresolved
no resolver link, observed 2026-08-07T14:09:12.381462Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:09:12.381462Z digest=sha256:2d1a6e2e59cc21f47bac9c6ba65d93d2c9d1479d4eeb5194e9fa4539d6d1ce99

Observation df645d7a-95e1-4249-b0ab-2baea805dd5e · outbound

This paper cites Qwen3: Think deeper, act faster, 2025.

Vad-R1: Towards Video Anomaly Reasoning via Perception-to-Cognition Chain-of-Thought Qwen3: Think deeper, act faster, 2025

Reference 58

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:09:21.383574Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T14:09:12.515845Z digest=sha256:87ad34db859b0b54b8ad234796236a5ace880e3d8ebd6e1de765c9eb1c6e984f

Observation 28b90eaf-839a-41af-9fbf-d26950833de2 · outbound

This paper cites LlamaV-o1: Rethinking Step-by-step Visual Reasoning in LLMs.

Vad-R1: Towards Video Anomaly Reasoning via Perception-to-Cognition Chain-of-Thought LlamaV-o1: Rethinking Step-by-step Visual Reasoning in LLMs

Reference 59

Resolution
unresolved
no resolver link, observed 2026-08-07T14:09:12.592946Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:09:12.592946Z digest=sha256:6bacad7f48db59ab0b5922d0707d30cdcda495987c6ece0166bd7b5dfc2696ec

Observation dd14da3b-8530-400d-aff9-1662944d3cef · outbound

This paper cites Federated weakly supervised video anomaly detection with multimodal prompt.

Vad-R1: Towards Video Anomaly Reasoning via Perception-to-Cognition Chain-of-Thought Federated weakly supervised video anomaly detection with multimodal prompt

Reference 60

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:09:21.196877Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T14:09:12.737393Z digest=sha256:c958e9bdd87a608e4823d4cd816e2d627199bba5c680651d4f90599269114455

Observation d5da0cb7-c562-42fe-a2f2-f59dad0ba0f5 · outbound

This paper cites Modelscope text-to-video technical report, 2023.

Vad-R1: Towards Video Anomaly Reasoning via Perception-to-Cognition Chain-of-Thought Modelscope text-to-video technical report, 2023

Reference 61

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:09:20.893590Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T14:09:12.826604Z digest=sha256:3488bd37b1052d29ca9fcca12ef9f54992776a138b44b381bc1470d1e14d0527

Observation 65b6ee68-ab31-4868-ac25-8f5e2446f386 · outbound

This paper cites Videolcm: Video latent consistency model, 2023.

Vad-R1: Towards Video Anomaly Reasoning via Perception-to-Cognition Chain-of-Thought Videolcm: Video latent consistency model, 2023

Reference 62

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:09:20.667580Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T14:09:12.954393Z digest=sha256:df762f6a89129ed2ba89ce0bfaf4566ff8523aff7085874a4ed9ea128d0e8da5

Observation 81712c2c-aba2-44c1-bd0a-5c9359dedc18 · outbound

This paper cites Open-r1-video, 2025.

Vad-R1: Towards Video Anomaly Reasoning via Perception-to-Cognition Chain-of-Thought Open-r1-video, 2025

Reference 63

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:09:20.421517Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T14:09:13.063090Z digest=sha256:f63106d8747e5bb8cb1a04fd539abc0dbf17173c557ac22f97f465fca7b2478a

Observation e45504bb-270e-4351-b01d-cf44a4d26041 · outbound

This paper cites Time-R1: Post-Training Large Vision Language Model for Temporal Video Grounding.

Vad-R1: Towards Video Anomaly Reasoning via Perception-to-Cognition Chain-of-Thought Time-R1: Post-Training Large Vision Language Model for Temporal Video Grounding

Reference 64

Resolution
unresolved
no resolver link, observed 2026-08-07T14:09:13.173718Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:09:13.173718Z digest=sha256:e9b2bd70953fcbac867bff94958eb48ce2e48eb3945517508cbc10689962a9dc

Observation 01a233a0-30b5-42cc-a9b7-5a273bf9102d · outbound

This paper cites InternVideo2.5: Empowering Video MLLMs with Long and Rich Context Modeling.

Vad-R1: Towards Video Anomaly Reasoning via Perception-to-Cognition Chain-of-Thought InternVideo2.5: Empowering Video MLLMs with Long and Rich Context Modeling

Reference 65

Resolution
unresolved
no resolver link, observed 2026-08-07T14:09:13.309276Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:09:13.309276Z digest=sha256:7daa0992f912a4b8bc89516358edfeceee3955d16357c99cb7151be7e3dd1881

Observation a4d1e605-c307-44f1-80bf-b179812f6c30 · outbound

This paper cites Not only look, but also listen: Learning multimodal violence detection under weak supervision.

Vad-R1: Towards Video Anomaly Reasoning via Perception-to-Cognition Chain-of-Thought Not only look, but also listen: Learning multimodal violence detection under weak supervision

Reference 66

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:09:20.273475Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T14:09:13.455450Z digest=sha256:26a0e01a498aab522765cc9fe2c6d8a2f4e695fd850ea4104230d1ce3fed6b18

Observation 7a19cb88-7667-4f19-9b35-15bf95a02bce · outbound

This paper cites Open-vocabulary video anomaly detection.

Vad-R1: Towards Video Anomaly Reasoning via Perception-to-Cognition Chain-of-Thought Open-vocabulary video anomaly detection

Reference 67

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:09:20.056783Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T14:09:13.666971Z digest=sha256:2806ff72001f3bfb89d92f8de7b2e406b17ad1e08f4d6a917e9192745877ec74

Observation c88bff23-6e4c-42ae-919e-f53d82d146fc · outbound

This paper cites Vadclip: Adapting vision-language models for weakly supervised video anomaly detection.

Vad-R1: Towards Video Anomaly Reasoning via Perception-to-Cognition Chain-of-Thought Vadclip: Adapting vision-language models for weakly supervised video anomaly detection

Reference 68

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:09:19.794855Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T14:09:13.834821Z digest=sha256:7a29523c2d0f5ae462a1f792535ba6bcf5759e0f81f0298d1886394ab6de3746

Observation e57d3da4-cce2-4184-93a0-e8ea7e0e0757 · outbound

This paper cites Streaming Video Understanding and Multi-round Interaction with Memory-enhanced Knowledge.

Vad-R1: Towards Video Anomaly Reasoning via Perception-to-Cognition Chain-of-Thought Streaming Video Understanding and Multi-round Interaction with Memory-enhanced Knowledge

Reference 69

Resolution
unresolved
no resolver link, observed 2026-08-07T14:09:13.898584Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:09:13.898584Z digest=sha256:402fe0966763e60ae7b7f00f38ab62ad55e92c1ebd89f67d139f2f4dade6f4ca

Observation 0b2d7b0d-c14e-40c1-afad-4bb99c6a607e · outbound

This paper cites LLaVA-CoT: Let Vision Language Models Reason Step-by-Step.

Vad-R1: Towards Video Anomaly Reasoning via Perception-to-Cognition Chain-of-Thought LLaVA-CoT: Let Vision Language Models Reason Step-by-Step

Reference 70

Resolution
unresolved
no resolver link, observed 2026-08-07T14:09:13.996611Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:09:13.996611Z digest=sha256:84c7f5f2b2b11b6881c569f2c9c171496aac704826a5f4b38a9146e0a888dab0

Observation acfe4bdc-31fb-4919-a07b-0e12ef58cd21 · outbound

This paper cites PLLaVA : Parameter-free LLaVA Extension from Images to Videos for Video Dense Captioning.

Vad-R1: Towards Video Anomaly Reasoning via Perception-to-Cognition Chain-of-Thought PLLaVA : Parameter-free LLaVA Extension from Images to Videos for Video Dense Captioning

Reference 71

Resolution
unresolved
no resolver link, observed 2026-08-07T14:09:14.086670Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:09:14.086670Z digest=sha256:c200b3fc943f19e6fe1889e3c192d1c1379f0fbb5196611f472ebe376a1f2761

Observation a43e15e6-d262-46ed-9cc9-c4d206fa7355 · outbound

This paper cites Feature prediction diffusion model for video anomaly detection.

Vad-R1: Towards Video Anomaly Reasoning via Perception-to-Cognition Chain-of-Thought Feature prediction diffusion model for video anomaly detection

Reference 72

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:09:19.515052Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T14:09:14.200572Z digest=sha256:89c2cc0e554ccf74df784f2ee6ffe06d3462fd3d10ca280499c61f0446bb1926

Observation 109c2b18-687d-4567-abbd-242fca3d7672 · outbound

This paper cites Follow the rules: reasoning for video anomaly detection with large language models.

Vad-R1: Towards Video Anomaly Reasoning via Perception-to-Cognition Chain-of-Thought Follow the rules: reasoning for video anomaly detection with large language models

Reference 73

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:09:19.308536Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T14:09:14.315573Z digest=sha256:bb033029efbcf3a62521d6f5080685460c10d64586ba0d7a500086e8694cb953

Observation 7da2edef-5179-49bc-a2d6-39b27bc27b8f · outbound

This paper cites Svbench: A benchmark with temporal multi-turn dialogues for streaming video understanding.

Vad-R1: Towards Video Anomaly Reasoning via Perception-to-Cognition Chain-of-Thought Svbench: A benchmark with temporal multi-turn dialogues for streaming video understanding

Reference 74

Resolution
unresolved
no resolver link, observed 2026-08-07T14:09:14.364645Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:09:14.364645Z digest=sha256:06400df6ab7044232f1b5c48db8424d9d07310ba7766ef19ea094c843bb6fad5

Observation 8cd34668-23b6-4bb2-9314-b1f26960f6b0 · outbound

This paper cites Dota: Unsupervised detection of traffic anomaly in driving videos.IEEE transactions on pattern analysis and machine intelligence, 45(1):444–459, 2022.

Vad-R1: Towards Video Anomaly Reasoning via Perception-to-Cognition Chain-of-Thought Dota: Unsupervised detection of traffic anomaly in driving videos.IEEE transactions on pattern analysis and machine intelligence, 45(1):444–459, 2022

Reference 75

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:09:19.059976Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T14:09:14.550526Z digest=sha256:9b82fa757f5cd5ae8d0a365d8b8af136520cc31dc3579bb150951d45dce2348c

Observation 92a88d12-b3c7-4dae-8ef0-d5de816ef36a · outbound

This paper cites VERA: Explainable Video Anomaly Detection via Verbalized Learning of Vision-Language Models.

Vad-R1: Towards Video Anomaly Reasoning via Perception-to-Cognition Chain-of-Thought VERA: Explainable Video Anomaly Detection via Verbalized Learning of Vision-Language Models

Reference 76

Resolution
unresolved
no resolver link, observed 2026-08-07T14:09:14.656477Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:09:14.656477Z digest=sha256:9ef195f583e2ee309658b3adf0968eac178584e03d548647ead76471f406175a

Observation a22424f5-40a8-4045-8435-a6ee1dab4d81 · outbound

This paper cites Unhackable Temporal Rewarding for Scalable Video MLLMs.

Vad-R1: Towards Video Anomaly Reasoning via Perception-to-Cognition Chain-of-Thought Unhackable Temporal Rewarding for Scalable Video MLLMs

Reference 77

Resolution
unresolved
no resolver link, observed 2026-08-07T14:09:14.787236Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:09:14.787236Z digest=sha256:235bb1dfd11a477f6fe6a590dbae655c8274b31deea99073ca700ea46f53a368

Observation 1f712552-3535-42cd-9da4-ffb909c68cd5 · outbound

This paper cites Towards surveillance video-and-language understanding: New dataset baselines and challenges.

Vad-R1: Towards Video Anomaly Reasoning via Perception-to-Cognition Chain-of-Thought Towards surveillance video-and-language understanding: New dataset baselines and challenges

Reference 78

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:09:18.816692Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T14:09:14.937077Z digest=sha256:a962d3e6093c2e3bc568b59ba54065dd5b26438b41462885139eb348085c858d

Observation deaa7724-d981-4235-b62b-1accd7085ef8 · outbound

This paper cites Generative cooperative learning for unsupervised video anomaly detection.

Vad-R1: Towards Video Anomaly Reasoning via Perception-to-Cognition Chain-of-Thought Generative cooperative learning for unsupervised video anomaly detection

Reference 79

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:09:18.571228Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T14:09:15.060104Z digest=sha256:2946d8a2c717be35bb0d1d553b323cae29fe0dff31f5113c093381d8e96f5b63

Observation c8ec1c8e-1397-447d-a78f-c2c221af5bab · outbound

This paper cites Har- nessing large language models for training-free video anomaly detection.

Vad-R1: Towards Video Anomaly Reasoning via Perception-to-Cognition Chain-of-Thought Har- nessing large language models for training-free video anomaly detection

Reference 80

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:09:18.409732Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T14:09:15.166896Z digest=sha256:654b7503db2a7c32b97f40acbd5b45c9182c1fbb601e6059bd0d17292f2bb3af

Observation 850ec3cd-0c61-4c3f-8211-822fc8fc9f39 · outbound

This paper cites Vision-R1: Evolving Human-Free Alignment in Large Vision-Language Models via Vision-Guided Reinforcement Learning.

Vad-R1: Towards Video Anomaly Reasoning via Perception-to-Cognition Chain-of-Thought Vision-R1: Evolving Human-Free Alignment in Large Vision-Language Models via Vision-Guided Reinforcement Learning

Reference 81

Resolution
unresolved
no resolver link, observed 2026-08-07T14:09:15.259254Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:09:15.259254Z digest=sha256:dc1bba7ac61fba4d7904c04d10cfc8eaf0e3197fb0434bde8e09977bc0356d0c

Observation 266dcb77-7c01-4542-9de0-21ee260189c1 · outbound

This paper cites VideoLLaMA 3: Frontier Multimodal Foundation Models for Image and Video Understanding.

Vad-R1: Towards Video Anomaly Reasoning via Perception-to-Cognition Chain-of-Thought VideoLLaMA 3: Frontier Multimodal Foundation Models for Image and Video Understanding

Reference 82

Resolution
unresolved
no resolver link, observed 2026-08-07T14:09:15.409059Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:09:15.409059Z digest=sha256:39584882c3710fefd5ea0bc2879c910aa82db29fae1049d39dfb4774687d6ff0

Observation 2eb827d8-4f25-4651-ad6f-4a6795047fed · outbound

This paper cites Video-LLaMA: An Instruction-tuned Audio-Visual Language Model for Video Understanding.

Vad-R1: Towards Video Anomaly Reasoning via Perception-to-Cognition Chain-of-Thought Video-LLaMA: An Instruction-tuned Audio-Visual Language Model for Video Understanding

Reference 83

Resolution
unresolved
no resolver link, observed 2026-08-07T14:09:15.530406Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:09:15.530406Z digest=sha256:fbfc8d2d1d2effdd85c290ed2a603f6c3b996bd14cbc70ab452aad2216029ab5

Observation 728085c7-9f30-4485-b642-9e33b7928028 · outbound

This paper cites Holmes-VAD: Towards Unbiased and Explainable Video Anomaly Detection via Multi-modal LLM.

Vad-R1: Towards Video Anomaly Reasoning via Perception-to-Cognition Chain-of-Thought Holmes-VAD: Towards Unbiased and Explainable Video Anomaly Detection via Multi-modal LLM

Reference 84

Resolution
unresolved
no resolver link, observed 2026-08-07T14:09:15.602472Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:09:15.602472Z digest=sha256:4b25e1a847757b07fa093c297df8aee241c270ba063b08742bc959108673a6cc

Observation af67379d-8252-48e8-9a0a-ce2e08eb9104 · outbound

This paper cites Holmes-VAU: Towards Long-term Video Anomaly Understanding at Any Granularity.

Vad-R1: Towards Video Anomaly Reasoning via Perception-to-Cognition Chain-of-Thought Holmes-VAU: Towards Long-term Video Anomaly Understanding at Any Granularity

Reference 85

Resolution
unresolved
no resolver link, observed 2026-08-07T14:09:15.659121Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:09:15.659121Z digest=sha256:6da4d51b3ff8f5494a83d9f31c037f3158888d9f0c1be702221cfb4ed64c5eec

Observation 094c0bcc-ac62-41ab-9ccc-b719cc5beb56 · outbound

This paper cites Long Context Transfer from Language to Vision.

Vad-R1: Towards Video Anomaly Reasoning via Perception-to-Cognition Chain-of-Thought Long Context Transfer from Language to Vision

Reference 86

Resolution
unresolved
no resolver link, observed 2026-08-07T14:09:15.746320Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:09:15.746320Z digest=sha256:91f1c427445535bc3c2aded029379ff7f9ae4c212d71a0f471c45077cfb58960

Observation bd3e1bb2-1c08-43a0-95df-5a76ad9fb136 · outbound

This paper cites LLaMA-Adapter: Efficient Fine-tuning of Language Models with Zero-init Attention.

Vad-R1: Towards Video Anomaly Reasoning via Perception-to-Cognition Chain-of-Thought LLaMA-Adapter: Efficient Fine-tuning of Language Models with Zero-init Attention

Reference 87

Resolution
unresolved
no resolver link, observed 2026-08-07T14:09:15.810537Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:09:15.810537Z digest=sha256:cb04addc4961b29dc957088da72e7225d2ea0d717813b3650a8661f8efb9caff

Observation 6c710c76-cef8-41fb-ad94-db1d04df5e4e · outbound

This paper cites TinyLLaVA-Video-R1: Towards Smaller LMMs for Video Reasoning.

Vad-R1: Towards Video Anomaly Reasoning via Perception-to-Cognition Chain-of-Thought TinyLLaVA-Video-R1: Towards Smaller LMMs for Video Reasoning

Reference 88

Resolution
unresolved
no resolver link, observed 2026-08-07T14:09:15.937040Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:09:15.937040Z digest=sha256:aa8af28e2580ce355d748d8b51e21b8d189af6e2acbe01d6fce52cd4376ad6ca

Observation 98a25c67-145d-45dc-be12-5c0ffaeba027 · outbound

This paper cites LLaVA-Video: Video Instruction Tuning With Synthetic Data.

Vad-R1: Towards Video Anomaly Reasoning via Perception-to-Cognition Chain-of-Thought LLaVA-Video: Video Instruction Tuning With Synthetic Data

Reference 89

Resolution
unresolved
no resolver link, observed 2026-08-07T14:09:16.030499Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:09:16.030499Z digest=sha256:435ffc593bbdc72e7939bbfaab4e829d415cb3932a1703811bfa848ae139ea33

Observation c33dd798-8d2a-44b6-a9c1-eb45f49350fb · outbound

This paper cites Graph convo- lutional label noise cleaner: Train a plug-and-play action classifier for anomaly detection.

Vad-R1: Towards Video Anomaly Reasoning via Perception-to-Cognition Chain-of-Thought Graph convo- lutional label noise cleaner: Train a plug-and-play action classifier for anomaly detection

Reference 90

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:09:18.081186Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T14:09:16.151891Z digest=sha256:88c2c0fd5b393dbb74a7b46aa99e54a87fdf867372d4856d15703bc055a6b1a5

Observation 90acc084-8ab1-42a2-a060-599dc0223d2b · outbound

This paper cites Dual memory units with uncertainty regulation for weakly supervised video anomaly detection.

Vad-R1: Towards Video Anomaly Reasoning via Perception-to-Cognition Chain-of-Thought Dual memory units with uncertainty regulation for weakly supervised video anomaly detection

Reference 91

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:09:17.901725Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T14:09:16.297176Z digest=sha256:eca102ca0ac3464a0200cf4adb90fd7b602f4b3036a792dbdcebff34003f6d84

Observation f059adeb-1584-42cc-8b4d-52f63e67b90f · outbound

This paper cites InternVL3: Exploring Advanced Training and Test-Time Recipes for Open-Source Multimodal Models.

Vad-R1: Towards Video Anomaly Reasoning via Perception-to-Cognition Chain-of-Thought InternVL3: Exploring Advanced Training and Test-Time Recipes for Open-Source Multimodal Models

Reference 92

Resolution
malformed identifier
no resolver link, observed 2026-08-07T14:09:16.421330Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:09:16.421330Z digest=sha256:614966e370e3af5b32b8a3bb9e261838ac259e19ee0586550de3634966c557e3

Observation 4b297de6-0e45-480e-b65c-2feff6bd5c22 · outbound

This paper cites an unresolved cited work.

Vad-R1: Towards Video Anomaly Reasoning via Perception-to-Cognition Chain-of-Thought Unresolved cited work

Reference 93

Resolution
unresolved
raw_fallback, observed 2026-08-07T14:09:17.593816Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T14:09:16.580056Z digest=sha256:64f7f7c3522e3f1c2b8cf93677723c618bc8916d64a3d9299fe8a77e3eaf6577

Observation c275f453-f17a-4f16-a048-a78a2af03a71 · outbound

This paper cites Abnormal\.

Vad-R1: Towards Video Anomaly Reasoning via Perception-to-Cognition Chain-of-Thought Abnormal\

Reference 94

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:09:17.440481Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T14:09:16.669323Z digest=sha256:6f8dfab08c78e3e0089da68b7ef5d0a8e4cbf17afba5bb41ec68e2a7b1a0f489

Pith citing papers

Observation 6857c0f6-5f78-4966-9f29-2684b9ee4c1b · inbound

DAMS:Dual-Branch Adaptive Multiscale Spatiotemporal Framework for Video Anomaly Detection cites this paper.

DAMS:Dual-Branch Adaptive Multiscale Spatiotemporal Framework for Video Anomaly Detection Vad-R1: Towards Video Anomaly Reasoning via Perception-to-Cognition Chain-of-Thought

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-06T13:30:09.804400Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T13:30:09.804400Z digest=sha256:953c273ef8fb94b594bf461d92a2afa2f8c0216749623c3edc8109ea1e93cd3b

Observation 2d58dce4-7863-4b4e-91da-9b45bf647994 · inbound

ESOM: Efficiently Understanding Streaming Video Anomalies with Open-world Dynamic Definitions cites this paper.

ESOM: Efficiently Understanding Streaming Video Anomalies with Open-world Dynamic Definitions Vad-R1: Towards Video Anomaly Reasoning via Perception-to-Cognition Chain-of-Thought

Reference 15

Resolution
verified exact
arxiv_id, observed 2026-05-11T05:51:03.085000Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-10T17:55:32.945721Z digest=sha256:1305203f23e08c858303d8759da012ff7f93de7fd56a39685110d795033a3445

Observation ba04cbef-78eb-49da-b50f-06c5f2bf4820 · inbound

MAVEN: A Multi-stage Agentic Annotation Pipeline for Video Reasoning Tasks cites this paper.

MAVEN: A Multi-stage Agentic Annotation Pipeline for Video Reasoning Tasks Vad-R1: Towards Video Anomaly Reasoning via Perception-to-Cognition Chain-of-Thought

Reference 13

Resolution
verified exact
arxiv_id, observed 2026-05-22T07:36:13.970967Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-22T07:36:09.562362Z digest=sha256:c8a6471c3a72b840b49392ee2bc9060657f1051d35eebd237f583cf5a80de364

Observation 7847a13a-472b-4470-a74c-608e47558be7 · inbound

MAVEN: A Multi-stage Agentic Annotation Pipeline for Video Reasoning Tasks cites this paper.

MAVEN: A Multi-stage Agentic Annotation Pipeline for Video Reasoning Tasks Vad-R1: Towards Video Anomaly Reasoning via Perception-to-Cognition Chain-of-Thought

Reference 13

Resolution
unresolved
no resolver link, observed 2026-07-12T16:17:21.074502Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T16:17:21.074502Z digest=sha256:7cf2dde8f5d1445a0603b02010155e235ced5f12c032d390014af9b60484750f

Observation 615540cb-879f-45fc-b9cc-2eaa3b41bae6 · inbound

O-VAD: Industrial Video Anomaly Detection through Object-Centric Tracking and Reasoning cites this paper.

O-VAD: Industrial Video Anomaly Detection through Object-Centric Tracking and Reasoning Vad-R1: Towards Video Anomaly Reasoning via Perception-to-Cognition Chain-of-Thought

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-01T15:56:41.756320Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T15:56:41.756320Z digest=sha256:b4cd67d29abd4d24ba41fe60c51f5635497b09b4249e8896f25d8af894b67493