Pith. sign in

Paper Citation Record · LEDGER

Vad-R1: Towards Video Anomaly Reasoning via Perception-to-Cognition Chain-of-Thought

As of 18 August 2026, this Paper Citation Record lists 94 of 94 outbound references and 6 inbound Pith citation observations for arXiv:2505.19877.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2505.19877 v1

Coverage vector

measured 94 of 94 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T14:09:16.669323Z

measured 100 of 100 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-17T06:30:58.91139+00:00

measured 6 of 6 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-14T04:15:47.172052Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-22T07:36:13.968526Z

Reference resolution

94 of 94 outbound references displayed

  • verified exact0
  • verified fuzzy40
  • unresolved53
  • parse uncertain0
  • malformed identifier1
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 74b36822-cd20-42bb-ad5f-b0b31c146b13 · outbound

This paper cites Ubnormal: New benchmark for supervised open-set video anomaly detection.

Vad-R1: Towards Video Anomaly Reasoning via Perception-to-Cognition Chain-of-Thought Ubnormal: New benchmark for supervised open-set video anomaly detection

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-07T14:09:05.474749Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:09:05.474749Z digest=sha256:04684f28ccd667f25810d0a58020fadbe6d84a865b404c09c00f2158dbadf3a6

Observation 0a812610-d509-40fb-af09-728b338fb3e6 · outbound

This paper cites Claude 3.5 haiku, 2024.

Vad-R1: Towards Video Anomaly Reasoning via Perception-to-Cognition Chain-of-Thought Claude 3.5 haiku, 2024

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-07T14:09:05.615261Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:09:05.615261Z digest=sha256:bc5060cef9a653a0ef69db8ee2599a67e84654db161109704b20736e272c06a5

Observation 26d5540a-20ea-4024-88c9-e95a8ebe8c35 · outbound

This paper cites Meteor: An automatic metric for mt evaluation with improved correlation with human judgments.

Vad-R1: Towards Video Anomaly Reasoning via Perception-to-Cognition Chain-of-Thought Meteor: An automatic metric for mt evaluation with improved correlation with human judgments

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-07T14:09:05.755439Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:09:05.755439Z digest=sha256:1a04b24c91e014e6b7dfcb488c7357711bb144451f64587ac46a3f99069e8207

Observation a97f2be7-6769-4435-b10e-77d3507e57ca · outbound

This paper cites Video generation models as world simulators.

Vad-R1: Towards Video Anomaly Reasoning via Perception-to-Cognition Chain-of-Thought Video generation models as world simulators

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-07T14:09:05.848111Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:09:05.848111Z digest=sha256:401bc34872a2c54784ef1b9a4aa469dfdbbce1715aaf41d6434a9002b3e298b3

Observation 007a6b95-d44e-496d-bac1-f5ee3bd80d15 · outbound

This paper cites A new comprehensive benchmark for semi-supervised video anomaly detection and anticipation.

Vad-R1: Towards Video Anomaly Reasoning via Perception-to-Cognition Chain-of-Thought A new comprehensive benchmark for semi-supervised video anomaly detection and anticipation

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-07T14:09:05.965114Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:09:05.965114Z digest=sha256:c9bfedf395b9cd6bc0f286551ea974ae1b6dc6e06f0bb588f53a6bab78c0707b

Observation 75f469a1-55fb-4e12-bfdd-de9b17badfb2 · outbound

This paper cites Videollm-online: Online video large language model for streaming video.

Vad-R1: Towards Video Anomaly Reasoning via Perception-to-Cognition Chain-of-Thought Videollm-online: Online video large language model for streaming video

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-07T14:09:06.064765Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:09:06.064765Z digest=sha256:82c6aeb528cf512c022853aad04ebda7de26d0f01261521db9c7a9c2828deb0c

Observation 4c726091-ed65-48b8-bb34-4b095c0e1dfe · outbound

This paper cites Prompt-enhanced multiple instance learning for weakly supervised video anomaly detection.

Vad-R1: Towards Video Anomaly Reasoning via Perception-to-Cognition Chain-of-Thought Prompt-enhanced multiple instance learning for weakly supervised video anomaly detection

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-07T14:09:06.185678Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:09:06.185678Z digest=sha256:e27694393aa98a8c0e7d132f42d3021ed16c9e879b7003f6b5ddd289eacae642

Observation e1e0d1f6-85ae-4ca4-9fe3-18bed1997561 · outbound

This paper cites Tevad: Improved video anomaly detection with captions.

Vad-R1: Towards Video Anomaly Reasoning via Perception-to-Cognition Chain-of-Thought Tevad: Improved video anomaly detection with captions

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-07T14:09:06.335558Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:09:06.335558Z digest=sha256:97003b825efdf27c0eb0a7fc88e730addfafc68237dfe6a0fa9c413eb1fc651c

Observation 2bea092f-ab99-4ae8-8334-e2ad17b0e165 · outbound

This paper cites DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning.

Vad-R1: Towards Video Anomaly Reasoning via Perception-to-Cognition Chain-of-Thought DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-07T14:09:06.452918Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:09:06.452918Z digest=sha256:30c58d5848204258765566f52f788df81c5eca49bdb12db2e7dfc9342de226e7

Observation 6fd8b443-55a2-4713-a11c-1352c1a8819a · outbound

This paper cites Streaming Video Question-Answering with In-context Video KV-Cache Retrieval.

Vad-R1: Towards Video Anomaly Reasoning via Perception-to-Cognition Chain-of-Thought Streaming Video Question-Answering with In-context Video KV-Cache Retrieval

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-07T14:09:06.602730Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:09:06.602730Z digest=sha256:c5a0018a523fcdd2ae54628268296b96c7ea8670ad0386e9297d5768c8399727

Observation 6f6e1000-48b4-49d2-9084-00d401747976 · outbound

This paper cites SlowFastVAD: Video Anomaly Detection via Integrating Simple Detector and RAG-Enhanced Vision-Language Model.

Vad-R1: Towards Video Anomaly Reasoning via Perception-to-Cognition Chain-of-Thought SlowFastVAD: Video Anomaly Detection via Integrating Simple Detector and RAG-Enhanced Vision-Language Model

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-07T14:09:06.744774Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:09:06.744774Z digest=sha256:8e2e24f688f05ef7cbca5b01f0a97095ef1084fa72cc7ece238d36a6ee0274ad

Observation d9af6d20-f371-4888-a015-7a7e9556afd8 · outbound

This paper cites Exploring What Why and How: A Multifaceted Benchmark for Causation Understanding of Video Anomaly.

Vad-R1: Towards Video Anomaly Reasoning via Perception-to-Cognition Chain-of-Thought Exploring What Why and How: A Multifaceted Benchmark for Causation Understanding of Video Anomaly

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-07T14:09:06.865811Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:09:06.865811Z digest=sha256:8c5a05693fa6131816b5d152a2fbc74c96acabad4190ce99137cbb5ae5f25463

Observation ae150774-d5e4-47b9-ad27-e5ab8f701bcb · outbound

This paper cites Uncovering what why and how: A comprehensive benchmark for causation understanding of video anomaly.

Vad-R1: Towards Video Anomaly Reasoning via Perception-to-Cognition Chain-of-Thought Uncovering what why and how: A comprehensive benchmark for causation understanding of video anomaly

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-07T14:09:06.941062Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:09:06.941062Z digest=sha256:00625028947bd36d018906a00be68819a1b6f8b343581a7978a623f830e19dc5

Observation c756657f-8bd2-4e55-a4e4-61d6835b3b62 · outbound

This paper cites Video-R1: Reinforcing Video Reasoning in MLLMs.

Vad-R1: Towards Video Anomaly Reasoning via Perception-to-Cognition Chain-of-Thought Video-R1: Reinforcing Video Reasoning in MLLMs

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-07T14:09:07.074909Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:09:07.074909Z digest=sha256:deba361b3c3f202a93a6d868778249a9a55e502a652a1412e50fa7d654420a34

Observation a502e150-fc80-453c-bf54-11fb568ebb92 · outbound

This paper cites Vane-bench: Video anomaly evaluation benchmark for conversational lmms.

Vad-R1: Towards Video Anomaly Reasoning via Perception-to-Cognition Chain-of-Thought Vane-bench: Video anomaly evaluation benchmark for conversational lmms

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-07T14:09:07.204754Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:09:07.204754Z digest=sha256:8140748ffcfdd721f454ff9e3915b3acd2c13c0b90cc9ed43bb642e862b9f872

Observation 488b2ac5-11a5-4341-8c92-923f5ce21e6e · outbound

This paper cites Open-sora: Democratizing efficient video production for all.

Vad-R1: Towards Video Anomaly Reasoning via Perception-to-Cognition Chain-of-Thought Open-sora: Democratizing efficient video production for all

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-07T14:09:07.324835Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:09:07.324835Z digest=sha256:e505e77a8819dac73b803f1cacf8dbc9137808fb5fdd1e79dacc75d257270877

Observation 0c1d6677-8766-4f32-9967-07573de7014b · outbound

This paper cites Abnormal event detection using deep contrastive learning for intelligent video surveillance system.

Vad-R1: Towards Video Anomaly Reasoning via Perception-to-Cognition Chain-of-Thought Abnormal event detection using deep contrastive learning for intelligent video surveillance system

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-07T14:09:07.457830Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:09:07.457830Z digest=sha256:8add5a787c7bc200066dac06673c45b784585b5f8cd24b607b8dacad8c37e319

Observation a6fdc99f-7990-47d5-9de8-e07f8c1467a5 · outbound

This paper cites Weakly supervised video anomaly detection via self-guided temporal discriminative transformer.

Vad-R1: Towards Video Anomaly Reasoning via Perception-to-Cognition Chain-of-Thought Weakly supervised video anomaly detection via self-guided temporal discriminative transformer

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-07T14:09:07.584828Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:09:07.584828Z digest=sha256:97e839dbad4b0ecadc7291fa9e2d2fe5432252b14d4dcd380e6d33ae292a6fe6

Observation e92d8aac-4946-45a9-a887-1d6e19654b5b · outbound

This paper cites Self-supervised attentive generative adversarial networks for video anomaly detection.

Vad-R1: Towards Video Anomaly Reasoning via Perception-to-Cognition Chain-of-Thought Self-supervised attentive generative adversarial networks for video anomaly detection

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-07T14:09:07.694327Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:09:07.694327Z digest=sha256:8e99bebcfa15fa49bdaf2875dc889bacb63dcd4eccb53c5ae7f2c424ef8e9be6

Observation 55e4a19f-f14a-4cd2-8baf-16877e66bd40 · outbound

This paper cites Long short-term dynamic prototype alignment learning for video anomaly detection.

Vad-R1: Towards Video Anomaly Reasoning via Perception-to-Cognition Chain-of-Thought Long short-term dynamic prototype alignment learning for video anomaly detection

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:09:26.839108Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T14:09:07.844750Z digest=sha256:30be0ca4392c12411f3b7b520df07a305f3a828098b4825913802087ce412a83

Observation df11acfd-9dd1-408b-8acd-3f5fef27bc4d · outbound

This paper cites Multi- modal evidential learning for open-world weakly-supervised video anomaly detection.

Vad-R1: Towards Video Anomaly Reasoning via Perception-to-Cognition Chain-of-Thought Multi- modal evidential learning for open-world weakly-supervised video anomaly detection

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:09:26.552206Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T14:09:07.974751Z digest=sha256:46ada9de55b497afcf97664fe370b13ed60c38860cf0f1957d3b1dae6f1e1664

Observation 5693514b-b470-4e39-8cc5-c8eb8e252e09 · outbound

This paper cites Vision-R1: Incentivizing Reasoning Capability in Multimodal Large Language Models.

Vad-R1: Towards Video Anomaly Reasoning via Perception-to-Cognition Chain-of-Thought Vision-R1: Incentivizing Reasoning Capability in Multimodal Large Language Models

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-07T14:09:08.062129Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:09:08.062129Z digest=sha256:20eeb2d62e62aec026848874f81689f9292967ec4d9b60765c30996a869b1001

Observation ea05da91-16eb-45dc-9691-113ca936dd84 · outbound

This paper cites Chat-univi: Unified visual representation empowers large language models with image and video understanding.

Vad-R1: Towards Video Anomaly Reasoning via Perception-to-Cognition Chain-of-Thought Chat-univi: Unified visual representation empowers large language models with image and video understanding

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:09:26.217216Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T14:09:08.224753Z digest=sha256:92342d395bc23645c3d1a3570c65ef2d34c4161567baccd50fef75cc5cecf1d3

Observation bd4c24f3-0769-4fee-a6d4-4e3fc60cebda · outbound

This paper cites Clip-tsa: Clip-assisted temporal self-attention for weakly-supervised video anomaly detection.

Vad-R1: Towards Video Anomaly Reasoning via Perception-to-Cognition Chain-of-Thought Clip-tsa: Clip-assisted temporal self-attention for weakly-supervised video anomaly detection

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:09:25.929611Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T14:09:08.484751Z digest=sha256:c3f003233adf9d6f9a7c2fd56747ec5a0fc33910b877eb55f428e70dc07cf02b

Observation 055477c3-d512-49c2-84aa-4834b945b15b · outbound

This paper cites VideoChat: Chat-Centric Video Understanding.

Vad-R1: Towards Video Anomaly Reasoning via Perception-to-Cognition Chain-of-Thought VideoChat: Chat-Centric Video Understanding

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-07T14:09:08.616938Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:09:08.616938Z digest=sha256:56d8c1e3d255369b68a8aa96d1a9db00da34f18765e42a287109c7f48aaac1bf

Observation 43d8a0f8-915c-4c81-9e1c-28d63e786563 · outbound

This paper cites Anomaly detection and localization in crowded scenes.

Vad-R1: Towards Video Anomaly Reasoning via Perception-to-Cognition Chain-of-Thought Anomaly detection and localization in crowded scenes

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:09:25.694757Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T14:09:08.795213Z digest=sha256:80f56449f27a0f1a4f50ef31feafcf688fb6b6e0de5d10096c3108c0ed8f890c

Observation 7cbb5882-cc3d-430f-b987-5b744e65f3ef · outbound

This paper cites VideoChat-Flash: Hierarchical Compression for Long-Context Video Modeling.

Vad-R1: Towards Video Anomaly Reasoning via Perception-to-Cognition Chain-of-Thought VideoChat-Flash: Hierarchical Compression for Long-Context Video Modeling

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-07T14:09:08.944838Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:09:08.944838Z digest=sha256:be7d684deb0767265ab91949040cf12956defb03fba502f495a840540e9db248

Observation 553478bb-aace-4b35-baf5-b767d1820f7c · outbound

This paper cites VideoChat-R1: Enhancing Spatio-Temporal Perception via Reinforcement Fine-Tuning.

Vad-R1: Towards Video Anomaly Reasoning via Perception-to-Cognition Chain-of-Thought VideoChat-R1: Enhancing Spatio-Temporal Perception via Reinforcement Fine-Tuning

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-07T14:09:09.154751Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:09:09.154751Z digest=sha256:9ce4aae58644f9a9d7b8016bff3ac9cdbdcb054ffee7d6dc4fb5ff4c62db8e6d

Observation 2d68afd5-ad4a-4810-ad0e-9ef78d163385 · outbound

This paper cites Llama-vid: An image is worth 2 tokens in large language models.

Vad-R1: Towards Video Anomaly Reasoning via Perception-to-Cognition Chain-of-Thought Llama-vid: An image is worth 2 tokens in large language models

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:09:25.440900Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T14:09:09.259223Z digest=sha256:5a0031a5c5fa2796efb69a09cffccfe4129a870300a78e9ab59042b116166545

Observation 691657a4-6481-4c8d-8176-03358e97511b · outbound

This paper cites Video-LLaVA: Learning United Visual Representation by Alignment Before Projection.

Vad-R1: Towards Video Anomaly Reasoning via Perception-to-Cognition Chain-of-Thought Video-LLaVA: Learning United Visual Representation by Alignment Before Projection

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-07T14:09:09.355173Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:09:09.355173Z digest=sha256:29a2f97d974dd450392ad63596be05981c4cedf0e9bef26e5a50efb8eb179f45

Observation 6d5d9df4-6d10-4750-9f61-435e7b11e8c2 · outbound

This paper cites Rouge: A package for automatic evaluation of summaries.

Vad-R1: Towards Video Anomaly Reasoning via Perception-to-Cognition Chain-of-Thought Rouge: A package for automatic evaluation of summaries

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-07T14:09:09.420025Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:09:09.420025Z digest=sha256:0093a7cea22fa010113a02f8b735a36cd84951233a7804a2461304cdd78b2e19

Observation 4ebeb344-590d-49e2-9024-331040fdedd9 · outbound

This paper cites Future frame prediction for anomaly detection–a new baseline.

Vad-R1: Towards Video Anomaly Reasoning via Perception-to-Cognition Chain-of-Thought Future frame prediction for anomaly detection–a new baseline

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:09:25.114118Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T14:09:09.505566Z digest=sha256:ed2aea4244328838ae099a4f0fe861a83f1c5654260743f3c5198f258cc469a4

Observation 05a3f08a-54d7-4589-8323-56bb18885531 · outbound

This paper cites Videomind: A chain-of-lora agent for long video reasoning.

Vad-R1: Towards Video Anomaly Reasoning via Perception-to-Cognition Chain-of-Thought Videomind: A chain-of-lora agent for long video reasoning

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-07T14:09:09.564739Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:09:09.564739Z digest=sha256:1d1d7f055c7ed2eadf17f63a9f430f10ae39dbe3abd2d02fb32b498191200739

Observation 7c41ad22-a1dd-4d6a-9c10-72f8d536cd85 · outbound

This paper cites A hybrid video anomaly detection framework via memory-augmented flow reconstruction and flow-guided frame prediction.

Vad-R1: Towards Video Anomaly Reasoning via Perception-to-Cognition Chain-of-Thought A hybrid video anomaly detection framework via memory-augmented flow reconstruction and flow-guided frame prediction

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:09:24.904990Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T14:09:09.627328Z digest=sha256:447d1e2cabfcba7d1fe8cef050b8404ce8ed9a2e3af70fe23b3e9e23d9165108

Observation fa7210e4-9495-44ff-8a9e-c199bdd8a1f7 · outbound

This paper cites Abnormal event detection at 150 fps in matlab.

Vad-R1: Towards Video Anomaly Reasoning via Perception-to-Cognition Chain-of-Thought Abnormal event detection at 150 fps in matlab

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:09:24.666630Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T14:09:09.713530Z digest=sha256:c1ae4aee2e3ebe13b5687607caf75d3af04426c3ace3833f171b2a62e2d7b532

Observation e63d5cbe-9a66-4311-9a70-570cae6a6718 · outbound

This paper cites Video Anomaly Detection and Explanation via Large Language Models.

Vad-R1: Towards Video Anomaly Reasoning via Perception-to-Cognition Chain-of-Thought Video Anomaly Detection and Explanation via Large Language Models

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-07T14:09:09.799641Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:09:09.799641Z digest=sha256:e795f8f4c0cd2b2a98ebbab24987103fae02154733a6a9111711b7fe6ef6f9e3

Observation b1165a0b-ab5b-460b-954b-803afd48c95a · outbound

This paper cites Localizing anomalies from weakly-labeled videos.

Vad-R1: Towards Video Anomaly Reasoning via Perception-to-Cognition Chain-of-Thought Localizing anomalies from weakly-labeled videos

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:09:24.434697Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T14:09:09.912993Z digest=sha256:17a28dda77547f1a1121b691a27caf118ead5e7c4c15625681b6efacd3d6462f

Observation b1f53c55-f8f8-4cc3-8378-0ffb6a2ddee9 · outbound

This paper cites Sherlock: Towards multi-scene video abnormal event extraction and localization via a global-local spatial-sensitive llm.

Vad-R1: Towards Video Anomaly Reasoning via Perception-to-Cognition Chain-of-Thought Sherlock: Towards multi-scene video abnormal event extraction and localization via a global-local spatial-sensitive llm

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:09:24.199170Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T14:09:10.058849Z digest=sha256:93ea3b211826b6d227b18cc2f17b086cb2641cace897d10f86a32167da4752a6

Observation b176954c-530a-42a8-bfc5-798933f5caae · outbound

This paper cites Video-ChatGPT: Towards Detailed Video Understanding via Large Vision and Language Models.

Vad-R1: Towards Video Anomaly Reasoning via Perception-to-Cognition Chain-of-Thought Video-ChatGPT: Towards Detailed Video Understanding via Large Vision and Language Models

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-07T14:09:10.196311Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:09:10.196311Z digest=sha256:a7857096a1727c6c2e6ec19a042982b22968a52f64254af64549f6c750186ec5

Observation 07fa7e2a-920e-4c79-92d3-5dcd7075e8a5 · outbound

This paper cites GPT-4o System Card.

Vad-R1: Towards Video Anomaly Reasoning via Perception-to-Cognition Chain-of-Thought GPT-4o System Card

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-07T14:09:10.302054Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:09:10.302054Z digest=sha256:c799f3ba5b21fe66663ad4e445ccd4925d9c4b8e36293be07854908f4f63c86c

Observation a5b0c482-f59f-4766-b458-7d87296821ea · outbound

This paper cites OpenAI o1 System Card.

Vad-R1: Towards Video Anomaly Reasoning via Perception-to-Cognition Chain-of-Thought OpenAI o1 System Card

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-07T14:09:10.444749Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:09:10.444749Z digest=sha256:1e39f5c4d8cd44664726c66a065d3cfd13524c1814c4ab1c57af2bb6533e5a5f

Observation ff741184-1820-4969-8740-24d993a00f77 · outbound

This paper cites Openai o3 and o4-mini system card, 2025.

Vad-R1: Towards Video Anomaly Reasoning via Perception-to-Cognition Chain-of-Thought Openai o3 and o4-mini system card, 2025

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:09:23.939270Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T14:09:10.505981Z digest=sha256:33678b49a36f85b8c95b375ee3278bcb14c543aa1cb5545a0b171f35935a9d59

Observation c6805c43-c11e-466e-8f1b-c68ee9a2e2c8 · outbound

This paper cites Bleu: a method for automatic evaluation of machine translation.

Vad-R1: Towards Video Anomaly Reasoning via Perception-to-Cognition Chain-of-Thought Bleu: a method for automatic evaluation of machine translation

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:09:23.714214Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T14:09:10.667970Z digest=sha256:d70464ad3057ef8d5fa705754797b90431d6aca9dd9a35f2d16896c5811f3518

Observation 97381c00-2406-43a0-a95b-e3928a48e55d · outbound

This paper cites Timechat: A time-sensitive multimodal large language model for long video understanding.

Vad-R1: Towards Video Anomaly Reasoning via Perception-to-Cognition Chain-of-Thought Timechat: A time-sensitive multimodal large language model for long video understanding

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:09:23.495945Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T14:09:10.756335Z digest=sha256:0bc85b593cb2625e7ed4bad95c2f9fca3b0129f44dafdb59e5bae0b138976490

Observation 5f14ff13-2a79-46b5-bef6-5d56de4466cf · outbound

This paper cites Self-distilled masked auto-encoders are efficient video anomaly detectors.

Vad-R1: Towards Video Anomaly Reasoning via Perception-to-Cognition Chain-of-Thought Self-distilled masked auto-encoders are efficient video anomaly detectors

Reference 45

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:09:23.327144Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T14:09:10.887212Z digest=sha256:f0d237c0248a5f8112ce604edbb41bbd6894287a43766775a2cd53cea0a7020f

Observation ff2d1d07-f497-4710-a4fb-0c452f49e8c7 · outbound

This paper cites Gen-2: The next step forward for generative ai.

Vad-R1: Towards Video Anomaly Reasoning via Perception-to-Cognition Chain-of-Thought Gen-2: The next step forward for generative ai

Reference 46

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:09:23.141205Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T14:09:11.035451Z digest=sha256:bbe956f26dd5c13c00c5fe0f482629274352ab2e8dd06f5689e9d396bbb3ae1b

Observation e02fa9d3-ca25-41f5-a676-0994b4c8d26d · outbound

This paper cites DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models.

Vad-R1: Towards Video Anomaly Reasoning via Perception-to-Cognition Chain-of-Thought DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-07T14:09:11.127199Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:09:11.127199Z digest=sha256:04dc2f91e3e06e5f96db311783d4da0abc66c6929ffeb835b734a23ef9e3ebe8

Observation 2614eb40-de40-4d40-b6e9-943ee598bd43 · outbound

This paper cites Moviechat: From dense token to sparse memory for long video understanding.

Vad-R1: Towards Video Anomaly Reasoning via Perception-to-Cognition Chain-of-Thought Moviechat: From dense token to sparse memory for long video understanding

Reference 48

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:09:22.897638Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T14:09:11.208115Z digest=sha256:ad808088f21145a56b1ebb4b5b0a88cf52083d78254797a63c384a16faf7f48d

Observation afb1d732-1205-4164-9414-c057c4dbcb29 · outbound

This paper cites Real-world anomaly detection in surveillance videos.

Vad-R1: Towards Video Anomaly Reasoning via Perception-to-Cognition Chain-of-Thought Real-world anomaly detection in surveillance videos

Reference 49

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:09:22.698560Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T14:09:11.296667Z digest=sha256:80954a2341c2f6990bea6f256d84265b41ad52bd050e66862e7d31615d08044e

Observation 20eda1fa-7d2a-4aa5-b4cd-1d375a38df29 · outbound

This paper cites Hawk: Learning to understand open-world video anomalies.

Vad-R1: Towards Video Anomaly Reasoning via Perception-to-Cognition Chain-of-Thought Hawk: Learning to understand open-world video anomalies

Reference 50

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:09:22.477709Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T14:09:11.427614Z digest=sha256:ceeb862a420699c50f597db47c4b238e97f44a119b2ff7ce545f1006954197b7

Observation 6f3f12c2-58de-4482-8856-24dfd35cde15 · outbound

This paper cites Gemini 2.5 flash preview model card, 2025.

Vad-R1: Towards Video Anomaly Reasoning via Perception-to-Cognition Chain-of-Thought Gemini 2.5 flash preview model card, 2025

Reference 51

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:09:22.248638Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T14:09:11.520760Z digest=sha256:c1ece83843a93b006c54044a817eede2ce8b02154faf21d638e10b13c1efb5df

Observation 2744d918-66fa-439e-aaa5-e79ce0f16cc2 · outbound

This paper cites Gemini 2.5 pro preview model card, 2025.

Vad-R1: Towards Video Anomaly Reasoning via Perception-to-Cognition Chain-of-Thought Gemini 2.5 pro preview model card, 2025

Reference 52

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:09:22.011914Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T14:09:11.656655Z digest=sha256:31867ad15d6e9f6ce46289144f181e5a1eab06e6734ad2e9fbefa4b9ff919291

Observation 8ee43beb-a40e-4c8c-a2b2-37ea29aec736 · outbound

This paper cites Kimi k1.5: Scaling Reinforcement Learning with LLMs.

Vad-R1: Towards Video Anomaly Reasoning via Perception-to-Cognition Chain-of-Thought Kimi k1.5: Scaling Reinforcement Learning with LLMs

Reference 53

Resolution
unresolved
no resolver link, observed 2026-08-07T14:09:11.779386Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:09:11.779386Z digest=sha256:2c8361f0d3f7956831ab7841d4d8708d1e172374bdcf68eb5fb4b2709f2cad60

Observation 07226a3d-d355-4915-b1de-611b370b356f · outbound

This paper cites QwQ: Reflect deeply on the boundaries of the unknown, 2024.

Vad-R1: Towards Video Anomaly Reasoning via Perception-to-Cognition Chain-of-Thought QwQ: Reflect deeply on the boundaries of the unknown, 2024

Reference 54

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:09:21.768204Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T14:09:11.905318Z digest=sha256:4a0ea2b784d03c8344604d4fc4908f697bd21bc675b7b9b55f36a298559b56ac

Observation 1f500267-5cd6-4dd0-8b16-da052f6e8663 · outbound

This paper cites Qwen2.5 Technical Report.

Vad-R1: Towards Video Anomaly Reasoning via Perception-to-Cognition Chain-of-Thought Qwen2.5 Technical Report

Reference 55

Resolution
unresolved
no resolver link, observed 2026-08-07T14:09:12.056747Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:09:12.056747Z digest=sha256:f8dcb5a5bcee36134bd4b0dae37afa3a609573d824d8767cb447381e034ad287

Observation 1ab641df-1237-4bfe-a4bb-f070f355621d · outbound

This paper cites QVQ-Max: Think with evidence, 2025.

Vad-R1: Towards Video Anomaly Reasoning via Perception-to-Cognition Chain-of-Thought QVQ-Max: Think with evidence, 2025

Reference 56

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:09:21.561752Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T14:09:12.193934Z digest=sha256:c9fb34ef7f92519a092bb8d29e4e8db8e74efc62d928f15dc45c35ef63658412

Observation 448e5500-10ea-4142-bb64-1724dfe07742 · outbound

This paper cites Qwen2.5-VL Technical Report.

Vad-R1: Towards Video Anomaly Reasoning via Perception-to-Cognition Chain-of-Thought Qwen2.5-VL Technical Report

Reference 57

Resolution
unresolved
no resolver link, observed 2026-08-07T14:09:12.381462Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:09:12.381462Z digest=sha256:9141028c659d4572798fe97f26ab29cbf3cf584319b492b22525154dceb1c182

Observation df645d7a-95e1-4249-b0ab-2baea805dd5e · outbound

This paper cites Qwen3: Think deeper, act faster, 2025.

Vad-R1: Towards Video Anomaly Reasoning via Perception-to-Cognition Chain-of-Thought Qwen3: Think deeper, act faster, 2025

Reference 58

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:09:21.383574Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T14:09:12.515845Z digest=sha256:0c503e4d7409defaa461ea13098eace91784083cfa3e1b336821add60d7154b6

Observation 28b90eaf-839a-41af-9fbf-d26950833de2 · outbound

This paper cites LlamaV-o1: Rethinking Step-by-step Visual Reasoning in LLMs.

Vad-R1: Towards Video Anomaly Reasoning via Perception-to-Cognition Chain-of-Thought LlamaV-o1: Rethinking Step-by-step Visual Reasoning in LLMs

Reference 59

Resolution
unresolved
no resolver link, observed 2026-08-07T14:09:12.592946Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:09:12.592946Z digest=sha256:17bf81a32a02b89c5e12dd7a366e7ee9b5bd7f548d055b36cd8d88aeecb38b42

Observation dd14da3b-8530-400d-aff9-1662944d3cef · outbound

This paper cites Federated weakly supervised video anomaly detection with multimodal prompt.

Vad-R1: Towards Video Anomaly Reasoning via Perception-to-Cognition Chain-of-Thought Federated weakly supervised video anomaly detection with multimodal prompt

Reference 60

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:09:21.196877Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T14:09:12.737393Z digest=sha256:b70ec1ac1652ea27bafd5bcd7be4fa18b25d145287ad22b511ab43c56be024ca

Observation d5da0cb7-c562-42fe-a2f2-f59dad0ba0f5 · outbound

This paper cites Modelscope text-to-video technical report, 2023.

Vad-R1: Towards Video Anomaly Reasoning via Perception-to-Cognition Chain-of-Thought Modelscope text-to-video technical report, 2023

Reference 61

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:09:20.893590Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T14:09:12.826604Z digest=sha256:826025e371366cfb935f082c5c98e1e3ab72e6ff4883842ba4954a83bfa91e52

Observation 65b6ee68-ab31-4868-ac25-8f5e2446f386 · outbound

This paper cites Videolcm: Video latent consistency model, 2023.

Vad-R1: Towards Video Anomaly Reasoning via Perception-to-Cognition Chain-of-Thought Videolcm: Video latent consistency model, 2023

Reference 62

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:09:20.667580Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T14:09:12.954393Z digest=sha256:62eba03ab11cdbe8190a1353c52ed36d9e7c52f2ca990c12a3daaedde784fd79

Observation 81712c2c-aba2-44c1-bd0a-5c9359dedc18 · outbound

This paper cites Open-r1-video, 2025.

Vad-R1: Towards Video Anomaly Reasoning via Perception-to-Cognition Chain-of-Thought Open-r1-video, 2025

Reference 63

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:09:20.421517Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T14:09:13.063090Z digest=sha256:5800856e9e7699910694e476b2a49cf746c5afa61e1bad64a59bc715f7311666

Observation e45504bb-270e-4351-b01d-cf44a4d26041 · outbound

This paper cites Time-R1: Post-Training Large Vision Language Model for Temporal Video Grounding.

Vad-R1: Towards Video Anomaly Reasoning via Perception-to-Cognition Chain-of-Thought Time-R1: Post-Training Large Vision Language Model for Temporal Video Grounding

Reference 64

Resolution
unresolved
no resolver link, observed 2026-08-07T14:09:13.173718Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:09:13.173718Z digest=sha256:688182e5d60b90fb624ba86c65ccfc2571c7afdaa0d8876112409b79e73438e0

Observation 01a233a0-30b5-42cc-a9b7-5a273bf9102d · outbound

This paper cites InternVideo2.5: Empowering Video MLLMs with Long and Rich Context Modeling.

Vad-R1: Towards Video Anomaly Reasoning via Perception-to-Cognition Chain-of-Thought InternVideo2.5: Empowering Video MLLMs with Long and Rich Context Modeling

Reference 65

Resolution
unresolved
no resolver link, observed 2026-08-07T14:09:13.309276Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:09:13.309276Z digest=sha256:21e3b625bfec548f3e863706ca227240482e4b3557c74a7884e8c0ca8d370918

Observation a4d1e605-c307-44f1-80bf-b179812f6c30 · outbound

This paper cites Not only look, but also listen: Learning multimodal violence detection under weak supervision.

Vad-R1: Towards Video Anomaly Reasoning via Perception-to-Cognition Chain-of-Thought Not only look, but also listen: Learning multimodal violence detection under weak supervision

Reference 66

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:09:20.273475Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T14:09:13.455450Z digest=sha256:a0770ea00472d154241769d166ee09fea4a15d9404770782e45cc27b2692ab46

Observation 7a19cb88-7667-4f19-9b35-15bf95a02bce · outbound

This paper cites Open-vocabulary video anomaly detection.

Vad-R1: Towards Video Anomaly Reasoning via Perception-to-Cognition Chain-of-Thought Open-vocabulary video anomaly detection

Reference 67

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:09:20.056783Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T14:09:13.666971Z digest=sha256:fad3197721b07a0ae8696f0423d69f500a0012c12ce46c358adf06791ca1b84f

Observation c88bff23-6e4c-42ae-919e-f53d82d146fc · outbound

This paper cites Vadclip: Adapting vision-language models for weakly supervised video anomaly detection.

Vad-R1: Towards Video Anomaly Reasoning via Perception-to-Cognition Chain-of-Thought Vadclip: Adapting vision-language models for weakly supervised video anomaly detection

Reference 68

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:09:19.794855Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T14:09:13.834821Z digest=sha256:cef2fd9cbe940001164bf84fd3bd42882cc8304164f6a8a69b5fa7086058cd4c

Observation e57d3da4-cce2-4184-93a0-e8ea7e0e0757 · outbound

This paper cites Streaming Video Understanding and Multi-round Interaction with Memory-enhanced Knowledge.

Vad-R1: Towards Video Anomaly Reasoning via Perception-to-Cognition Chain-of-Thought Streaming Video Understanding and Multi-round Interaction with Memory-enhanced Knowledge

Reference 69

Resolution
unresolved
no resolver link, observed 2026-08-07T14:09:13.898584Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:09:13.898584Z digest=sha256:98d11cc49f0c159d6064451c1adec5e9712a9f08c1450a8c6274c939b4621ab3

Observation 0b2d7b0d-c14e-40c1-afad-4bb99c6a607e · outbound

This paper cites LLaVA-CoT: Let Vision Language Models Reason Step-by-Step.

Vad-R1: Towards Video Anomaly Reasoning via Perception-to-Cognition Chain-of-Thought LLaVA-CoT: Let Vision Language Models Reason Step-by-Step

Reference 70

Resolution
unresolved
no resolver link, observed 2026-08-07T14:09:13.996611Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:09:13.996611Z digest=sha256:fd5998af17d172abcfc5574679c31986efe82f6a1cad31b0eeab3f246ac49fd6

Observation acfe4bdc-31fb-4919-a07b-0e12ef58cd21 · outbound

This paper cites PLLaVA : Parameter-free LLaVA Extension from Images to Videos for Video Dense Captioning.

Vad-R1: Towards Video Anomaly Reasoning via Perception-to-Cognition Chain-of-Thought PLLaVA : Parameter-free LLaVA Extension from Images to Videos for Video Dense Captioning

Reference 71

Resolution
unresolved
no resolver link, observed 2026-08-07T14:09:14.086670Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:09:14.086670Z digest=sha256:53bd2f2ee32f0e45f4424bba1fc3aedc5cb65c538222a884e2455200f0174efa

Observation a43e15e6-d262-46ed-9cc9-c4d206fa7355 · outbound

This paper cites Feature prediction diffusion model for video anomaly detection.

Vad-R1: Towards Video Anomaly Reasoning via Perception-to-Cognition Chain-of-Thought Feature prediction diffusion model for video anomaly detection

Reference 72

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:09:19.515052Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T14:09:14.200572Z digest=sha256:39139d3132d82bd132f369d30f9a237e9726eade376d8d17fc46a1aea9911225

Observation 109c2b18-687d-4567-abbd-242fca3d7672 · outbound

This paper cites Follow the rules: reasoning for video anomaly detection with large language models.

Vad-R1: Towards Video Anomaly Reasoning via Perception-to-Cognition Chain-of-Thought Follow the rules: reasoning for video anomaly detection with large language models

Reference 73

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:09:19.308536Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T14:09:14.315573Z digest=sha256:734d1e75e53a847a71b5bf8c32e4c2867082e1cd07e525d8ed0aea4733130ed4

Observation 7da2edef-5179-49bc-a2d6-39b27bc27b8f · outbound

This paper cites Svbench: A benchmark with temporal multi-turn dialogues for streaming video understanding.

Vad-R1: Towards Video Anomaly Reasoning via Perception-to-Cognition Chain-of-Thought Svbench: A benchmark with temporal multi-turn dialogues for streaming video understanding

Reference 74

Resolution
unresolved
no resolver link, observed 2026-08-07T14:09:14.364645Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:09:14.364645Z digest=sha256:f69e571f605d8a87e350657e67e229910a9016cbdc622f1d8547cf1fdbe96492

Observation 8cd34668-23b6-4bb2-9314-b1f26960f6b0 · outbound

This paper cites Dota: Unsupervised detection of traffic anomaly in driving videos.IEEE transactions on pattern analysis and machine intelligence, 45(1):444–459, 2022.

Vad-R1: Towards Video Anomaly Reasoning via Perception-to-Cognition Chain-of-Thought Dota: Unsupervised detection of traffic anomaly in driving videos.IEEE transactions on pattern analysis and machine intelligence, 45(1):444–459, 2022

Reference 75

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:09:19.059976Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T14:09:14.550526Z digest=sha256:a30d46d2cf2bb8fc36fafd3f102931e5cf721e08b618b6ab701fe4717a9ed0a8

Observation 92a88d12-b3c7-4dae-8ef0-d5de816ef36a · outbound

This paper cites VERA: Explainable Video Anomaly Detection via Verbalized Learning of Vision-Language Models.

Vad-R1: Towards Video Anomaly Reasoning via Perception-to-Cognition Chain-of-Thought VERA: Explainable Video Anomaly Detection via Verbalized Learning of Vision-Language Models

Reference 76

Resolution
unresolved
no resolver link, observed 2026-08-07T14:09:14.656477Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:09:14.656477Z digest=sha256:0540ddea29dea2431b37bd6d6e675b40ffd2750c4a165033d3f560c7d17198df

Observation a22424f5-40a8-4045-8435-a6ee1dab4d81 · outbound

This paper cites Unhackable Temporal Rewarding for Scalable Video MLLMs.

Vad-R1: Towards Video Anomaly Reasoning via Perception-to-Cognition Chain-of-Thought Unhackable Temporal Rewarding for Scalable Video MLLMs

Reference 77

Resolution
unresolved
no resolver link, observed 2026-08-07T14:09:14.787236Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:09:14.787236Z digest=sha256:dbfdfaf69d8c05e726b995cf286ae03b31669cadee7445f9b34b3cc62daa5660

Observation 1f712552-3535-42cd-9da4-ffb909c68cd5 · outbound

This paper cites Towards surveillance video-and-language understanding: New dataset baselines and challenges.

Vad-R1: Towards Video Anomaly Reasoning via Perception-to-Cognition Chain-of-Thought Towards surveillance video-and-language understanding: New dataset baselines and challenges

Reference 78

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:09:18.816692Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T14:09:14.937077Z digest=sha256:3de17d8286170cd84a18b0c07e69f9161d3800cadc211816605cc7683d0735f3

Observation deaa7724-d981-4235-b62b-1accd7085ef8 · outbound

This paper cites Generative cooperative learning for unsupervised video anomaly detection.

Vad-R1: Towards Video Anomaly Reasoning via Perception-to-Cognition Chain-of-Thought Generative cooperative learning for unsupervised video anomaly detection

Reference 79

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:09:18.571228Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T14:09:15.060104Z digest=sha256:23fb6a5854441b147c76c5892251850fbb4318b9c17bd52d6aaa18b2c84329b7

Observation c8ec1c8e-1397-447d-a78f-c2c221af5bab · outbound

This paper cites Har- nessing large language models for training-free video anomaly detection.

Vad-R1: Towards Video Anomaly Reasoning via Perception-to-Cognition Chain-of-Thought Har- nessing large language models for training-free video anomaly detection

Reference 80

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:09:18.409732Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T14:09:15.166896Z digest=sha256:07c1466c550bdbae8964716e3db7cc29066df9a678d7311e30c4bd2d431f44b2

Observation 850ec3cd-0c61-4c3f-8211-822fc8fc9f39 · outbound

This paper cites Vision-R1: Evolving Human-Free Alignment in Large Vision-Language Models via Vision-Guided Reinforcement Learning.

Vad-R1: Towards Video Anomaly Reasoning via Perception-to-Cognition Chain-of-Thought Vision-R1: Evolving Human-Free Alignment in Large Vision-Language Models via Vision-Guided Reinforcement Learning

Reference 81

Resolution
unresolved
no resolver link, observed 2026-08-07T14:09:15.259254Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:09:15.259254Z digest=sha256:4f76a15f856a5067a74a13ec7c373f50ad821f1fdcfb211d28f50bed2cceae8d

Observation 266dcb77-7c01-4542-9de0-21ee260189c1 · outbound

This paper cites VideoLLaMA 3: Frontier Multimodal Foundation Models for Image and Video Understanding.

Vad-R1: Towards Video Anomaly Reasoning via Perception-to-Cognition Chain-of-Thought VideoLLaMA 3: Frontier Multimodal Foundation Models for Image and Video Understanding

Reference 82

Resolution
unresolved
no resolver link, observed 2026-08-07T14:09:15.409059Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:09:15.409059Z digest=sha256:d3f9c7381cf5eec9c4794980503e4e3df926227218cae7140cb6308664e3c1d9

Observation 2eb827d8-4f25-4651-ad6f-4a6795047fed · outbound

This paper cites Video-LLaMA: An Instruction-tuned Audio-Visual Language Model for Video Understanding.

Vad-R1: Towards Video Anomaly Reasoning via Perception-to-Cognition Chain-of-Thought Video-LLaMA: An Instruction-tuned Audio-Visual Language Model for Video Understanding

Reference 83

Resolution
unresolved
no resolver link, observed 2026-08-07T14:09:15.530406Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:09:15.530406Z digest=sha256:07a79a7693d1f39cfefdeb161f85cf69d4527beb28533e88cc3fc97b9c08d5ab

Observation 728085c7-9f30-4485-b642-9e33b7928028 · outbound

This paper cites Holmes-VAD: Towards Unbiased and Explainable Video Anomaly Detection via Multi-modal LLM.

Vad-R1: Towards Video Anomaly Reasoning via Perception-to-Cognition Chain-of-Thought Holmes-VAD: Towards Unbiased and Explainable Video Anomaly Detection via Multi-modal LLM

Reference 84

Resolution
unresolved
no resolver link, observed 2026-08-07T14:09:15.602472Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:09:15.602472Z digest=sha256:adc8f720c6f952aa1a9d5b01649e3986feed537358d681c5feb15e4742db05bf

Observation af67379d-8252-48e8-9a0a-ce2e08eb9104 · outbound

This paper cites Holmes-VAU: Towards Long-term Video Anomaly Understanding at Any Granularity.

Vad-R1: Towards Video Anomaly Reasoning via Perception-to-Cognition Chain-of-Thought Holmes-VAU: Towards Long-term Video Anomaly Understanding at Any Granularity

Reference 85

Resolution
unresolved
no resolver link, observed 2026-08-07T14:09:15.659121Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:09:15.659121Z digest=sha256:c2dcf27a53d905806c32883fd6041becc2ecf984aaa22819ad701ec76a4cb0a2

Observation 094c0bcc-ac62-41ab-9ccc-b719cc5beb56 · outbound

This paper cites Long Context Transfer from Language to Vision.

Vad-R1: Towards Video Anomaly Reasoning via Perception-to-Cognition Chain-of-Thought Long Context Transfer from Language to Vision

Reference 86

Resolution
unresolved
no resolver link, observed 2026-08-07T14:09:15.746320Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:09:15.746320Z digest=sha256:7309bd699e8ded63ca7843328c7e296b4ea80decdf19bcea290146a06130ce82

Observation bd3e1bb2-1c08-43a0-95df-5a76ad9fb136 · outbound

This paper cites LLaMA-Adapter: Efficient Fine-tuning of Language Models with Zero-init Attention.

Vad-R1: Towards Video Anomaly Reasoning via Perception-to-Cognition Chain-of-Thought LLaMA-Adapter: Efficient Fine-tuning of Language Models with Zero-init Attention

Reference 87

Resolution
unresolved
no resolver link, observed 2026-08-07T14:09:15.810537Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:09:15.810537Z digest=sha256:08dd90574fe83e8af0704becfddcaf7dacc80d0e567007d553a244cd1a548f3e

Observation 6c710c76-cef8-41fb-ad94-db1d04df5e4e · outbound

This paper cites TinyLLaVA-Video-R1: Towards Smaller LMMs for Video Reasoning.

Vad-R1: Towards Video Anomaly Reasoning via Perception-to-Cognition Chain-of-Thought TinyLLaVA-Video-R1: Towards Smaller LMMs for Video Reasoning

Reference 88

Resolution
unresolved
no resolver link, observed 2026-08-07T14:09:15.937040Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:09:15.937040Z digest=sha256:aa839ce4b597b4668305e93188ccc4cd63973440ba790ecabd864c7f3e801b0d

Observation 98a25c67-145d-45dc-be12-5c0ffaeba027 · outbound

This paper cites LLaVA-Video: Video Instruction Tuning With Synthetic Data.

Vad-R1: Towards Video Anomaly Reasoning via Perception-to-Cognition Chain-of-Thought LLaVA-Video: Video Instruction Tuning With Synthetic Data

Reference 89

Resolution
unresolved
no resolver link, observed 2026-08-07T14:09:16.030499Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:09:16.030499Z digest=sha256:eae7402a4d46fecb39878842630f7eb03e05e930130a4ab8c473d5a31fb2bd22

Observation c33dd798-8d2a-44b6-a9c1-eb45f49350fb · outbound

This paper cites Graph convo- lutional label noise cleaner: Train a plug-and-play action classifier for anomaly detection.

Vad-R1: Towards Video Anomaly Reasoning via Perception-to-Cognition Chain-of-Thought Graph convo- lutional label noise cleaner: Train a plug-and-play action classifier for anomaly detection

Reference 90

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:09:18.081186Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T14:09:16.151891Z digest=sha256:2ed75144115c82a1c8b9790007518a9a343339d4ef151359ab71d567c355da24

Observation 90acc084-8ab1-42a2-a060-599dc0223d2b · outbound

This paper cites Dual memory units with uncertainty regulation for weakly supervised video anomaly detection.

Vad-R1: Towards Video Anomaly Reasoning via Perception-to-Cognition Chain-of-Thought Dual memory units with uncertainty regulation for weakly supervised video anomaly detection

Reference 91

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:09:17.901725Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T14:09:16.297176Z digest=sha256:374e788cd96b2e6c94b5584e3fa94a4a2377d121b71a70541859e2e843be7040

Observation f059adeb-1584-42cc-8b4d-52f63e67b90f · outbound

This paper cites InternVL3: Exploring Advanced Training and Test-Time Recipes for Open-Source Multimodal Models.

Vad-R1: Towards Video Anomaly Reasoning via Perception-to-Cognition Chain-of-Thought InternVL3: Exploring Advanced Training and Test-Time Recipes for Open-Source Multimodal Models

Reference 92

Resolution
malformed identifier
no resolver link, observed 2026-08-07T14:09:16.421330Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:09:16.421330Z digest=sha256:357c13969641610e6e8022aa9a22c6feef6ae3992feb5478ccfde3b10c7ae435

Observation 4b297de6-0e45-480e-b65c-2feff6bd5c22 · outbound

This paper cites an unresolved cited work.

Vad-R1: Towards Video Anomaly Reasoning via Perception-to-Cognition Chain-of-Thought Unresolved cited work

Reference 93

Resolution
unresolved
raw_fallback, observed 2026-08-07T14:09:17.593816Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T14:09:16.580056Z digest=sha256:c5cac8097f9351470ec93d6d4ae081507c24a996ddb9c27b3dea5f5eb03b7252

Observation c275f453-f17a-4f16-a048-a78a2af03a71 · outbound

This paper cites Abnormal\.

Vad-R1: Towards Video Anomaly Reasoning via Perception-to-Cognition Chain-of-Thought Abnormal\

Reference 94

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:09:17.440481Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T14:09:16.669323Z digest=sha256:6e16f326e3980b5a853582d7a0d55408b6eb070608665b8f7a286e9013d4b58c

Pith citing papers

Observation 6857c0f6-5f78-4966-9f29-2684b9ee4c1b · inbound

DAMS:Dual-Branch Adaptive Multiscale Spatiotemporal Framework for Video Anomaly Detection cites this paper.

DAMS:Dual-Branch Adaptive Multiscale Spatiotemporal Framework for Video Anomaly Detection Vad-R1: Towards Video Anomaly Reasoning via Perception-to-Cognition Chain-of-Thought

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-06T13:30:09.804400Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T13:30:09.804400Z digest=sha256:6fdc1811da0c938396511735e2344a5350cfe8ef39e870b89614d2318a85fe01

Observation 2d58dce4-7863-4b4e-91da-9b45bf647994 · inbound

ESOM: Efficiently Understanding Streaming Video Anomalies with Open-world Dynamic Definitions cites this paper.

ESOM: Efficiently Understanding Streaming Video Anomalies with Open-world Dynamic Definitions Vad-R1: Towards Video Anomaly Reasoning via Perception-to-Cognition Chain-of-Thought

Reference 15

Resolution
verified exact
arxiv_id, observed 2026-05-11T05:51:03.085000Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-05-10T17:55:32.945721Z digest=sha256:8872228a1f6f1305bc82a1f69eaeb8dad5144b600c13453d180cf1148c9d35e0

Observation ba04cbef-78eb-49da-b50f-06c5f2bf4820 · inbound

MAVEN: A Multi-stage Agentic Annotation Pipeline for Video Reasoning Tasks cites this paper.

MAVEN: A Multi-stage Agentic Annotation Pipeline for Video Reasoning Tasks Vad-R1: Towards Video Anomaly Reasoning via Perception-to-Cognition Chain-of-Thought

Reference 13

Resolution
verified exact
arxiv_id, observed 2026-05-22T07:36:13.970967Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-05-22T07:36:09.562362Z digest=sha256:ef1380a31ad7fa1dd3e293913e1ebba5e2eb9cc77b9c79c9e542b79147c3e71f

Observation 7847a13a-472b-4470-a74c-608e47558be7 · inbound

MAVEN: A Multi-stage Agentic Annotation Pipeline for Video Reasoning Tasks cites this paper.

MAVEN: A Multi-stage Agentic Annotation Pipeline for Video Reasoning Tasks Vad-R1: Towards Video Anomaly Reasoning via Perception-to-Cognition Chain-of-Thought

Reference 13

Resolution
unresolved
no resolver link, observed 2026-07-12T16:17:21.074502Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T16:17:21.074502Z digest=sha256:98a5386aac06be36c1a35760054739dc1a33497ecc910d924a7aa151880a2f78

Observation 615540cb-879f-45fc-b9cc-2eaa3b41bae6 · inbound

O-VAD: Industrial Video Anomaly Detection through Object-Centric Tracking and Reasoning cites this paper.

O-VAD: Industrial Video Anomaly Detection through Object-Centric Tracking and Reasoning Vad-R1: Towards Video Anomaly Reasoning via Perception-to-Cognition Chain-of-Thought

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-01T15:56:41.756320Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T15:56:41.756320Z digest=sha256:e03f27301d8336f08962faf9c123a99b0c33fb517b16f035c84c7c459b72d2a1

Observation e5c0632f-ce6a-42b3-a323-029901b15531 · inbound

From Detection to Understanding: TAR and TAR-Bench for Multi-Task Traffic Anomaly Reasoning cites this paper.

From Detection to Understanding: TAR and TAR-Bench for Multi-Task Traffic Anomaly Reasoning Vad-R1: Towards Video Anomaly Reasoning via Perception-to-Cognition Chain-of-Thought

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-14T04:15:47.172052Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-14T04:15:47.172052Z digest=sha256:a79fffdc55e87217c219842b5cbc6c57b15a2fa826aaf1a011bcd47b41c9ac84