Pith. sign in

Paper Citation Record · LEDGER

Large Language Models for Crash Detection in Video: A Survey of Methods, Datasets, and Challenges

As of 11 August 2026, this Paper Citation Record lists 97 of 97 outbound references and 2 inbound Pith citation observations for arXiv:2507.02074.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2507.02074 v2

Coverage vector

measured 97 of 97 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-06T20:43:07.779736Z

measured 99 of 99 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-10T06:31:04.303077+00:00

measured 2 of 2 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-06-28T14:39:27.242094Z

measured 1 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Reference resolution

97 of 97 outbound references displayed

  • verified exact8
  • verified fuzzy38
  • unresolved46
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch5

External citation measurements

0
arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Outbound references

Observation 7a0a023e-1640-479c-8230-95ecd1fee14f · outbound

This paper cites Traffic monitoring and accident detection at intersections,.

Large Language Models for Crash Detection in Video: A Survey of Methods, Datasets, and Challenges Traffic monitoring and accident detection at intersections,

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-06T20:43:01.707781Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:43:01.707781Z digest=sha256:62730cc42e02ccf724c9e521ee29137a5df02d13a23be85fc891e53050add918

Observation c181588d-b6cb-4a3f-bc51-1a7a23a84639 · outbound

This paper cites A survey of vision-based trajec- tory learning and analysis for surveillance,.

Large Language Models for Crash Detection in Video: A Survey of Methods, Datasets, and Challenges A survey of vision-based trajec- tory learning and analysis for surveillance,

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-06T20:43:01.751587Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:43:01.751587Z digest=sha256:18f3db5e55472859d7839e1e91fa822e95629760aef5a821c02f113f4655efe0

Observation eb5f2631-df82-4955-818f-e45ad5f2f003 · outbound

This paper cites Trajectory-based anomalous event detection,.

Large Language Models for Crash Detection in Video: A Survey of Methods, Datasets, and Challenges Trajectory-based anomalous event detection,

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-06T20:43:01.863066Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:43:01.863066Z digest=sha256:6707180ad8be83ea1db8c075a139e1a39e808500432b33f6c3ac2ba9745785a4

Observation af03812a-3686-4184-84ee-a0e3a288fe7d · outbound

This paper cites Development of artificial neural network models to predict driver injury severity in traffic accidents at signalized intersections,.

Large Language Models for Crash Detection in Video: A Survey of Methods, Datasets, and Challenges Development of artificial neural network models to predict driver injury severity in traffic accidents at signalized intersections,

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-06T20:43:01.915387Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:43:01.915387Z digest=sha256:5355cf445ec91d30c611c5f188dd1d336c01790eb14bd05c40447204b3c4bd43

Observation 7475dead-8473-4927-bb23-ac060d15e6e1 · outbound

This paper cites Severity of driver injury and vehicle damage in traffic crashes at intersections: a bayesian hierarchical analysis,.

Large Language Models for Crash Detection in Video: A Survey of Methods, Datasets, and Challenges Severity of driver injury and vehicle damage in traffic crashes at intersections: a bayesian hierarchical analysis,

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-06T20:43:01.988904Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:43:01.988904Z digest=sha256:0181a4c42495f66d6d682a1daa81f21d2bbe98a182a05849e6e509485524d35e

Observation 2dcb33f7-1c66-470b-8d73-e91b9984e679 · outbound

This paper cites Two-stream convolutional networks for action recognition in videos,.

Large Language Models for Crash Detection in Video: A Survey of Methods, Datasets, and Challenges Two-stream convolutional networks for action recognition in videos,

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-06T20:43:02.099954Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:43:02.099954Z digest=sha256:f73355da26b8d347571fa96976f209626a540f0377a9d3a46883bcd7430ccd74

Observation b9e77609-0058-4ad8-8a89-7f5eecde20c5 · outbound

This paper cites Learning spatiotemporal features with 3d convolutional networks,.

Large Language Models for Crash Detection in Video: A Survey of Methods, Datasets, and Challenges Learning spatiotemporal features with 3d convolutional networks,

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-06T20:43:02.171761Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:43:02.171761Z digest=sha256:ec9bf572872cb1197ffed6d5ebbc7b25eaac0aeb374cd12d9e8d5c28fc8976ea

Observation c2a74191-4297-4f9a-acf7-95cd6c5c69a5 · outbound

This paper cites Quo vadis, action recognition? a new model and the kinetics dataset,.

Large Language Models for Crash Detection in Video: A Survey of Methods, Datasets, and Challenges Quo vadis, action recognition? a new model and the kinetics dataset,

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-06T20:43:02.230067Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:43:02.230067Z digest=sha256:c311295251ca0d9ef4ef04edbac7d98de3b7a76da3ca40fc3c8530de01418d83

Observation c1e8de97-b561-499b-9a09-5acac53269e2 · outbound

This paper cites Slowfast networks for video recognition,.

Large Language Models for Crash Detection in Video: A Survey of Methods, Datasets, and Challenges Slowfast networks for video recognition,

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-06T20:43:02.302154Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:43:02.302154Z digest=sha256:a1f0694c5219cb5daf4bbb22467850eff56ad9d71316cdff9ca62cc63d20abd3

Observation 8ae9d92e-0c55-4c14-b260-63172779b037 · outbound

This paper cites Real-world anomaly detection in surveillance videos,.

Large Language Models for Crash Detection in Video: A Survey of Methods, Datasets, and Challenges Real-world anomaly detection in surveillance videos,

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-06T20:43:02.364456Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:43:02.364456Z digest=sha256:09e3a0aac59622faf989c6c0fe8c43aac27b34d9f4beebc5fadd6a3ae37cf3a2

Observation 762e0d57-d62c-4959-8d6e-688d5eb3f115 · outbound

This paper cites Future frame prediction for anomaly detection–a new baseline,.

Large Language Models for Crash Detection in Video: A Survey of Methods, Datasets, and Challenges Future frame prediction for anomaly detection–a new baseline,

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-06T20:43:02.398110Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:43:02.398110Z digest=sha256:7d6f13f624bc38a67601075270a47ad9de315ab9e6c1ced1b69900cbce2eb609

Observation 45769cbc-de9b-4547-9175-371aa7ada1b0 · outbound

This paper cites Learning memory-guided nor- mality for anomaly detection,.

Large Language Models for Crash Detection in Video: A Survey of Methods, Datasets, and Challenges Learning memory-guided nor- mality for anomaly detection,

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-06T20:43:02.461330Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:43:02.461330Z digest=sha256:760a3baf2c792586cd48b961739d74c4125c2963da447341a392d7e7e5615c95

Observation 8a4ed46c-ec87-4d43-9dbf-1810b13cdf36 · outbound

This paper cites Uniter: Universal image-text representation learning,.

Large Language Models for Crash Detection in Video: A Survey of Methods, Datasets, and Challenges Uniter: Universal image-text representation learning,

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-06T20:43:02.527630Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:43:02.527630Z digest=sha256:736ee07b7f5703130ad3d70dc6b1486f7fe319796249336a75e3f1032e529cbf

Observation 0d915eb2-cc12-4d91-a601-a02dd232d72c · outbound

This paper cites Align before fuse: Vision and language representation learning with momentum distillation,.

Large Language Models for Crash Detection in Video: A Survey of Methods, Datasets, and Challenges Align before fuse: Vision and language representation learning with momentum distillation,

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-06T20:43:02.565716Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:43:02.565716Z digest=sha256:203fbf3f41af14b5f4faa8521bb1510bdec30133774ab109d2df5fd4106f50c9

Observation cc6bc847-2d23-4df1-be3b-f5ccf41a3dda · outbound

This paper cites Less is more: Clipbert for video-and-language learning via sparse sampling,.

Large Language Models for Crash Detection in Video: A Survey of Methods, Datasets, and Challenges Less is more: Clipbert for video-and-language learning via sparse sampling,

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-06T20:43:02.649490Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:43:02.649490Z digest=sha256:d7fb5929866ecc90de90ea241c9f9e370e969ddac034a2a7244ce80170e53f7c

Observation a7cd5b0e-f429-4709-8374-5fc24d5e9207 · outbound

This paper cites Videoclip: Contrastive pre-training for zero-shot video-text understanding,.

Large Language Models for Crash Detection in Video: A Survey of Methods, Datasets, and Challenges Videoclip: Contrastive pre-training for zero-shot video-text understanding,

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-06T20:43:02.726302Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:43:02.726302Z digest=sha256:8a442e4bd3a287c24143f892418d1ac6e142f5f06989d5184c3b7875f76cdeb0

Observation ca12865d-0f16-49f5-8ce6-5b324cb69aa9 · outbound

This paper cites GPT-4 Technical Report.

Large Language Models for Crash Detection in Video: A Survey of Methods, Datasets, and Challenges GPT-4 Technical Report

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-06T20:43:02.777750Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:43:02.777750Z digest=sha256:0389f592e68f14714457a855e7fbbedf981f4a3ca4799c1cc9218931806978fa

Observation 2320f2c0-64dc-4ff6-a19c-0b4923c6b4b7 · outbound

This paper cites Flamingo: a visual language model for few-shot learning,.

Large Language Models for Crash Detection in Video: A Survey of Methods, Datasets, and Challenges Flamingo: a visual language model for few-shot learning,

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-06T20:43:02.860733Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:43:02.860733Z digest=sha256:5bfaa1b515b85340abb9f20b3dfca5f87b3b44e68cdfb66d0b17fb3c966317b2

Observation ef9a3f3c-0c3d-4a13-a339-aecce4443c8a · outbound

This paper cites Blip-2: Bootstrapping language-image pre-training with frozen image encoders and large language models,.

Large Language Models for Crash Detection in Video: A Survey of Methods, Datasets, and Challenges Blip-2: Bootstrapping language-image pre-training with frozen image encoders and large language models,

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:43:18.064337Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T20:43:02.905256Z digest=sha256:0c8403aff7934e472248e1dcdabc4d8c1f1b84b425abe1e4eb53c0fb5fb8e369

Observation 7da91b94-5ee5-4c89-92c3-19dd261cdbed · outbound

This paper cites Anomalous video event detection using spatiotemporal context,.

Large Language Models for Crash Detection in Video: A Survey of Methods, Datasets, and Challenges Anomalous video event detection using spatiotemporal context,

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:43:17.890932Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T20:43:02.943513Z digest=sha256:c61472142ec83358da44c84533bf11c5e084c99decfce1b47de14abc63ebf0f6

Observation 46d4f0b0-2231-4793-95df-c756a1ab9b1c · outbound

This paper cites Detecting anomalies in people’s trajectories using spectral graph analysis,.

Large Language Models for Crash Detection in Video: A Survey of Methods, Datasets, and Challenges Detecting anomalies in people’s trajectories using spectral graph analysis,

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:43:17.678953Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T20:43:02.974628Z digest=sha256:aaffa67a7eb01220767a6e4f77443c02f18f478f92a8956327906823ddf8efc9

Observation 314d5834-1e69-40c3-9d01-9abd50c48051 · outbound

This paper cites Using support vector machine models for crash injury severity analysis,.

Large Language Models for Crash Detection in Video: A Survey of Methods, Datasets, and Challenges Using support vector machine models for crash injury severity analysis,

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:43:17.464120Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T20:43:03.022540Z digest=sha256:4eab81d8cbdd427b355a23926bfea1c5abc91caffd34e6c05d263f304e5192da

Observation 9540b7bf-4036-4920-9371-8967656132ea · outbound

This paper cites Accident detection system using image processing and mdr,.

Large Language Models for Crash Detection in Video: A Survey of Methods, Datasets, and Challenges Accident detection system using image processing and mdr,

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:43:17.285156Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T20:43:03.096962Z digest=sha256:a3e80cf8f254e79200cd9d65084c67d59b4b80ea31a59e5217a6771d31c74e0b

Observation 66facb1f-0cab-46fa-ac98-36c9e68c0d40 · outbound

This paper cites Temporal segment networks: Towards good practices for deep action recognition,.

Large Language Models for Crash Detection in Video: A Survey of Methods, Datasets, and Challenges Temporal segment networks: Towards good practices for deep action recognition,

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:43:17.095290Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T20:43:03.146316Z digest=sha256:44041cf669187b09fa043d45e475129263d05eb0d38be9924f1200cec44aa388

Observation c6f9de2a-893b-4ad6-a4ce-5c6cf1c02cf2 · outbound

This paper cites Multimodal machine learning: A survey and taxonomy,.

Large Language Models for Crash Detection in Video: A Survey of Methods, Datasets, and Challenges Multimodal machine learning: A survey and taxonomy,

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:43:16.847350Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T20:43:03.180713Z digest=sha256:68c62d807d084bc8c589a2444e7db09a2a8a78d801265e908818a8dddfd04f7b

Observation f350ee47-b512-4855-9074-cec27d8e6948 · outbound

This paper cites Multimodal deep learning,.

Large Language Models for Crash Detection in Video: A Survey of Methods, Datasets, and Challenges Multimodal deep learning,

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:43:16.714894Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T20:43:03.231501Z digest=sha256:7445b985d14d34b0d2fd318ba6dec6587fdca33c3d48264acabf86a42033b377

Observation 545500c2-2f2f-4e0a-b70f-1704dd55d722 · outbound

This paper cites Deep multimodal learning: A survey on recent advances and trends,.

Large Language Models for Crash Detection in Video: A Survey of Methods, Datasets, and Challenges Deep multimodal learning: A survey on recent advances and trends,

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:43:16.548682Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T20:43:03.284396Z digest=sha256:27c8206ae9c4aa456de61ed6d9c156fd97692ecf0bd29c98c476af1ffede869e

Observation 354e4939-3157-42f6-b4cd-52aaf15c6c4d · outbound

This paper cites SimVLM: Simple Visual Language Model Pretraining with Weak Supervision.

Large Language Models for Crash Detection in Video: A Survey of Methods, Datasets, and Challenges SimVLM: Simple Visual Language Model Pretraining with Weak Supervision

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-06T20:43:03.339629Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:43:03.339629Z digest=sha256:d1f8ac5ef16dc5215ad5d56088099d6cd43d95cbae655a6778873e8606e2bf32

Observation ccb2bec7-8af5-441a-aa14-603706962798 · outbound

This paper cites All in one: Exploring unified video-language pre-training,.

Large Language Models for Crash Detection in Video: A Survey of Methods, Datasets, and Challenges All in one: Exploring unified video-language pre-training,

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:43:16.365975Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T20:43:03.378057Z digest=sha256:e472bfc39fcb2aa7d300336789876cc30cf2b18e4ad3baf29df584cfaa8d9b9d

Observation 6da57f0e-1e22-43e5-89db-908caa2860bc · outbound

This paper cites The cityscapes dataset for semantic urban scene understanding,.

Large Language Models for Crash Detection in Video: A Survey of Methods, Datasets, and Challenges The cityscapes dataset for semantic urban scene understanding,

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:43:16.104420Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T20:43:03.426557Z digest=sha256:f00d8b73f9d8d1df41e4e729f8ecccdfa022914ec84fc1fce64267578f69e05a

Observation 542aa650-d6ad-4e3c-8889-155d0b67985d · outbound

This paper cites nuscenes: A multimodal dataset for autonomous driving,.

Large Language Models for Crash Detection in Video: A Survey of Methods, Datasets, and Challenges nuscenes: A multimodal dataset for autonomous driving,

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-06T20:43:03.452329Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:43:03.452329Z digest=sha256:f72105f1751b02f283d14e636954d173e764eddf078390ff785349bdefce450e

Observation ba54830f-f768-497f-bf37-d5e80076dc7e · outbound

This paper cites Are we ready for autonomous driving? the kitti vision benchmark suite,.

Large Language Models for Crash Detection in Video: A Survey of Methods, Datasets, and Challenges Are we ready for autonomous driving? the kitti vision benchmark suite,

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:43:15.881441Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T20:43:03.511845Z digest=sha256:70afbd89aa5ab5d20ee415274bba6d38be011c11ac9bb58aad94957d6087fdc7

Observation 19ce6e3b-dbdf-4dd9-901a-bc713e0c3bdf · outbound

This paper cites Edge computing: Vision and challenges,.

Large Language Models for Crash Detection in Video: A Survey of Methods, Datasets, and Challenges Edge computing: Vision and challenges,

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-06T20:43:03.545890Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:43:03.545890Z digest=sha256:38972bb011b95eecff4c69101dea9556012c88515b535692591b135aba24e4c1

Observation e2e764c7-9eec-447c-9be3-1f8a6ff2552c · outbound

This paper cites Edge intelligence: Paving the last mile of artificial intelligence with edge computing,.

Large Language Models for Crash Detection in Video: A Survey of Methods, Datasets, and Challenges Edge intelligence: Paving the last mile of artificial intelligence with edge computing,

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-06T20:43:03.594786Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:43:03.594786Z digest=sha256:9655ec2f1b455029c9496f181025703ebba6affc65894be1a94da91f843177d0

Observation 3bba2dc6-21de-4146-94df-ce35b2cb5aaa · outbound

This paper cites Videobert: A joint model for video and language representa- tion learning,.

Large Language Models for Crash Detection in Video: A Survey of Methods, Datasets, and Challenges Videobert: A joint model for video and language representa- tion learning,

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:43:15.615669Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T20:43:03.634399Z digest=sha256:bef76a9ffb123875dde401a645860bca13e3d6a855a616821ff947385e6d8ffb

Observation bd6011d1-151f-439f-a5ed-162557e87440 · outbound

This paper cites Vivit: A video vision transformer,.

Large Language Models for Crash Detection in Video: A Survey of Methods, Datasets, and Challenges Vivit: A video vision transformer,

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:43:15.456053Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T20:43:03.667379Z digest=sha256:15181694942dcb8e6f0deb85a0749292ee3e9b9fa283a42132e48e598dadd3dc

Observation 9d5659d3-d7e4-42a9-93ba-4ad4b01dfb20 · outbound

This paper cites Videogpt: Video generation using vq-vae and transformers,.

Large Language Models for Crash Detection in Video: A Survey of Methods, Datasets, and Challenges Videogpt: Video generation using vq-vae and transformers,

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:43:15.309881Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T20:43:03.731997Z digest=sha256:118beb0ec5b30001e458a9826308e3557a26528954e3ca6de108a037ec1d2595

Observation 7ef8defa-413b-433e-9a5a-4870e5fa5414 · outbound

This paper cites Video-LLaMA: An Instruction-tuned Audio-Visual Language Model for Video Understanding.

Large Language Models for Crash Detection in Video: A Survey of Methods, Datasets, and Challenges Video-LLaMA: An Instruction-tuned Audio-Visual Language Model for Video Understanding

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-06T20:43:03.778153Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:43:03.778153Z digest=sha256:9ab106caf17ab8558167311ef6cfe0446147a5c5f1013aebfcd98cb311926570

Observation 288793db-74eb-4782-b296-5890ae443464 · outbound

This paper cites Improved Baselines with Visual Instruction Tuning.

Large Language Models for Crash Detection in Video: A Survey of Methods, Datasets, and Challenges Improved Baselines with Visual Instruction Tuning

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-06T20:43:03.846253Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:43:03.846253Z digest=sha256:de68f9bbfbc474f77e025531987ff349bfb6fb3848cb4ddb2122b886c75cd913

Observation d137ac91-53fc-401f-9595-d28fe3e60348 · outbound

This paper cites PaLM-E: An Embodied Multimodal Language Model.

Large Language Models for Crash Detection in Video: A Survey of Methods, Datasets, and Challenges PaLM-E: An Embodied Multimodal Language Model

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-06T20:43:03.913715Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:43:03.913715Z digest=sha256:adf13b7a91aa39dcbf378c842500f294b86b8b34200b748023502186ee9fc004

Observation 6fad97ba-8ab2-499e-b3b6-d74d9cb07c50 · outbound

This paper cites Gpt-4v(ision) system card,.

Large Language Models for Crash Detection in Video: A Survey of Methods, Datasets, and Challenges Gpt-4v(ision) system card,

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:43:15.210868Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T20:43:03.964175Z digest=sha256:4c168451b834f57441b353ca65c45721671f331baf6d7cdbcf1b5beb795d57ac

Observation 00e885d6-1256-4256-aa5c-44a0ad06ce9e · outbound

This paper cites Gemini 1.5: Unlocking multimodal understanding across millions of tokens of context,.

Large Language Models for Crash Detection in Video: A Survey of Methods, Datasets, and Challenges Gemini 1.5: Unlocking multimodal understanding across millions of tokens of context,

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:43:15.151435Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T20:43:04.030238Z digest=sha256:1519c8b961021bbfdccb7a6d3df25b258752007b38354fa5ca5eaaccdac2f865

Observation b7be538e-2e9e-4368-a47b-acbe008365b6 · outbound

This paper cites Sora: Creating video from text,.

Large Language Models for Crash Detection in Video: A Survey of Methods, Datasets, and Challenges Sora: Creating video from text,

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:43:14.993735Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T20:43:04.098209Z digest=sha256:d4fd27c9f6a7fdaa747cacbce578751b2c5fcae11149c0f7ec8bd24260589807

Observation a12d3737-8413-4032-9198-c63c793d755e · outbound

This paper cites Crash: Crash recognition and anticipation system harnessing with context-aware and temporal focus attentions,.

Large Language Models for Crash Detection in Video: A Survey of Methods, Datasets, and Challenges Crash: Crash recognition and anticipation system harnessing with context-aware and temporal focus attentions,

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:43:14.782918Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T20:43:04.154363Z digest=sha256:49c1ee767387d8720b5d0bee21b9bfd25f176b60929b6e6a247dc05b761a3f3c

Observation f0149b61-9c89-4859-b042-ebcfd31e88a7 · outbound

This paper cites When language and vision meet road safety: leveraging multimodal large language models for video-based traffic accident analysis.

Large Language Models for Crash Detection in Video: A Survey of Methods, Datasets, and Challenges When language and vision meet road safety: leveraging multimodal large language models for video-based traffic accident analysis

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-06T20:43:04.297014Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:43:04.297014Z digest=sha256:3177b4e4c70e5b9fb8dc7f39604cb217bbe13b91f5fb40ce121790c7abe2fc8b

Observation 4a95e254-d30b-42cf-a757-3dc95da00af2 · outbound

This paper cites Learning transferable visual models from natural language supervision,.

Large Language Models for Crash Detection in Video: A Survey of Methods, Datasets, and Challenges Learning transferable visual models from natural language supervision,

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-06T20:43:04.364558Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:43:04.364558Z digest=sha256:1fe61a8176a1d02bc27e718901783c072037fdf62c8006ecb82370ef403f02a9

Observation 17b3a461-14f7-47f8-a8ff-fbd05e9e0b22 · outbound

This paper cites CLIP4Clip: An Empirical Study of CLIP for End to End Video Clip Retrieval.

Large Language Models for Crash Detection in Video: A Survey of Methods, Datasets, and Challenges CLIP4Clip: An Empirical Study of CLIP for End to End Video Clip Retrieval

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-06T20:43:04.406744Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:43:04.406744Z digest=sha256:88ec153f518a1b3079878805681e629522dc32619e6b9870e66fc85eb4689235

Observation 448b17bc-4b82-4f38-aa8e-10cba87233f4 · outbound

This paper cites Llama 2: Open Foundation and Fine-Tuned Chat Models.

Large Language Models for Crash Detection in Video: A Survey of Methods, Datasets, and Challenges Llama 2: Open Foundation and Fine-Tuned Chat Models

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-06T20:43:04.442302Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:43:04.442302Z digest=sha256:df7ca5e446ab46fd3269e4b474d0d51c7459b685bbf450d4e772c8f851004e48

Observation abd4cb4e-e83c-44dc-8829-7d0d6ddcc42b · outbound

This paper cites Quick and robust detection of crash events from video streams,.

Large Language Models for Crash Detection in Video: A Survey of Methods, Datasets, and Challenges Quick and robust detection of crash events from video streams,

Reference 49

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:43:14.537088Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T20:43:04.505518Z digest=sha256:a881d23a986b383492f6024758e9400a9760231d4db74e2e015fbbfd4f7d8864

Observation 1de8b4b1-d6cf-4634-9aca-c2428ce4f0ad · outbound

This paper cites A real-time system for detection of road accidents from video streams,.

Large Language Models for Crash Detection in Video: A Survey of Methods, Datasets, and Challenges A real-time system for detection of road accidents from video streams,

Reference 50

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:43:14.256997Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T20:43:04.570143Z digest=sha256:ce4f9b5cca924c94378de4b2f17fa71ae34e44054b926ba3bd6fa4da56d37dc8

Observation 3a3c7fb1-97ac-4f3a-987e-eb14d147d8a6 · outbound

This paper cites Video-based traffic accident detection: A survey,.

Large Language Models for Crash Detection in Video: A Survey of Methods, Datasets, and Challenges Video-based traffic accident detection: A survey,

Reference 51

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:43:13.942348Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T20:43:04.622770Z digest=sha256:4be387f69af4f9f5c2711320115b4d83c15256980364f45ff1071ca13d5a9dcf

Observation 87610921-fe17-4f72-9711-b5b7c8f55caf · outbound

This paper cites LoRA: Low-Rank Adaptation of Large Language Models.

Large Language Models for Crash Detection in Video: A Survey of Methods, Datasets, and Challenges LoRA: Low-Rank Adaptation of Large Language Models

Reference 52

Resolution
unresolved
no resolver link, observed 2026-08-06T20:43:04.679275Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:43:04.679275Z digest=sha256:816cf6007ed9271e22e1f97ac58cb8b966aaedbc862ab6769cd8814fb5e3a47e

Observation 4ad41515-89e2-4ee8-aa36-f17d7c4f00e8 · outbound

This paper cites Pentagon relation and Biedenharn-Elliott identity.

Large Language Models for Crash Detection in Video: A Survey of Methods, Datasets, and Challenges Pentagon relation and Biedenharn-Elliott identity

Reference 53

Resolution
metadata mismatch
local_arxiv, observed 2026-08-06T20:43:10.153649Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T20:43:04.749137Z digest=sha256:12a7d818fa38f777eea060f5e07166fd8fe4a01d97a26f5a0328dfe0c55b2a16

Observation a45733c2-3731-4332-a915-3102c9cd5cf3 · outbound

This paper cites Harnessing Neuron Stability to Improve DNN Verification.

Large Language Models for Crash Detection in Video: A Survey of Methods, Datasets, and Challenges Harnessing Neuron Stability to Improve DNN Verification

Reference 54

Resolution
metadata mismatch
local_arxiv, observed 2026-08-06T20:43:10.016965Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T20:43:04.793014Z digest=sha256:76f3dfdb5700d38aac9652b5dfe79edc0b9c4633465b8ab9aa48e973b3203966

Observation 240f0f88-0037-4bec-9b7f-087d9756e0e4 · outbound

This paper cites Partially hyperbolic diffeomorphisms that are center fixing.

Large Language Models for Crash Detection in Video: A Survey of Methods, Datasets, and Challenges Partially hyperbolic diffeomorphisms that are center fixing

Reference 55

Resolution
metadata mismatch
local_arxiv, observed 2026-08-06T20:43:09.921329Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T20:43:04.825480Z digest=sha256:0baea6f2b914f65625cf056ad54a6a5aa15458ec916f6123b12afec18216e13f

Observation 4956d886-5e93-4272-b982-2d98a7eb761a · outbound

This paper cites On the consistency of relative facts.

Large Language Models for Crash Detection in Video: A Survey of Methods, Datasets, and Challenges On the consistency of relative facts

Reference 56

Resolution
unresolved
no resolver link, observed 2026-08-06T20:43:04.866116Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:43:04.866116Z digest=sha256:16f1b2459ac7bee5b72f5e6487a955368e914bb64be73df722ee8d70f0c1d672

Observation 645beef6-394d-46aa-8ece-d71fc68a0b65 · outbound

This paper cites Leveraging video-llms for crash detection and narrative generation: Performance analysis and challenges,.

Large Language Models for Crash Detection in Video: A Survey of Methods, Datasets, and Challenges Leveraging video-llms for crash detection and narrative generation: Performance analysis and challenges,

Reference 57

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:43:13.654406Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T20:43:04.920193Z digest=sha256:c56bb7887c6fe8696f6a34a31dab971174a214859087cce9c9ca574de5e21aaa

Observation 35443ac5-fef4-4937-b880-52f9a49d9bfa · outbound

This paper cites Joint Discriminative and Generative Learning for Person Re-identification.

Large Language Models for Crash Detection in Video: A Survey of Methods, Datasets, and Challenges Joint Discriminative and Generative Learning for Person Re-identification

Reference 58

Resolution
verified exact
local_arxiv, observed 2026-08-06T20:43:09.775500Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T20:43:04.980461Z digest=sha256:b183631a5899b8607a9470e5457ab18286df837f66d070a59476cb220b2bba33

Observation 7b496f6e-da04-4812-97ca-db92a850519c · outbound

This paper cites Trafficlens: A novel system for video-to-text conver- sion in multi-camera traffic environments,.

Large Language Models for Crash Detection in Video: A Survey of Methods, Datasets, and Challenges Trafficlens: A novel system for video-to-text conver- sion in multi-camera traffic environments,

Reference 59

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:43:13.417130Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T20:43:05.013039Z digest=sha256:95ff89ee67ba954bb697e4d38957d5b01196a13cff5105033831fdf564153abb

Observation 7d91cfaf-5882-4ede-88c1-5be018cfc9ee · outbound

This paper cites Enhancing Multilingual Capabilities of Large Language Models through Self-Distillation from Resource-Rich Languages.

Large Language Models for Crash Detection in Video: A Survey of Methods, Datasets, and Challenges Enhancing Multilingual Capabilities of Large Language Models through Self-Distillation from Resource-Rich Languages

Reference 60

Resolution
verified exact
local_arxiv, observed 2026-08-06T20:43:09.652431Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T20:43:05.071380Z digest=sha256:f3bf1827bbdae53a0aa93125a67d1ea5a81cbbdad9f74ae6e7bf98bc447affec

Observation 3377578f-c309-4f20-8dff-a41c5b3d58d6 · outbound

This paper cites Dad: A dashcam accident dataset,.

Large Language Models for Crash Detection in Video: A Survey of Methods, Datasets, and Challenges Dad: A dashcam accident dataset,

Reference 61

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:43:13.156123Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T20:43:05.117717Z digest=sha256:8a39a9e87a7dafc45346fd50dbc6eb352149221a75c2d904715214bfa99e38cc

Observation 7023c4c6-fb00-47d1-bced-8fdcceaeaa7c · outbound

This paper cites Cadp: A novel dataset for car accident detection and prediction from police reports,.

Large Language Models for Crash Detection in Video: A Survey of Methods, Datasets, and Challenges Cadp: A novel dataset for car accident detection and prediction from police reports,

Reference 62

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:43:12.935784Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T20:43:05.153296Z digest=sha256:ecde5151be7bdf308be5d700a172e2ecdf4a7e9d75f7f45086ecbeff7700f7cb

Observation ac938ad0-26ac-48b0-bb4c-abbdf1f6b792 · outbound

This paper cites BDD100K: A Diverse Driving Dataset for Heterogeneous Multitask Learning.

Large Language Models for Crash Detection in Video: A Survey of Methods, Datasets, and Challenges BDD100K: A Diverse Driving Dataset for Heterogeneous Multitask Learning

Reference 63

Resolution
unresolved
no resolver link, observed 2026-08-06T20:43:05.206991Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:43:05.206991Z digest=sha256:b2db1c3841551f7381fe6bbe4be49d4d9a0d2f7a938aef6df90e42cbcb04b82e

Observation 37a20d0b-c064-4295-a336-bf917aa640d9 · outbound

This paper cites Bulk-Boundary Correspondence in the Quantum Hall Effect.

Large Language Models for Crash Detection in Video: A Survey of Methods, Datasets, and Challenges Bulk-Boundary Correspondence in the Quantum Hall Effect

Reference 64

Resolution
metadata mismatch
local_arxiv, observed 2026-08-06T20:43:09.499953Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T20:43:05.265806Z digest=sha256:7059c816211c4702a2c2924f01cda11046556b935a867879e163a11633f3a4e3

Observation c8fcad50-7dac-412b-834f-4a7807d1b478 · outbound

This paper cites ScVLM: Enhancing Vision-Language Model for Safety-Critical Event Understanding.

Large Language Models for Crash Detection in Video: A Survey of Methods, Datasets, and Challenges ScVLM: Enhancing Vision-Language Model for Safety-Critical Event Understanding

Reference 65

Resolution
verified exact
local_arxiv, observed 2026-08-06T20:43:09.342500Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T20:43:05.322723Z digest=sha256:e689160b509cf3754a01dfb62e7bc1098aee3857ff090a329b4009383a08ef0f

Observation 0a3fc6ec-0e31-46ea-8ae6-ceae26073d34 · outbound

This paper cites TrafficVLM: A Controllable Visual Language Model for Traffic Video Captioning.

Large Language Models for Crash Detection in Video: A Survey of Methods, Datasets, and Challenges TrafficVLM: A Controllable Visual Language Model for Traffic Video Captioning

Reference 66

Resolution
verified exact
local_arxiv, observed 2026-08-06T20:43:09.183265Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T20:43:05.399630Z digest=sha256:77635601cfaa07778f71eb5b2a97349dc4e99b745eb92a5e492e3da0de365a39

Observation bd857ff8-0747-497f-9992-a83301ee16c4 · outbound

This paper cites Crash: A context-aware attention-based framework for crash anticipation,.

Large Language Models for Crash Detection in Video: A Survey of Methods, Datasets, and Challenges Crash: A context-aware attention-based framework for crash anticipation,

Reference 67

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:43:12.649915Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T20:43:05.463586Z digest=sha256:ba4a262264a91eb6d27cd0de0b0f0c924b1ece89db8b857e036ec15956d423f8

Observation db865e62-ee75-4544-9a9b-fff300e955c1 · outbound

This paper cites MiniGPT-4: Enhancing Vision-Language Understanding with Advanced Large Language Models.

Large Language Models for Crash Detection in Video: A Survey of Methods, Datasets, and Challenges MiniGPT-4: Enhancing Vision-Language Understanding with Advanced Large Language Models

Reference 68

Resolution
unresolved
no resolver link, observed 2026-08-06T20:43:05.505355Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:43:05.505355Z digest=sha256:ec54fdea8e920617fd193d5c63699149cc6c6801d2216dc39a8a89d3624ba6f3

Observation d6bb9525-c2e1-4c40-85f6-0f802bb7578c · outbound

This paper cites Slowfast networks for video recognition,.

Large Language Models for Crash Detection in Video: A Survey of Methods, Datasets, and Challenges Slowfast networks for video recognition,

Reference 69

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:43:12.373008Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T20:43:05.565916Z digest=sha256:98077281216ec62423570a3053c793857efe18467367693e9de868f2f09042bd

Observation 2bceff8e-4009-4032-a3e3-a4939c79406e · outbound

This paper cites Self-supervised anomaly detection: A survey and outlook,.

Large Language Models for Crash Detection in Video: A Survey of Methods, Datasets, and Challenges Self-supervised anomaly detection: A survey and outlook,

Reference 70

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:43:12.073096Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T20:43:05.645978Z digest=sha256:7bed0f803276342e23874a3daec928e9527bd6089d0d7b68e8b0f839bac485dd

Observation eccf9c1e-c30b-4a11-8f5b-eb6dad9d780b · outbound

This paper cites Few-shot fast-adaptive anomaly detection,.

Large Language Models for Crash Detection in Video: A Survey of Methods, Datasets, and Challenges Few-shot fast-adaptive anomaly detection,

Reference 71

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:43:11.794299Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T20:43:05.691625Z digest=sha256:4a5be601d25c6e37dcdf2dff17695296ca0ffbcbc5f9c5591e5bdd61ee478c9e

Observation accd368e-1106-4b7f-afb7-66e42b0905cd · outbound

This paper cites An Image is Worth 16x16 Words: Transformers for Image Recognition at Scale.

Large Language Models for Crash Detection in Video: A Survey of Methods, Datasets, and Challenges An Image is Worth 16x16 Words: Transformers for Image Recognition at Scale

Reference 72

Resolution
unresolved
no resolver link, observed 2026-08-06T20:43:05.739969Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:43:05.739969Z digest=sha256:c2ad1a9dc0b567339bd38a955d9dc94ccf66985934634eda474fcce225c5aeea

Observation 7175c93d-b69a-401d-bf3f-7b819aa45d1c · outbound

This paper cites A Survey on Deep Learning Techniques for Video Anomaly Detection.

Large Language Models for Crash Detection in Video: A Survey of Methods, Datasets, and Challenges A Survey on Deep Learning Techniques for Video Anomaly Detection

Reference 73

Resolution
verified exact
local_arxiv, observed 2026-08-06T20:43:08.974237Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T20:43:05.823209Z digest=sha256:ef68ef0d5c63f8592e52cc1da7dec4261eb89fa64a797320195b863303b3b991

Observation e8ae7848-09dd-479d-af08-841db39a7e10 · outbound

This paper cites Swin transformer: Hierarchical vision transformer using shifted windows,.

Large Language Models for Crash Detection in Video: A Survey of Methods, Datasets, and Challenges Swin transformer: Hierarchical vision transformer using shifted windows,

Reference 74

Resolution
unresolved
no resolver link, observed 2026-08-06T20:43:05.873151Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:43:05.873151Z digest=sha256:c39e939c57c230d8378d4cf5b5a4b3cf8bc9d18808d728ff0efb477fd13e139a

Observation d4f5159c-5cb2-409d-821f-611f7c353325 · outbound

This paper cites Convolutional lstm network: A machine learning approach for precipitation nowcasting,.

Large Language Models for Crash Detection in Video: A Survey of Methods, Datasets, and Challenges Convolutional lstm network: A machine learning approach for precipitation nowcasting,

Reference 75

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:43:11.532931Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T20:43:05.939940Z digest=sha256:bf09655e63afad27f45fecc1f3857c12c58e149589a8fadd45c6bc5dce552be6

Observation 4d7b8a6e-051a-4240-a8bb-97af8b39e9e3 · outbound

This paper cites Video Anomaly Detection in 10 Years: A Survey and Outlook.

Large Language Models for Crash Detection in Video: A Survey of Methods, Datasets, and Challenges Video Anomaly Detection in 10 Years: A Survey and Outlook

Reference 76

Resolution
unresolved
no resolver link, observed 2026-08-06T20:43:06.026238Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:43:06.026238Z digest=sha256:4f8ae115becc9dd4630179ff9cb08be051da559b1cabdc36197b5e3a98612818

Observation 2c97d037-7e18-4489-90a6-291bf823bb0d · outbound

This paper cites ViViT: A Video Vision Transformer.

Large Language Models for Crash Detection in Video: A Survey of Methods, Datasets, and Challenges ViViT: A Video Vision Transformer

Reference 77

Resolution
unresolved
no resolver link, observed 2026-08-06T20:43:06.121709Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:43:06.121709Z digest=sha256:f8b4d6bedf0328a548d2a3a7b81c6c2adb4bf065d2f4590f2280c2db5c9d53cc

Observation 2f72261b-2234-4b57-95af-49b625d86f30 · outbound

This paper cites Bleu: a method for automatic evaluation of machine translation,.

Large Language Models for Crash Detection in Video: A Survey of Methods, Datasets, and Challenges Bleu: a method for automatic evaluation of machine translation,

Reference 78

Resolution
unresolved
no resolver link, observed 2026-08-06T20:43:06.207625Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:43:06.207625Z digest=sha256:164fb46d49c65bd5ec592a72033ee84590f5a8eb57fa0af379b17f1f0139ab4e

Observation a7d6492b-55f1-461d-8dc3-ba0780cc329b · outbound

This paper cites Meteor: An automatic metric for mt evaluation with improved correlation with human judgments,.

Large Language Models for Crash Detection in Video: A Survey of Methods, Datasets, and Challenges Meteor: An automatic metric for mt evaluation with improved correlation with human judgments,

Reference 79

Resolution
unresolved
no resolver link, observed 2026-08-06T20:43:06.264040Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:43:06.264040Z digest=sha256:fd1240f13325d039851a6be6a2950e02b8defc37928e4eac2e6f423b0444fb5e

Observation 2a39a169-d8a1-46ae-b00c-51b6e5263cc6 · outbound

This paper cites Cider: Consensus-based image description evaluation,.

Large Language Models for Crash Detection in Video: A Survey of Methods, Datasets, and Challenges Cider: Consensus-based image description evaluation,

Reference 80

Resolution
unresolved
no resolver link, observed 2026-08-06T20:43:06.334184Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:43:06.334184Z digest=sha256:8c25ce1e338d0eb72edb7e9bcf94c31003a06c19e03bcd0e42eb41a3d50bd84c

Observation c1740d64-d080-4406-84b4-848f689fc2d8 · outbound

This paper cites Sutd-trafficqa: A question answering benchmark and an efficient network for autonomous driving,.

Large Language Models for Crash Detection in Video: A Survey of Methods, Datasets, and Challenges Sutd-trafficqa: A question answering benchmark and an efficient network for autonomous driving,

Reference 81

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:43:11.302437Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T20:43:06.436310Z digest=sha256:24166696cb2a9002b2bb1c3034e437ba15e6a591be1b1685b3cb935c83c6689f

Observation f1022a6c-eaa7-465b-98f8-9d7c89144724 · outbound

This paper cites Causallp: Visual causal question answering with knowledge graphs,.

Large Language Models for Crash Detection in Video: A Survey of Methods, Datasets, and Challenges Causallp: Visual causal question answering with knowledge graphs,

Reference 82

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:43:11.171794Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T20:43:06.486551Z digest=sha256:52f42bae0d7977c0fb06adfca8243733ad1716152f32a13f77d7ad4f4fc918c9

Observation f5e412ca-fdd3-4dc9-9d33-1c1f18023792 · outbound

This paper cites CausalChaos! Dataset for Comprehensive Causal Action Question Answering Over Longer Causal Chains Grounded in Dynamic Visual Scenes.

Large Language Models for Crash Detection in Video: A Survey of Methods, Datasets, and Challenges CausalChaos! Dataset for Comprehensive Causal Action Question Answering Over Longer Causal Chains Grounded in Dynamic Visual Scenes

Reference 83

Resolution
verified exact
local_arxiv, observed 2026-08-06T20:43:08.600662Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T20:43:06.555774Z digest=sha256:774b97efb918d5d874e419b57142f7bcbfc1092665b342e1976746cc4f68bfed

Observation 046aa1f3-2440-482a-965b-f55308a449ac · outbound

This paper cites Bolstering causal reasoning in video question answering with multi-event causal discovery,.

Large Language Models for Crash Detection in Video: A Survey of Methods, Datasets, and Challenges Bolstering causal reasoning in video question answering with multi-event causal discovery,

Reference 84

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:43:11.013528Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T20:43:06.649287Z digest=sha256:9a7be7652bc80c2ecc804295814ccbdc8516caf6b99d47ff0a074fa69e80290b

Observation 5e2b02e0-e7f2-42f7-ab20-4f32a1d7f23b · outbound

This paper cites Holmes-VAD: Towards Unbiased and Explainable Video Anomaly Detection via Multi-modal LLM.

Large Language Models for Crash Detection in Video: A Survey of Methods, Datasets, and Challenges Holmes-VAD: Towards Unbiased and Explainable Video Anomaly Detection via Multi-modal LLM

Reference 85

Resolution
unresolved
no resolver link, observed 2026-08-06T20:43:06.699602Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:43:06.699602Z digest=sha256:3f14a3a37aa5e34c0b7edbd5e44260059e12ddae515f6fd4ad6850970973a126

Observation fa4a90d0-3923-4b4d-9ace-6a6bdbaffac6 · outbound

This paper cites Crash time matters: Hybridmamba for fine-grained temporal localization in traffic surveillance footage,.

Large Language Models for Crash Detection in Video: A Survey of Methods, Datasets, and Challenges Crash time matters: Hybridmamba for fine-grained temporal localization in traffic surveillance footage,

Reference 86

Resolution
verified exact
raw_fallback, observed 2026-08-06T20:43:08.409542Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T20:43:06.823875Z digest=sha256:74712d250bdfbb422a1395eb20eca30a8c03c640e3460fea69b97fbd75cb9f9d

Observation de6cdf92-6b1d-4796-9fd7-13e3091f4760 · outbound

This paper cites Deep Learning for Video Anomaly Detection: A Review.

Large Language Models for Crash Detection in Video: A Survey of Methods, Datasets, and Challenges Deep Learning for Video Anomaly Detection: A Review

Reference 87

Resolution
unresolved
no resolver link, observed 2026-08-06T20:43:06.875392Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:43:06.875392Z digest=sha256:fbee458bc0daca1aec86d4481ddd6c6f60f1da906142abf886c62758fe71393a

Observation 0b092ca4-8f8d-4826-b6ab-28e387c2b98d · outbound

This paper cites Edge-Based Video Analytics: A Survey.

Large Language Models for Crash Detection in Video: A Survey of Methods, Datasets, and Challenges Edge-Based Video Analytics: A Survey

Reference 88

Resolution
verified exact
local_arxiv, observed 2026-08-06T20:43:08.050272Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T20:43:06.975537Z digest=sha256:e41f0845d72184b7e593ce431d0ddf8af3ac4cac8eba6f38c52978a056658a5f

Observation 717e465f-ae8d-4ec1-b57d-2bf8175191c6 · outbound

This paper cites Synthetic datasets for autonomous driving: IEEE TRANSACTIONS ON INTELLIGENT TRANSPORTATION SYSTEMS, VOL. XX, NO. X, MONTH YEAR 24 A survey,.

Large Language Models for Crash Detection in Video: A Survey of Methods, Datasets, and Challenges Synthetic datasets for autonomous driving: IEEE TRANSACTIONS ON INTELLIGENT TRANSPORTATION SYSTEMS, VOL. XX, NO. X, MONTH YEAR 24 A survey,

Reference 89

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:43:10.906404Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T20:43:07.129418Z digest=sha256:cb3cb23d5060696ceecde27fb3bf2b56ba5953f676548ba093e8d26e8a2a7ea2

Observation b0c01428-eb14-473d-abfa-5624e1ff7f9c · outbound

This paper cites Domain general- ization through meta-learning: a survey,.

Large Language Models for Crash Detection in Video: A Survey of Methods, Datasets, and Challenges Domain general- ization through meta-learning: a survey,

Reference 90

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:43:10.802463Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T20:43:07.219424Z digest=sha256:087ee413dad38a2ca84622cef45116e67ca3ac45d8d8f8075d9e1f9dda0ae39f

Observation d82ac71d-4534-4271-872a-011aa335da6f · outbound

This paper cites Generalized Out-of-Distribution Detection: A Survey.

Large Language Models for Crash Detection in Video: A Survey of Methods, Datasets, and Challenges Generalized Out-of-Distribution Detection: A Survey

Reference 91

Resolution
unresolved
no resolver link, observed 2026-08-06T20:43:07.278200Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:43:07.278200Z digest=sha256:c2c1b1b3c3e64b654428f424d597e1ebafbe59b42dca378f3d42b0d2d688be8c

Observation 8303e124-e7be-4be8-b8b6-ea50d8515671 · outbound

This paper cites Carla: An open urban driving simulator,.

Large Language Models for Crash Detection in Video: A Survey of Methods, Datasets, and Challenges Carla: An open urban driving simulator,

Reference 92

Resolution
unresolved
no resolver link, observed 2026-08-06T20:43:07.353734Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:43:07.353734Z digest=sha256:c56f4e84319aa40b4ec324e48f4178707291c1bc4cf7536c50f905afe3cbc9af

Observation 703f5a5a-6ad8-41be-8d29-cfa24306dade · outbound

This paper cites Perception, planning, control, and coordination for autonomous vehicles,.

Large Language Models for Crash Detection in Video: A Survey of Methods, Datasets, and Challenges Perception, planning, control, and coordination for autonomous vehicles,

Reference 93

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:43:10.701373Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T20:43:07.464300Z digest=sha256:9c8a89df66b6c29f5aa6e1d2726dac24d021b93f22334e63f5e5e83c995c6486

Observation a8c7cc71-d1b7-48aa-930d-07d82ab268c5 · outbound

This paper cites A survey of the multi-sensor fusion object detection task in autonomous driving,.

Large Language Models for Crash Detection in Video: A Survey of Methods, Datasets, and Challenges A survey of the multi-sensor fusion object detection task in autonomous driving,

Reference 94

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:43:10.573511Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T20:43:07.581462Z digest=sha256:9badae88192656e19966cbc25c2976275346ada5afdeda0d97cccaa66b980519

Observation c44dce93-6c64-4a7e-b261-b9159e25900e · outbound

This paper cites A Survey on Efficient Vision-Language Models.

Large Language Models for Crash Detection in Video: A Survey of Methods, Datasets, and Challenges A Survey on Efficient Vision-Language Models

Reference 95

Resolution
unresolved
no resolver link, observed 2026-08-06T20:43:07.662405Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:43:07.662405Z digest=sha256:7050f3d6a15135f42ecbc5a412612715dc7d52f22bc2ffbb6b1b0d62dae8b9e8

Observation e3e4edb3-e917-46a5-bf44-7a4e139f5ef5 · outbound

This paper cites Privacy-Preserving Video Anomaly Detection: A Survey.

Large Language Models for Crash Detection in Video: A Survey of Methods, Datasets, and Challenges Privacy-Preserving Video Anomaly Detection: A Survey

Reference 96

Resolution
unresolved
no resolver link, observed 2026-08-06T20:43:07.779736Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:43:07.779736Z digest=sha256:60c8deae4fac9f590197609a9cc0bf03b7be0ba5c2ddf792a1453af0acd74339

Observation 2ff580c3-6c20-4f3b-beec-639e3d156ff0 · outbound

This paper cites CRASH: Crash Recognition and Anticipation System Harnessing with Context-Aware and Temporal Focus Attentions.

Large Language Models for Crash Detection in Video: A Survey of Methods, Datasets, and Challenges CRASH: Crash Recognition and Anticipation System Harnessing with Context-Aware and Temporal Focus Attentions

Reference 2024

Resolution
metadata mismatch
local_arxiv, observed 2026-08-06T20:43:10.363170Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T20:43:04.229060Z digest=sha256:64fb7a4f8017899e0864a8fc514601aafb3e170cee2faace9896b9b545643b1e

Pith citing papers

Observation 290e8f37-2e24-426b-9fd2-ef980f75c9ae · inbound

Automating Crash Diagram Generation Using Vision-Language Models: A Case Study on Multi-Lane Roundabouts cites this paper.

Automating Crash Diagram Generation Using Vision-Language Models: A Case Study on Multi-Lane Roundabouts Large Language Models for Crash Detection in Video: A Survey of Methods, Datasets, and Challenges

Reference 9

Resolution
verified exact
arxiv_id, observed 2026-05-15T14:50:04.412900Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-15T14:46:41.877762Z digest=sha256:1ab79556d037ca7218edad81629ba639970c1c7094f5c1cd9ba8a9749199f61a

Observation eaa701fc-b87b-4e5a-b9bf-14c87c2f02c3 · inbound

TrafficRAG: A Multimodal RAG Framework for Traffic Accident Liability Determination cites this paper.

TrafficRAG: A Multimodal RAG Framework for Traffic Accident Liability Determination Large Language Models for Crash Detection in Video: A Survey of Methods, Datasets, and Challenges

Reference 3

Resolution
metadata mismatch
arxiv_id, observed 2026-07-01T23:06:20.629333Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-06-28T14:39:27.242094Z digest=sha256:f59d4d1147306ec9339358d61a2d53460a19656b5f419e9f5f1892ef5d4ebab0