Pith. sign in

Paper Citation Record · LEDGER

Weakly Supervised Temporal Action Localization via Dual-Prior Collaborative Learning Guided by Multimodal Large Language Models

As of 21 August 2026, this Paper Citation Record lists 49 of 49 outbound references and 6 inbound Pith citation observations for arXiv:2411.08466.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2411.08466 v2

Coverage vector

measured 49 of 49 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-12T21:36:11.822339Z

measured 55 of 55 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-20T06:33:59.587034+00:00

measured 6 of 6 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-10T18:41:46.263737Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-08-06T23:35:05.394191Z

Reference resolution

49 of 49 outbound references displayed

  • verified exact0
  • verified fuzzy34
  • unresolved15
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 97085144-945e-4d1c-9114-61480fd36d05 · outbound

This paper cites an unresolved cited work.

Weakly Supervised Temporal Action Localization via Dual-Prior Collaborative Learning Guided by Multimodal Large Language Models Unresolved cited work

Reference 1

Resolution
unresolved
raw_fallback, observed 2026-08-12T21:36:12.701361Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-12T21:36:11.558097Z digest=sha256:7105c82cd0abc48cb96be2f148a2cce6f628ad08d7964065694e4dc24266a4f6

Observation 261affba-45d6-4c7d-a464-6ff2955a3d2e · outbound

This paper cites Flamingo: a visual language model for few-shot learning.

Weakly Supervised Temporal Action Localization via Dual-Prior Collaborative Learning Guided by Multimodal Large Language Models Flamingo: a visual language model for few-shot learning

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-12T21:36:11.564654Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T21:36:11.564654Z digest=sha256:6277cc7c36daa10fd2dc352a893853c256966459ae2ef1cb8ede16accb88456c

Observation 38334622-3801-429a-8883-069d586db59c · outbound

This paper cites Boundary content graph neural network for temporal action proposal generation.

Weakly Supervised Temporal Action Localization via Dual-Prior Collaborative Learning Guided by Multimodal Large Language Models Boundary content graph neural network for temporal action proposal generation

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T21:36:12.673530Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-12T21:36:11.570633Z digest=sha256:4e9221441b12c25c177f4d61e3a9e615beb8b8bbfbe068a1d9e4326fd24f6d4d

Observation 09336ad7-3c97-48fb-ae14-e9c972815aed · outbound

This paper cites Lan- guage models are few-shot learners.

Weakly Supervised Temporal Action Localization via Dual-Prior Collaborative Learning Guided by Multimodal Large Language Models Lan- guage models are few-shot learners

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-12T21:36:11.576291Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T21:36:11.576291Z digest=sha256:415354f397d7b8d02d32fcf799cc7fe4ad039b82cd016551e49c3322d4c36f20

Observation 36779a57-94f6-4842-9b6a-eb4cb1d12b87 · outbound

This paper cites Activitynet: A large-scale video benchmark for human activity understanding.

Weakly Supervised Temporal Action Localization via Dual-Prior Collaborative Learning Guided by Multimodal Large Language Models Activitynet: A large-scale video benchmark for human activity understanding

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-12T21:36:11.582375Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T21:36:11.582375Z digest=sha256:8eda4213c576769e8e6645d387df40e08847ff66582d9b24d6aa55803a6fa7d4

Observation 600d8cc2-c59a-4ac2-a54c-1a88469f2027 · outbound

This paper cites Ross, Jia Deng, and Rahul Sukthankar.

Weakly Supervised Temporal Action Localization via Dual-Prior Collaborative Learning Guided by Multimodal Large Language Models Ross, Jia Deng, and Rahul Sukthankar

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T21:36:12.636096Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-12T21:36:11.588189Z digest=sha256:5868524c856c5c3b1349c88f7e090285c0110f38ca4b6c7e06a844f51a6180a5

Observation 39eaff1e-1cf9-4801-953a-17a6821f4a71 · outbound

This paper cites Dual-evidential learning for weakly-supervised tempo- ral action localization.

Weakly Supervised Temporal Action Localization via Dual-Prior Collaborative Learning Guided by Multimodal Large Language Models Dual-evidential learning for weakly-supervised tempo- ral action localization

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T21:36:12.619954Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-12T21:36:11.593446Z digest=sha256:d8dfbf9f3c701026b3a233d1a085eea826321eee95a52247c19300130aa00149

Observation 86716db9-bca9-4a6e-a39c-32888fd56e20 · outbound

This paper cites GIM: A Million-scale Benchmark for Generative Image Manipulation Detection and Localization.

Weakly Supervised Temporal Action Localization via Dual-Prior Collaborative Learning Guided by Multimodal Large Language Models GIM: A Million-scale Benchmark for Generative Image Manipulation Detection and Localization

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-12T21:36:11.599192Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T21:36:11.599192Z digest=sha256:4111fce3015f7297c165d8e3be3401026fac5f6937cf0a5251d9371d5b91c5c4

Observation 20deaa4e-b21d-4519-a17a-389b31bc4c32 · outbound

This paper cites Vicuna: An open-source chatbot impressing gpt-4 with 90%* chatgpt quality.

Weakly Supervised Temporal Action Localization via Dual-Prior Collaborative Learning Guided by Multimodal Large Language Models Vicuna: An open-source chatbot impressing gpt-4 with 90%* chatgpt quality

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T21:36:12.603245Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-12T21:36:11.604874Z digest=sha256:20d1443170907f198f7547ddd8dc8b12dc10eb8825f32774da971610af54ed2e

Observation 8a4c02c5-c110-49fb-9a0e-b6932ae9c350 · outbound

This paper cites Imagebind: One embedding space to bind them all.

Weakly Supervised Temporal Action Localization via Dual-Prior Collaborative Learning Guided by Multimodal Large Language Models Imagebind: One embedding space to bind them all

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-12T21:36:11.610123Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T21:36:11.610123Z digest=sha256:20e8f3f5ee150e5e453f2f6d4bd90ae7e93a11b03dee04661507b7e039856cf8

Observation 1d45a2f4-09e0-433e-9012-465930d6bdd2 · outbound

This paper cites Cross-modal consensus network for weakly supervised temporal action localization.

Weakly Supervised Temporal Action Localization via Dual-Prior Collaborative Learning Guided by Multimodal Large Language Models Cross-modal consensus network for weakly supervised temporal action localization

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T21:36:12.576722Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-12T21:36:11.615303Z digest=sha256:6148fc7e1cda9dcbd9704682a7ddd46799b7ab4505da13b796a4cffb4cb06fcc

Observation 4db7b917-08a1-49c1-a973-da2380f8982d · outbound

This paper cites in the wild.

Weakly Supervised Temporal Action Localization via Dual-Prior Collaborative Learning Guided by Multimodal Large Language Models in the wild

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T21:36:12.559970Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-12T21:36:11.620447Z digest=sha256:eba8bde306576bb732f9eaa73919949a269dc27de3e490735a8f32448a8d87ed

Observation b17a1e44-be00-404a-b803-fb4244765137 · outbound

This paper cites A hybrid attention mechanism for weakly-supervised temporal action localization.

Weakly Supervised Temporal Action Localization via Dual-Prior Collaborative Learning Guided by Multimodal Large Language Models A hybrid attention mechanism for weakly-supervised temporal action localization

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T21:36:12.542981Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-12T21:36:11.625374Z digest=sha256:db180284b2ffe352f906fdd46f846905743a270ad6f1663e6331429fbb7b389c

Observation 1a5ea49f-278e-4fd8-9397-0d422a9a5af1 · outbound

This paper cites Distill- ing vision-language pre-training to collaborate with weakly- supervised temporal action localization.

Weakly Supervised Temporal Action Localization via Dual-Prior Collaborative Learning Guided by Multimodal Large Language Models Distill- ing vision-language pre-training to collaborate with weakly- supervised temporal action localization

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T21:36:12.525359Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-12T21:36:11.630650Z digest=sha256:48d5caa6a8cc725f1a81933400e4b9339048db0cf4d5332575a511e765e61642

Observation bcc98d08-6822-4593-825b-9bc4ab4a172b · outbound

This paper cites Weakly-supervised temporal action localization by uncer- tainty modeling.

Weakly Supervised Temporal Action Localization via Dual-Prior Collaborative Learning Guided by Multimodal Large Language Models Weakly-supervised temporal action localization by uncer- tainty modeling

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T21:36:12.506597Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-12T21:36:11.635969Z digest=sha256:768d71d34f9dcf64164e3e0ded8aedb58315e260bec3a9e7040658ba2af80646

Observation bc4f8965-66ec-469e-a4ac-e83353684fac · outbound

This paper cites Otter: A Multi-Modal Model with In-Context Instruction Tuning.

Weakly Supervised Temporal Action Localization via Dual-Prior Collaborative Learning Guided by Multimodal Large Language Models Otter: A Multi-Modal Model with In-Context Instruction Tuning

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-12T21:36:11.641219Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T21:36:11.641219Z digest=sha256:5831d1da2685a47c3e8cb4d60d447fe70fd8564491cf4b56a366e3e92ab56849

Observation da9dbb40-df71-4b10-863d-1ff79759fef7 · outbound

This paper cites Boosting weakly-supervised temporal action localization with text information.

Weakly Supervised Temporal Action Localization via Dual-Prior Collaborative Learning Guided by Multimodal Large Language Models Boosting weakly-supervised temporal action localization with text information

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T21:36:12.489870Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-12T21:36:11.646549Z digest=sha256:43f542adda93a20dd7b519873e63c5322af164ec74a4f20c6103469adff5e0ea

Observation d0ca7528-3d02-4a9f-92d1-4a5850410aab · outbound

This paper cites Exploring denoised cross-video contrast for weakly- supervised temporal action localization.

Weakly Supervised Temporal Action Localization via Dual-Prior Collaborative Learning Guided by Multimodal Large Language Models Exploring denoised cross-video contrast for weakly- supervised temporal action localization

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T21:36:12.471770Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-12T21:36:11.652359Z digest=sha256:088ce6ce526eaf45be2eaa9236aad9d3ae5e984380f5406440cbb226076fc398

Observation 3831a3f9-736f-4389-a001-9eb0cdca190e · outbound

This paper cites BLIP-2: Bootstrapping Language-Image Pre-training with Frozen Image Encoders and Large Language Models.

Weakly Supervised Temporal Action Localization via Dual-Prior Collaborative Learning Guided by Multimodal Large Language Models BLIP-2: Bootstrapping Language-Image Pre-training with Frozen Image Encoders and Large Language Models

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-12T21:36:11.657110Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T21:36:11.657110Z digest=sha256:249063617941bfa5b9f29c85ce9c42ab66b2aaf6c6dfe02b0d947f48218f710d

Observation d3735fa2-fcab-41b4-8dc6-b610c4e096e3 · outbound

This paper cites Revisiting foreground and background sepa- ration in weakly-supervised temporal action localization: A clustering-based approach.

Weakly Supervised Temporal Action Localization via Dual-Prior Collaborative Learning Guided by Multimodal Large Language Models Revisiting foreground and background sepa- ration in weakly-supervised temporal action localization: A clustering-based approach

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T21:36:12.452071Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-12T21:36:11.662227Z digest=sha256:764b83edb697da8d9eed667704815fae8dada4852b0f2392231964ba6a7f0a5a

Observation 0e568180-5dce-46e2-b60c-d930976fead8 · outbound

This paper cites Pivotal: Prior-driven supervision for weakly-supervised tem- poral action localization.

Weakly Supervised Temporal Action Localization via Dual-Prior Collaborative Learning Guided by Multimodal Large Language Models Pivotal: Prior-driven supervision for weakly-supervised tem- poral action localization

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T21:36:12.433741Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-12T21:36:11.667684Z digest=sha256:b9661357ff072f3fd9a96b9adce1c1196725e354a92f5de8da7c612e182437dd

Observation ac17e594-17ca-45c6-93ce-91f7937f0051 · outbound

This paper cites Weakly supervised action localization by sparse temporal pooling network.

Weakly Supervised Temporal Action Localization via Dual-Prior Collaborative Learning Guided by Multimodal Large Language Models Weakly supervised action localization by sparse temporal pooling network

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T21:36:12.410662Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-12T21:36:11.673029Z digest=sha256:40e994292c09d32dba2908762b16fef7060d1b1348f9a85caf619c385af4c85c

Observation 52fb9456-c874-41e7-997b-e4be013c4002 · outbound

This paper cites GPT-4 Technical Report.

Weakly Supervised Temporal Action Localization via Dual-Prior Collaborative Learning Guided by Multimodal Large Language Models GPT-4 Technical Report

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T21:36:12.388809Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-12T21:36:11.678221Z digest=sha256:c38d9872bcae16728dd1a45066e6cf869a1137cba25cd7139bd128a2b8109407

Observation 2ac5cd9a-1fda-4223-9a70-a8d8336a720b · outbound

This paper cites W- talc: Weakly-supervised temporal activity localization and classification.

Weakly Supervised Temporal Action Localization via Dual-Prior Collaborative Learning Guided by Multimodal Large Language Models W- talc: Weakly-supervised temporal activity localization and classification

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T21:36:12.368915Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-12T21:36:11.683356Z digest=sha256:e667d6c49517d047995d6d92687a2cb104a3a60803c70b43ea23f0a76b3c380a

Observation 122c751a-a5cc-4fde-97d2-cace70bf17a2 · outbound

This paper cites Glove: Global vectors for word representation.

Weakly Supervised Temporal Action Localization via Dual-Prior Collaborative Learning Guided by Multimodal Large Language Models Glove: Global vectors for word representation

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-12T21:36:11.688551Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T21:36:11.688551Z digest=sha256:34b2b8feb51e8886f6cf3608cd37552f920bb598367a4e2ee37840945f724bc9

Observation d54afbb7-1d90-4d32-8e50-a14c043255c1 · outbound

This paper cites Proposal-based multiple instance learning for weakly-supervised temporal action localization.

Weakly Supervised Temporal Action Localization via Dual-Prior Collaborative Learning Guided by Multimodal Large Language Models Proposal-based multiple instance learning for weakly-supervised temporal action localization

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T21:36:12.331497Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-12T21:36:11.693613Z digest=sha256:08b0dda7b8fec0c0c0481417c70585ccc3396594c5689f0c515a156b07eb73c0

Observation 5fc1b78f-ec68-474d-91cb-c97c1e7414f7 · outbound

This paper cites Dynamic graph modeling for weakly-supervised temporal action localization.

Weakly Supervised Temporal Action Localization via Dual-Prior Collaborative Learning Guided by Multimodal Large Language Models Dynamic graph modeling for weakly-supervised temporal action localization

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T21:36:12.312279Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-12T21:36:11.701763Z digest=sha256:733438b09ad6b72025d61326f415661ac1fa39bd032364d0a95750a174aa12f2

Observation 4d5875be-b636-47dd-9302-bcfec9d5cc99 · outbound

This paper cites Ddg- net: Discriminability-driven graph network for weakly- supervised temporal action localization.

Weakly Supervised Temporal Action Localization via Dual-Prior Collaborative Learning Guided by Multimodal Large Language Models Ddg- net: Discriminability-driven graph network for weakly- supervised temporal action localization

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T21:36:12.292693Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-12T21:36:11.708125Z digest=sha256:2c5d22fafc585d0ae2121e22a69f4582f5c19fe90de2b32e45ac7395383379a4

Observation 8c245d6f-d29c-41fb-bcc3-e5165c5a259a · outbound

This paper cites Llama 2: Open Foundation and Fine-Tuned Chat Models.

Weakly Supervised Temporal Action Localization via Dual-Prior Collaborative Learning Guided by Multimodal Large Language Models Llama 2: Open Foundation and Fine-Tuned Chat Models

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-12T21:36:11.713590Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T21:36:11.713590Z digest=sha256:6160f4008cab16c92f5c25b005858bd300c1ef67a5cd141636d3d7df2181355d

Observation bf17ea4f-44ce-4617-b95e-5b76be27a0d3 · outbound

This paper cites ChatVideo: A Tracklet-centric Multimodal and Versatile Video Understanding System.

Weakly Supervised Temporal Action Localization via Dual-Prior Collaborative Learning Guided by Multimodal Large Language Models ChatVideo: A Tracklet-centric Multimodal and Versatile Video Understanding System

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-12T21:36:11.718934Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T21:36:11.718934Z digest=sha256:e2a14c86a0a328e33f96a7f45cbb69348b4064dc977721e7f5e8320a548c43f1

Observation 447632aa-23a4-4aca-9ad1-322b9f8c00b9 · outbound

This paper cites Untrimmednets for weakly supervised action recognition and detection.

Weakly Supervised Temporal Action Localization via Dual-Prior Collaborative Learning Guided by Multimodal Large Language Models Untrimmednets for weakly supervised action recognition and detection

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T21:36:12.273423Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-12T21:36:11.725188Z digest=sha256:3a0b6997a1efb8811138054ac4b73e21ad4df0c23307f19f47fd72c4434ee2ad

Observation d2fb7309-8681-4875-ba20-01c17a89716d · outbound

This paper cites Two-stream net- works for weakly-supervised temporal action localization with semantic-aware mechanisms.

Weakly Supervised Temporal Action Localization via Dual-Prior Collaborative Learning Guided by Multimodal Large Language Models Two-stream net- works for weakly-supervised temporal action localization with semantic-aware mechanisms

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T21:36:12.255394Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-12T21:36:11.731447Z digest=sha256:5f8933fda414e5284a0bfbc93ad6503c1410cbeab065f2da6b59314d3dd1b9e9

Observation f2c5a812-8387-4bf7-a140-1ced0dc93316 · outbound

This paper cites Task groupings regulariza- tion: Data-free meta-learning with heterogeneous pre-trained models.

Weakly Supervised Temporal Action Localization via Dual-Prior Collaborative Learning Guided by Multimodal Large Language Models Task groupings regulariza- tion: Data-free meta-learning with heterogeneous pre-trained models

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T21:36:12.237316Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-12T21:36:11.737018Z digest=sha256:b37b597533f73d5d21557ab0694784a52f6d9c9cf28d8eaa98eec0864bc7d6f3

Observation 56ebf1f4-8c3e-42b9-a47e-99c3b4e58fe4 · outbound

This paper cites Free: Faster and better data-free meta-learning.

Weakly Supervised Temporal Action Localization via Dual-Prior Collaborative Learning Guided by Multimodal Large Language Models Free: Faster and better data-free meta-learning

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T21:36:12.220371Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-12T21:36:11.742306Z digest=sha256:e2fb10a2126e2dc89b1f85d94857784815cfcf7f3f6ab4f042fe393e7ee6a561

Observation 036d4442-99ef-4ce7-b05d-10dbac69da3b · outbound

This paper cites Modeling Multi-Task Model Merging as Adaptive Projective Gradient Descent.

Weakly Supervised Temporal Action Localization via Dual-Prior Collaborative Learning Guided by Multimodal Large Language Models Modeling Multi-Task Model Merging as Adaptive Projective Gradient Descent

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-12T21:36:11.747711Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T21:36:11.747711Z digest=sha256:e12be5d1bb6dcdde78e407b5f99cc8800485751f34a93d9d7890fa4778a910d7

Observation 8dbc99b1-c9b6-4c0a-8c9b-7afe697693fe · outbound

This paper cites G-tad: Sub-graph localization for tempo- ral action detection.

Weakly Supervised Temporal Action Localization via Dual-Prior Collaborative Learning Guided by Multimodal Large Language Models G-tad: Sub-graph localization for tempo- ral action detection

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T21:36:12.204095Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-12T21:36:11.752928Z digest=sha256:44e2de739d56adf2bfe6afdd09ff376b1890f055777493e60a5cfb513667e830

Observation d95f7a06-1112-4c7b-91bb-d767273b5e8e · outbound

This paper cites Uncertainty guided collaborative training for weakly supervised temporal action detection.

Weakly Supervised Temporal Action Localization via Dual-Prior Collaborative Learning Guided by Multimodal Large Language Models Uncertainty guided collaborative training for weakly supervised temporal action detection

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T21:36:12.187262Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-12T21:36:11.758271Z digest=sha256:131a8c90ad659409973cb97545706d276fb57304e09f3617cc160e8ca38aac56

Observation c305d2a5-eeec-4385-9b24-2ee6712a2e21 · outbound

This paper cites Weakly-supervised temporal action localization by in- ferring salient snippet-feature.

Weakly Supervised Temporal Action Localization via Dual-Prior Collaborative Learning Guided by Multimodal Large Language Models Weakly-supervised temporal action localization by in- ferring salient snippet-feature

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T21:36:12.170720Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-12T21:36:11.763190Z digest=sha256:58658d4b13549cbdb94d64c1ac983ae03d80f8f8a489590f878a211b87555018

Observation 814cf87b-15ff-4084-9b97-a233af6b960e · outbound

This paper cites Graph con- volutional networks for temporal action localization.

Weakly Supervised Temporal Action Localization via Dual-Prior Collaborative Learning Guided by Multimodal Large Language Models Graph con- volutional networks for temporal action localization

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T21:36:12.153636Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-12T21:36:11.768607Z digest=sha256:792aabe65974580bb00be39a7890d56fc91efa4bc471300704761b0913fc01aa

Observation a1ff3c2c-7018-45bb-a709-2818e5f398b3 · outbound

This paper cites Two-stream consensus network for weakly-supervised temporal action localization.

Weakly Supervised Temporal Action Localization via Dual-Prior Collaborative Learning Guided by Multimodal Large Language Models Two-stream consensus network for weakly-supervised temporal action localization

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T21:36:12.137071Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-12T21:36:11.773598Z digest=sha256:1d1ac0f99b1cff4534be8e62fada954ac1daac005a86600bd5d603aa0fbc2a94

Observation 55af6a6f-961e-4397-af5d-5c24577db356 · outbound

This paper cites Cola: Weakly-supervised temporal action lo- calization with snippet contrastive learning.

Weakly Supervised Temporal Action Localization via Dual-Prior Collaborative Learning Guided by Multimodal Large Language Models Cola: Weakly-supervised temporal action lo- calization with snippet contrastive learning

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T21:36:12.117654Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-12T21:36:11.779695Z digest=sha256:44c865893c7a77079aa2bc0cb1fa8e2f860ca05d83d672f71840bec958c54e0f

Observation 28b6fad0-024c-4177-be9a-da63dfcbf2f0 · outbound

This paper cites Video-LLaMA: An Instruction-tuned Audio-Visual Language Model for Video Understanding.

Weakly Supervised Temporal Action Localization via Dual-Prior Collaborative Learning Guided by Multimodal Large Language Models Video-LLaMA: An Instruction-tuned Audio-Visual Language Model for Video Understanding

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-12T21:36:11.784885Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T21:36:11.784885Z digest=sha256:0dc5425a3e574e6859bf9724c4c3f608c4e1a0752d6171caee07f3064a20661b

Observation dcc7179f-a301-4d6b-b156-3b52753db662 · outbound

This paper cites Distilling semantic priors from sam to efficient image restoration models.

Weakly Supervised Temporal Action Localization via Dual-Prior Collaborative Learning Guided by Multimodal Large Language Models Distilling semantic priors from sam to efficient image restoration models

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T21:36:12.098688Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-12T21:36:11.790123Z digest=sha256:07b91cf01e351afb26e36f8851813ebbe376d1eded53fe9da4804eaa8f116527

Observation bf015b08-6896-48a0-a4d3-683d187819c0 · outbound

This paper cites IMDPrompter: Adapting SAM to image manipulation detection by cross-view automated prompt learning.

Weakly Supervised Temporal Action Localization via Dual-Prior Collaborative Learning Guided by Multimodal Large Language Models IMDPrompter: Adapting SAM to image manipulation detection by cross-view automated prompt learning

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T21:36:12.081702Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-12T21:36:11.795002Z digest=sha256:5548f8a17838f3ce39f693e265ff4a224f560081143a965cc6c818d022dc4fa0

Observation c340076f-f02f-47af-8969-591674eb56e7 · outbound

This paper cites Rethinking Pseudo-Label Guided Learning for Weakly Supervised Temporal Action Localization from the Perspective of Noise Correction.

Weakly Supervised Temporal Action Localization via Dual-Prior Collaborative Learning Guided by Multimodal Large Language Models Rethinking Pseudo-Label Guided Learning for Weakly Supervised Temporal Action Localization from the Perspective of Noise Correction

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-12T21:36:11.800447Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T21:36:11.800447Z digest=sha256:a19a0d0d1302b68f7e3422bf1676db71f56a71ac2ea54c7b1c7ed40b4f4b2725

Observation 240a1cd1-2724-49ab-bbb3-5ffb4d8bd26a · outbound

This paper cites Equivalent classification mapping for weakly supervised temporal action localization.

Weakly Supervised Temporal Action Localization via Dual-Prior Collaborative Learning Guided by Multimodal Large Language Models Equivalent classification mapping for weakly supervised temporal action localization

Reference 46

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T21:36:12.063208Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-12T21:36:11.807573Z digest=sha256:e85549e1eba14d829e707987b2ac1f0cf3e79b22e650316c282f3ee7f7538ad9

Observation a5af33e2-847c-46b9-b007-6d250328252d · outbound

This paper cites Temporal action detection with structured segment networks.

Weakly Supervised Temporal Action Localization via Dual-Prior Collaborative Learning Guided by Multimodal Large Language Models Temporal action detection with structured segment networks

Reference 47

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T21:36:12.046226Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-12T21:36:11.812636Z digest=sha256:1c78fdfe6ea8b9d000de385bb3ba938d9bdf83e74870ce8880ca34db843f16f5

Observation 81c3e45f-dca6-41b3-bff4-ac02220ac7c6 · outbound

This paper cites Improving weakly supervised temporal ac- tion localization by bridging train-test gap in pseudo labels.

Weakly Supervised Temporal Action Localization via Dual-Prior Collaborative Learning Guided by Multimodal Large Language Models Improving weakly supervised temporal ac- tion localization by bridging train-test gap in pseudo labels

Reference 48

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T21:36:12.028875Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-12T21:36:11.817668Z digest=sha256:18603ca5560f14f566a7f0c7e9d4f88a35722aa79d021792144956bfa0719bca

Observation c12398e9-0c05-4d2b-ae6b-78b3c5635b34 · outbound

This paper cites MiniGPT-4: Enhancing Vision-Language Understanding with Advanced Large Language Models.

Weakly Supervised Temporal Action Localization via Dual-Prior Collaborative Learning Guided by Multimodal Large Language Models MiniGPT-4: Enhancing Vision-Language Understanding with Advanced Large Language Models

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-12T21:36:11.822339Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T21:36:11.822339Z digest=sha256:516450ab209d9c64b778e135ec07a169db99a5a55f8fa99bd29a661b2b02bd8c

Pith citing papers

Observation f56252a4-3235-4d7b-86b6-5dadfd0e446e · inbound

Rethinking Pseudo-Label Guided Learning for Weakly Supervised Temporal Action Localization from the Perspective of Noise Correction cites this paper.

Rethinking Pseudo-Label Guided Learning for Weakly Supervised Temporal Action Localization from the Perspective of Noise Correction Weakly Supervised Temporal Action Localization via Dual-Prior Collaborative Learning Guided by Multimodal Large Language Models

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-10T18:41:46.263737Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T18:41:46.263737Z digest=sha256:8eed5df5eb30f596625991158179918890bb53ad169baa62eaafe2d8cd2ddf5b

Observation 3678d028-2e12-4e16-9437-163810683175 · inbound

IMDPrompter: Adapting SAM to Image Manipulation Detection by Cross-View Automated Prompt Learning cites this paper.

IMDPrompter: Adapting SAM to Image Manipulation Detection by Cross-View Automated Prompt Learning Weakly Supervised Temporal Action Localization via Dual-Prior Collaborative Learning Guided by Multimodal Large Language Models

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-09T12:09:14.625616Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T12:09:14.625616Z digest=sha256:2923c091c53e7c76e25c0a5b94ca9e74ad4c6bac26339d870dcfb2471f71ec45

Observation 7a53f4de-b7b7-40fc-88cd-a4056a624f73 · inbound

$K^2$VAE: A Koopman-Kalman Enhanced Variational AutoEncoder for Probabilistic Time Series Forecasting cites this paper.

$K^2$VAE: A Koopman-Kalman Enhanced Variational AutoEncoder for Probabilistic Time Series Forecasting Weakly Supervised Temporal Action Localization via Dual-Prior Collaborative Learning Guided by Multimodal Large Language Models

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-07T13:02:35.241122Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:02:35.241122Z digest=sha256:b0ba34d43f81e0a9cc0e0e8e51ffcacf52200eb7c4936ef4d0554a1f5dde8c6f

Observation bedf9601-f849-45c1-893d-4433e4ec5471 · inbound

CLIP-AE: CLIP-assisted Cross-view Audio-Visual Enhancement for Unsupervised Temporal Action Localization cites this paper.

CLIP-AE: CLIP-assisted Cross-view Audio-Visual Enhancement for Unsupervised Temporal Action Localization Weakly Supervised Temporal Action Localization via Dual-Prior Collaborative Learning Guided by Multimodal Large Language Models

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-07T12:48:46.650668Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:48:46.650668Z digest=sha256:3ecebf1362c46872253f7d036488e2394dc254dfe93fd3d3b9c8122d161b5e07

Observation a3ed2881-9cd1-469e-bad0-9fa6ec7dcede · inbound

ComRoPE: Scalable and Robust Rotary Position Embedding Parameterized by Trainable Commuting Angle Matrices cites this paper.

ComRoPE: Scalable and Robust Rotary Position Embedding Parameterized by Trainable Commuting Angle Matrices Weakly Supervised Temporal Action Localization via Dual-Prior Collaborative Learning Guided by Multimodal Large Language Models

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-07T11:03:17.468809Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:03:17.468809Z digest=sha256:b47a32a52987f95f68a42aa35655ded451b6da4a53ff2d808f3974c7c441c89f

Observation 25f38f2e-2f51-40c5-92f9-cb4d5f11d566 · inbound

EASE: Embodied Active Event Perception via Self-Supervised Energy Minimization cites this paper.

EASE: Embodied Active Event Perception via Self-Supervised Energy Minimization Weakly Supervised Temporal Action Localization via Dual-Prior Collaborative Learning Guided by Multimodal Large Language Models

Reference 22

Resolution
verified exact
local_arxiv, observed 2026-08-06T23:35:05.494183Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-06T23:35:02.407618Z digest=sha256:5583ba202b78ebb21405b77f4eefb1f577a5a9afead9f8e40f3380fdeab79c6c