Pith. sign in

Paper Citation Record · LEDGER

Weakly Supervised Temporal Action Localization via Dual-Prior Collaborative Learning Guided by Multimodal Large Language Models

As of 21 August 2026, this Paper Citation Record lists 49 of 49 outbound references and 6 inbound Pith citation observations for arXiv:2411.08466.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2411.08466 v2

Coverage vector

measured 49 of 49 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-12T21:36:11.822339Z

measured 55 of 55 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-21T06:32:19.484+00:00

measured 6 of 6 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-10T18:41:46.263737Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-08-06T23:35:05.394191Z

Reference resolution

49 of 49 outbound references displayed

  • verified exact0
  • verified fuzzy34
  • unresolved15
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 97085144-945e-4d1c-9114-61480fd36d05 · outbound

This paper cites an unresolved cited work.

Weakly Supervised Temporal Action Localization via Dual-Prior Collaborative Learning Guided by Multimodal Large Language Models Unresolved cited work

Reference 1

Resolution
unresolved
raw_fallback, observed 2026-08-12T21:36:12.701361Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-12T21:36:11.558097Z digest=sha256:590c990466088f0679c662fc2004cc443e79c90bf929f2f0aa9b036bdc9644ef

Observation 261affba-45d6-4c7d-a464-6ff2955a3d2e · outbound

This paper cites Flamingo: a visual language model for few-shot learning.

Weakly Supervised Temporal Action Localization via Dual-Prior Collaborative Learning Guided by Multimodal Large Language Models Flamingo: a visual language model for few-shot learning

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-12T21:36:11.564654Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T21:36:11.564654Z digest=sha256:6277cc7c36daa10fd2dc352a893853c256966459ae2ef1cb8ede16accb88456c

Observation 38334622-3801-429a-8883-069d586db59c · outbound

This paper cites Boundary content graph neural network for temporal action proposal generation.

Weakly Supervised Temporal Action Localization via Dual-Prior Collaborative Learning Guided by Multimodal Large Language Models Boundary content graph neural network for temporal action proposal generation

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T21:36:12.673530Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-12T21:36:11.570633Z digest=sha256:a266cb4aa943c7229440cbd3529069ca5be175888e14063497c3409959e91ef4

Observation 09336ad7-3c97-48fb-ae14-e9c972815aed · outbound

This paper cites Lan- guage models are few-shot learners.

Weakly Supervised Temporal Action Localization via Dual-Prior Collaborative Learning Guided by Multimodal Large Language Models Lan- guage models are few-shot learners

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-12T21:36:11.576291Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T21:36:11.576291Z digest=sha256:415354f397d7b8d02d32fcf799cc7fe4ad039b82cd016551e49c3322d4c36f20

Observation 36779a57-94f6-4842-9b6a-eb4cb1d12b87 · outbound

This paper cites Activitynet: A large-scale video benchmark for human activity understanding.

Weakly Supervised Temporal Action Localization via Dual-Prior Collaborative Learning Guided by Multimodal Large Language Models Activitynet: A large-scale video benchmark for human activity understanding

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-12T21:36:11.582375Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T21:36:11.582375Z digest=sha256:8eda4213c576769e8e6645d387df40e08847ff66582d9b24d6aa55803a6fa7d4

Observation 600d8cc2-c59a-4ac2-a54c-1a88469f2027 · outbound

This paper cites Ross, Jia Deng, and Rahul Sukthankar.

Weakly Supervised Temporal Action Localization via Dual-Prior Collaborative Learning Guided by Multimodal Large Language Models Ross, Jia Deng, and Rahul Sukthankar

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T21:36:12.636096Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-12T21:36:11.588189Z digest=sha256:0b18bb50c94f4a0eeee297313816f94fb523bacc29fa10dfe0fcc9b1c98e5b88

Observation 39eaff1e-1cf9-4801-953a-17a6821f4a71 · outbound

This paper cites Dual-evidential learning for weakly-supervised tempo- ral action localization.

Weakly Supervised Temporal Action Localization via Dual-Prior Collaborative Learning Guided by Multimodal Large Language Models Dual-evidential learning for weakly-supervised tempo- ral action localization

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T21:36:12.619954Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-12T21:36:11.593446Z digest=sha256:d633c540410bb9edd248ed7ec9772535ce6f900a50256b374b74e3dd052969ad

Observation 86716db9-bca9-4a6e-a39c-32888fd56e20 · outbound

This paper cites GIM: A Million-scale Benchmark for Generative Image Manipulation Detection and Localization.

Weakly Supervised Temporal Action Localization via Dual-Prior Collaborative Learning Guided by Multimodal Large Language Models GIM: A Million-scale Benchmark for Generative Image Manipulation Detection and Localization

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-12T21:36:11.599192Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T21:36:11.599192Z digest=sha256:4111fce3015f7297c165d8e3be3401026fac5f6937cf0a5251d9371d5b91c5c4

Observation 20deaa4e-b21d-4519-a17a-389b31bc4c32 · outbound

This paper cites Vicuna: An open-source chatbot impressing gpt-4 with 90%* chatgpt quality.

Weakly Supervised Temporal Action Localization via Dual-Prior Collaborative Learning Guided by Multimodal Large Language Models Vicuna: An open-source chatbot impressing gpt-4 with 90%* chatgpt quality

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T21:36:12.603245Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-12T21:36:11.604874Z digest=sha256:68d11db948e6dc83158f43392a1b76b0c3b451a868794f5acfea3a3a9cf84e23

Observation 8a4c02c5-c110-49fb-9a0e-b6932ae9c350 · outbound

This paper cites Imagebind: One embedding space to bind them all.

Weakly Supervised Temporal Action Localization via Dual-Prior Collaborative Learning Guided by Multimodal Large Language Models Imagebind: One embedding space to bind them all

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-12T21:36:11.610123Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T21:36:11.610123Z digest=sha256:20e8f3f5ee150e5e453f2f6d4bd90ae7e93a11b03dee04661507b7e039856cf8

Observation 1d45a2f4-09e0-433e-9012-465930d6bdd2 · outbound

This paper cites Cross-modal consensus network for weakly supervised temporal action localization.

Weakly Supervised Temporal Action Localization via Dual-Prior Collaborative Learning Guided by Multimodal Large Language Models Cross-modal consensus network for weakly supervised temporal action localization

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T21:36:12.576722Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-12T21:36:11.615303Z digest=sha256:42d99aa527068c82eb82b31cf98051b509c21e4bfca4a86250a47fefafa53cb7

Observation 4db7b917-08a1-49c1-a973-da2380f8982d · outbound

This paper cites in the wild.

Weakly Supervised Temporal Action Localization via Dual-Prior Collaborative Learning Guided by Multimodal Large Language Models in the wild

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T21:36:12.559970Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-12T21:36:11.620447Z digest=sha256:647159b0f57e0ed805c4e7b809545fc23761dbbf2c4a80914497bcc3dae8c08e

Observation b17a1e44-be00-404a-b803-fb4244765137 · outbound

This paper cites A hybrid attention mechanism for weakly-supervised temporal action localization.

Weakly Supervised Temporal Action Localization via Dual-Prior Collaborative Learning Guided by Multimodal Large Language Models A hybrid attention mechanism for weakly-supervised temporal action localization

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T21:36:12.542981Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-12T21:36:11.625374Z digest=sha256:f8c62e2cdc05caf7ce25e29bc655ef42ecdfceb11773617b96ccffbdf246a174

Observation 1a5ea49f-278e-4fd8-9397-0d422a9a5af1 · outbound

This paper cites Distill- ing vision-language pre-training to collaborate with weakly- supervised temporal action localization.

Weakly Supervised Temporal Action Localization via Dual-Prior Collaborative Learning Guided by Multimodal Large Language Models Distill- ing vision-language pre-training to collaborate with weakly- supervised temporal action localization

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T21:36:12.525359Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-12T21:36:11.630650Z digest=sha256:fe1d560958d63dcc3dd8bdc62367a730364096c0c5914b602ec0749d671297eb

Observation bcc98d08-6822-4593-825b-9bc4ab4a172b · outbound

This paper cites Weakly-supervised temporal action localization by uncer- tainty modeling.

Weakly Supervised Temporal Action Localization via Dual-Prior Collaborative Learning Guided by Multimodal Large Language Models Weakly-supervised temporal action localization by uncer- tainty modeling

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T21:36:12.506597Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-12T21:36:11.635969Z digest=sha256:4bf35b3466ffb9dd11f557610eec9e83df24352d8a99955c9f73d7210f520328

Observation bc4f8965-66ec-469e-a4ac-e83353684fac · outbound

This paper cites Otter: A Multi-Modal Model with In-Context Instruction Tuning.

Weakly Supervised Temporal Action Localization via Dual-Prior Collaborative Learning Guided by Multimodal Large Language Models Otter: A Multi-Modal Model with In-Context Instruction Tuning

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-12T21:36:11.641219Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T21:36:11.641219Z digest=sha256:5831d1da2685a47c3e8cb4d60d447fe70fd8564491cf4b56a366e3e92ab56849

Observation da9dbb40-df71-4b10-863d-1ff79759fef7 · outbound

This paper cites Boosting weakly-supervised temporal action localization with text information.

Weakly Supervised Temporal Action Localization via Dual-Prior Collaborative Learning Guided by Multimodal Large Language Models Boosting weakly-supervised temporal action localization with text information

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T21:36:12.489870Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-12T21:36:11.646549Z digest=sha256:e833ddfdb1fbf68e733b4d44cb300a0e6a133d89ae1e89f2ed96ff4eb540cec3

Observation d0ca7528-3d02-4a9f-92d1-4a5850410aab · outbound

This paper cites Exploring denoised cross-video contrast for weakly- supervised temporal action localization.

Weakly Supervised Temporal Action Localization via Dual-Prior Collaborative Learning Guided by Multimodal Large Language Models Exploring denoised cross-video contrast for weakly- supervised temporal action localization

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T21:36:12.471770Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-12T21:36:11.652359Z digest=sha256:d3712de3c0d45b674de8ce8c8e6dcd55a1667105ba780873d02d6f3c5860d4e8

Observation 3831a3f9-736f-4389-a001-9eb0cdca190e · outbound

This paper cites BLIP-2: Bootstrapping Language-Image Pre-training with Frozen Image Encoders and Large Language Models.

Weakly Supervised Temporal Action Localization via Dual-Prior Collaborative Learning Guided by Multimodal Large Language Models BLIP-2: Bootstrapping Language-Image Pre-training with Frozen Image Encoders and Large Language Models

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-12T21:36:11.657110Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T21:36:11.657110Z digest=sha256:249063617941bfa5b9f29c85ce9c42ab66b2aaf6c6dfe02b0d947f48218f710d

Observation d3735fa2-fcab-41b4-8dc6-b610c4e096e3 · outbound

This paper cites Revisiting foreground and background sepa- ration in weakly-supervised temporal action localization: A clustering-based approach.

Weakly Supervised Temporal Action Localization via Dual-Prior Collaborative Learning Guided by Multimodal Large Language Models Revisiting foreground and background sepa- ration in weakly-supervised temporal action localization: A clustering-based approach

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T21:36:12.452071Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-12T21:36:11.662227Z digest=sha256:9c512b17a6b5a6a509ea76e1ccf850c4ce81525e9a63bb2b6447405d02b99752

Observation 0e568180-5dce-46e2-b60c-d930976fead8 · outbound

This paper cites Pivotal: Prior-driven supervision for weakly-supervised tem- poral action localization.

Weakly Supervised Temporal Action Localization via Dual-Prior Collaborative Learning Guided by Multimodal Large Language Models Pivotal: Prior-driven supervision for weakly-supervised tem- poral action localization

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T21:36:12.433741Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-12T21:36:11.667684Z digest=sha256:765c23c572ef69a2e04eff9138a0f09c08f18a55a8f676e52e58f7b060a3b673

Observation ac17e594-17ca-45c6-93ce-91f7937f0051 · outbound

This paper cites Weakly supervised action localization by sparse temporal pooling network.

Weakly Supervised Temporal Action Localization via Dual-Prior Collaborative Learning Guided by Multimodal Large Language Models Weakly supervised action localization by sparse temporal pooling network

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T21:36:12.410662Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-12T21:36:11.673029Z digest=sha256:6e70f87296e96f8f2fd06cb6e8a3084d2cb053359ca16282a7b6ecf071b67d8d

Observation 52fb9456-c874-41e7-997b-e4be013c4002 · outbound

This paper cites GPT-4 Technical Report.

Weakly Supervised Temporal Action Localization via Dual-Prior Collaborative Learning Guided by Multimodal Large Language Models GPT-4 Technical Report

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T21:36:12.388809Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-12T21:36:11.678221Z digest=sha256:fedf3a8c9c2e7ac8c77e2e7423c66f445eea3433dddb4571c616ef14ff027800

Observation 2ac5cd9a-1fda-4223-9a70-a8d8336a720b · outbound

This paper cites W- talc: Weakly-supervised temporal activity localization and classification.

Weakly Supervised Temporal Action Localization via Dual-Prior Collaborative Learning Guided by Multimodal Large Language Models W- talc: Weakly-supervised temporal activity localization and classification

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T21:36:12.368915Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-12T21:36:11.683356Z digest=sha256:116f01448480cf20064607d823fbb1d3277e961d3898bde2ab9f46a9ff2b6454

Observation 122c751a-a5cc-4fde-97d2-cace70bf17a2 · outbound

This paper cites Glove: Global vectors for word representation.

Weakly Supervised Temporal Action Localization via Dual-Prior Collaborative Learning Guided by Multimodal Large Language Models Glove: Global vectors for word representation

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-12T21:36:11.688551Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T21:36:11.688551Z digest=sha256:34b2b8feb51e8886f6cf3608cd37552f920bb598367a4e2ee37840945f724bc9

Observation d54afbb7-1d90-4d32-8e50-a14c043255c1 · outbound

This paper cites Proposal-based multiple instance learning for weakly-supervised temporal action localization.

Weakly Supervised Temporal Action Localization via Dual-Prior Collaborative Learning Guided by Multimodal Large Language Models Proposal-based multiple instance learning for weakly-supervised temporal action localization

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T21:36:12.331497Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-12T21:36:11.693613Z digest=sha256:e7a3c1ed7eeca286640885994a68d59cbd99a654c5f40a722430eed306e0fa52

Observation 5fc1b78f-ec68-474d-91cb-c97c1e7414f7 · outbound

This paper cites Dynamic graph modeling for weakly-supervised temporal action localization.

Weakly Supervised Temporal Action Localization via Dual-Prior Collaborative Learning Guided by Multimodal Large Language Models Dynamic graph modeling for weakly-supervised temporal action localization

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T21:36:12.312279Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-12T21:36:11.701763Z digest=sha256:d344078157cb6c60289ade3f77e2edc376081ac15b847c2c5a94f68b897f84c2

Observation 4d5875be-b636-47dd-9302-bcfec9d5cc99 · outbound

This paper cites Ddg- net: Discriminability-driven graph network for weakly- supervised temporal action localization.

Weakly Supervised Temporal Action Localization via Dual-Prior Collaborative Learning Guided by Multimodal Large Language Models Ddg- net: Discriminability-driven graph network for weakly- supervised temporal action localization

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T21:36:12.292693Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-12T21:36:11.708125Z digest=sha256:8c09b1912e29c931d99c803357fac79c157a8af2380f08e1fcfd250a0b6c167e

Observation 8c245d6f-d29c-41fb-bcc3-e5165c5a259a · outbound

This paper cites Llama 2: Open Foundation and Fine-Tuned Chat Models.

Weakly Supervised Temporal Action Localization via Dual-Prior Collaborative Learning Guided by Multimodal Large Language Models Llama 2: Open Foundation and Fine-Tuned Chat Models

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-12T21:36:11.713590Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T21:36:11.713590Z digest=sha256:6160f4008cab16c92f5c25b005858bd300c1ef67a5cd141636d3d7df2181355d

Observation bf17ea4f-44ce-4617-b95e-5b76be27a0d3 · outbound

This paper cites ChatVideo: A Tracklet-centric Multimodal and Versatile Video Understanding System.

Weakly Supervised Temporal Action Localization via Dual-Prior Collaborative Learning Guided by Multimodal Large Language Models ChatVideo: A Tracklet-centric Multimodal and Versatile Video Understanding System

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-12T21:36:11.718934Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T21:36:11.718934Z digest=sha256:e2a14c86a0a328e33f96a7f45cbb69348b4064dc977721e7f5e8320a548c43f1

Observation 447632aa-23a4-4aca-9ad1-322b9f8c00b9 · outbound

This paper cites Untrimmednets for weakly supervised action recognition and detection.

Weakly Supervised Temporal Action Localization via Dual-Prior Collaborative Learning Guided by Multimodal Large Language Models Untrimmednets for weakly supervised action recognition and detection

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T21:36:12.273423Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-12T21:36:11.725188Z digest=sha256:d4ce6f4a94fc9cf6f038350196bbe6c3e56259a8343497a1ff068dde02fae84f

Observation d2fb7309-8681-4875-ba20-01c17a89716d · outbound

This paper cites Two-stream net- works for weakly-supervised temporal action localization with semantic-aware mechanisms.

Weakly Supervised Temporal Action Localization via Dual-Prior Collaborative Learning Guided by Multimodal Large Language Models Two-stream net- works for weakly-supervised temporal action localization with semantic-aware mechanisms

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T21:36:12.255394Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-12T21:36:11.731447Z digest=sha256:815aafccdfcd05fd1ac9cddcf8ae00546563d14892ef92fc5ea463715ed7da9b

Observation f2c5a812-8387-4bf7-a140-1ced0dc93316 · outbound

This paper cites Task groupings regulariza- tion: Data-free meta-learning with heterogeneous pre-trained models.

Weakly Supervised Temporal Action Localization via Dual-Prior Collaborative Learning Guided by Multimodal Large Language Models Task groupings regulariza- tion: Data-free meta-learning with heterogeneous pre-trained models

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T21:36:12.237316Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-12T21:36:11.737018Z digest=sha256:38f6f3e90237552796d0e417539c25f26c5cfc7aae9632f9fb357d03f616a77d

Observation 56ebf1f4-8c3e-42b9-a47e-99c3b4e58fe4 · outbound

This paper cites Free: Faster and better data-free meta-learning.

Weakly Supervised Temporal Action Localization via Dual-Prior Collaborative Learning Guided by Multimodal Large Language Models Free: Faster and better data-free meta-learning

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T21:36:12.220371Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-12T21:36:11.742306Z digest=sha256:303340542776b735a4515d1cf673d772500f9642634a1f645d0be182cc07ec58

Observation 036d4442-99ef-4ce7-b05d-10dbac69da3b · outbound

This paper cites Modeling Multi-Task Model Merging as Adaptive Projective Gradient Descent.

Weakly Supervised Temporal Action Localization via Dual-Prior Collaborative Learning Guided by Multimodal Large Language Models Modeling Multi-Task Model Merging as Adaptive Projective Gradient Descent

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-12T21:36:11.747711Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T21:36:11.747711Z digest=sha256:e12be5d1bb6dcdde78e407b5f99cc8800485751f34a93d9d7890fa4778a910d7

Observation 8dbc99b1-c9b6-4c0a-8c9b-7afe697693fe · outbound

This paper cites G-tad: Sub-graph localization for tempo- ral action detection.

Weakly Supervised Temporal Action Localization via Dual-Prior Collaborative Learning Guided by Multimodal Large Language Models G-tad: Sub-graph localization for tempo- ral action detection

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T21:36:12.204095Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-12T21:36:11.752928Z digest=sha256:85cd39472719b48bfea2fe29313dd394e72f43442883a9667f2505a1e5559fc5

Observation d95f7a06-1112-4c7b-91bb-d767273b5e8e · outbound

This paper cites Uncertainty guided collaborative training for weakly supervised temporal action detection.

Weakly Supervised Temporal Action Localization via Dual-Prior Collaborative Learning Guided by Multimodal Large Language Models Uncertainty guided collaborative training for weakly supervised temporal action detection

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T21:36:12.187262Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-12T21:36:11.758271Z digest=sha256:36e47716e93a6164f5b524a3e35fe4cc143578133a1a7f2f536f540f32d71926

Observation c305d2a5-eeec-4385-9b24-2ee6712a2e21 · outbound

This paper cites Weakly-supervised temporal action localization by in- ferring salient snippet-feature.

Weakly Supervised Temporal Action Localization via Dual-Prior Collaborative Learning Guided by Multimodal Large Language Models Weakly-supervised temporal action localization by in- ferring salient snippet-feature

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T21:36:12.170720Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-12T21:36:11.763190Z digest=sha256:c49503a0f8f5b45bb98e79b322403a297a2710df9a62b2816a656620fe2d560b

Observation 814cf87b-15ff-4084-9b97-a233af6b960e · outbound

This paper cites Graph con- volutional networks for temporal action localization.

Weakly Supervised Temporal Action Localization via Dual-Prior Collaborative Learning Guided by Multimodal Large Language Models Graph con- volutional networks for temporal action localization

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T21:36:12.153636Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-12T21:36:11.768607Z digest=sha256:5efc7181c2257a86dce1be4db11c9d1a3c8d1677089c372065015a5c3275861f

Observation a1ff3c2c-7018-45bb-a709-2818e5f398b3 · outbound

This paper cites Two-stream consensus network for weakly-supervised temporal action localization.

Weakly Supervised Temporal Action Localization via Dual-Prior Collaborative Learning Guided by Multimodal Large Language Models Two-stream consensus network for weakly-supervised temporal action localization

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T21:36:12.137071Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-12T21:36:11.773598Z digest=sha256:f88403a15b7df10d100375529dfab1da444f2386718386955f6c04e2e47ea4d0

Observation 55af6a6f-961e-4397-af5d-5c24577db356 · outbound

This paper cites Cola: Weakly-supervised temporal action lo- calization with snippet contrastive learning.

Weakly Supervised Temporal Action Localization via Dual-Prior Collaborative Learning Guided by Multimodal Large Language Models Cola: Weakly-supervised temporal action lo- calization with snippet contrastive learning

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T21:36:12.117654Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-12T21:36:11.779695Z digest=sha256:24978ffb066ae7d95ded0cb86e71c4b386c6daa0b971b59fa33ef6ddf196619d

Observation 28b6fad0-024c-4177-be9a-da63dfcbf2f0 · outbound

This paper cites Video-LLaMA: An Instruction-tuned Audio-Visual Language Model for Video Understanding.

Weakly Supervised Temporal Action Localization via Dual-Prior Collaborative Learning Guided by Multimodal Large Language Models Video-LLaMA: An Instruction-tuned Audio-Visual Language Model for Video Understanding

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-12T21:36:11.784885Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T21:36:11.784885Z digest=sha256:0dc5425a3e574e6859bf9724c4c3f608c4e1a0752d6171caee07f3064a20661b

Observation dcc7179f-a301-4d6b-b156-3b52753db662 · outbound

This paper cites Distilling semantic priors from sam to efficient image restoration models.

Weakly Supervised Temporal Action Localization via Dual-Prior Collaborative Learning Guided by Multimodal Large Language Models Distilling semantic priors from sam to efficient image restoration models

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T21:36:12.098688Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-12T21:36:11.790123Z digest=sha256:aba95db18af011740ed9c4376d8f1cd967aa985716015b3c5cde1cb880d9ad5d

Observation bf015b08-6896-48a0-a4d3-683d187819c0 · outbound

This paper cites IMDPrompter: Adapting SAM to image manipulation detection by cross-view automated prompt learning.

Weakly Supervised Temporal Action Localization via Dual-Prior Collaborative Learning Guided by Multimodal Large Language Models IMDPrompter: Adapting SAM to image manipulation detection by cross-view automated prompt learning

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T21:36:12.081702Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-12T21:36:11.795002Z digest=sha256:d222ff34a8823b4761b19708f5f8d037f4963ff56e8b776dfb154fc8b7070c89

Observation c340076f-f02f-47af-8969-591674eb56e7 · outbound

This paper cites Rethinking Pseudo-Label Guided Learning for Weakly Supervised Temporal Action Localization from the Perspective of Noise Correction.

Weakly Supervised Temporal Action Localization via Dual-Prior Collaborative Learning Guided by Multimodal Large Language Models Rethinking Pseudo-Label Guided Learning for Weakly Supervised Temporal Action Localization from the Perspective of Noise Correction

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-12T21:36:11.800447Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T21:36:11.800447Z digest=sha256:a19a0d0d1302b68f7e3422bf1676db71f56a71ac2ea54c7b1c7ed40b4f4b2725

Observation 240a1cd1-2724-49ab-bbb3-5ffb4d8bd26a · outbound

This paper cites Equivalent classification mapping for weakly supervised temporal action localization.

Weakly Supervised Temporal Action Localization via Dual-Prior Collaborative Learning Guided by Multimodal Large Language Models Equivalent classification mapping for weakly supervised temporal action localization

Reference 46

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T21:36:12.063208Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-12T21:36:11.807573Z digest=sha256:c48a137c09642deab3f526c591783879fe82626752303f992652d3e2d3743e16

Observation a5af33e2-847c-46b9-b007-6d250328252d · outbound

This paper cites Temporal action detection with structured segment networks.

Weakly Supervised Temporal Action Localization via Dual-Prior Collaborative Learning Guided by Multimodal Large Language Models Temporal action detection with structured segment networks

Reference 47

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T21:36:12.046226Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-12T21:36:11.812636Z digest=sha256:92251fc27c2da3f2969edb6ab0cb48298f08ca2a9fae47ab155ccdbaaff72c24

Observation 81c3e45f-dca6-41b3-bff4-ac02220ac7c6 · outbound

This paper cites Improving weakly supervised temporal ac- tion localization by bridging train-test gap in pseudo labels.

Weakly Supervised Temporal Action Localization via Dual-Prior Collaborative Learning Guided by Multimodal Large Language Models Improving weakly supervised temporal ac- tion localization by bridging train-test gap in pseudo labels

Reference 48

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T21:36:12.028875Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-12T21:36:11.817668Z digest=sha256:80467b19f163572a356c4b3a0b0098617450ec60222165c32fd7c61a45bde0a5

Observation c12398e9-0c05-4d2b-ae6b-78b3c5635b34 · outbound

This paper cites MiniGPT-4: Enhancing Vision-Language Understanding with Advanced Large Language Models.

Weakly Supervised Temporal Action Localization via Dual-Prior Collaborative Learning Guided by Multimodal Large Language Models MiniGPT-4: Enhancing Vision-Language Understanding with Advanced Large Language Models

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-12T21:36:11.822339Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T21:36:11.822339Z digest=sha256:516450ab209d9c64b778e135ec07a169db99a5a55f8fa99bd29a661b2b02bd8c

Pith citing papers

Observation f56252a4-3235-4d7b-86b6-5dadfd0e446e · inbound

Rethinking Pseudo-Label Guided Learning for Weakly Supervised Temporal Action Localization from the Perspective of Noise Correction cites this paper.

Rethinking Pseudo-Label Guided Learning for Weakly Supervised Temporal Action Localization from the Perspective of Noise Correction Weakly Supervised Temporal Action Localization via Dual-Prior Collaborative Learning Guided by Multimodal Large Language Models

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-10T18:41:46.263737Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T18:41:46.263737Z digest=sha256:8eed5df5eb30f596625991158179918890bb53ad169baa62eaafe2d8cd2ddf5b

Observation 3678d028-2e12-4e16-9437-163810683175 · inbound

IMDPrompter: Adapting SAM to Image Manipulation Detection by Cross-View Automated Prompt Learning cites this paper.

IMDPrompter: Adapting SAM to Image Manipulation Detection by Cross-View Automated Prompt Learning Weakly Supervised Temporal Action Localization via Dual-Prior Collaborative Learning Guided by Multimodal Large Language Models

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-09T12:09:14.625616Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T12:09:14.625616Z digest=sha256:2923c091c53e7c76e25c0a5b94ca9e74ad4c6bac26339d870dcfb2471f71ec45

Observation 7a53f4de-b7b7-40fc-88cd-a4056a624f73 · inbound

$K^2$VAE: A Koopman-Kalman Enhanced Variational AutoEncoder for Probabilistic Time Series Forecasting cites this paper.

$K^2$VAE: A Koopman-Kalman Enhanced Variational AutoEncoder for Probabilistic Time Series Forecasting Weakly Supervised Temporal Action Localization via Dual-Prior Collaborative Learning Guided by Multimodal Large Language Models

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-07T13:02:35.241122Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:02:35.241122Z digest=sha256:b0ba34d43f81e0a9cc0e0e8e51ffcacf52200eb7c4936ef4d0554a1f5dde8c6f

Observation bedf9601-f849-45c1-893d-4433e4ec5471 · inbound

CLIP-AE: CLIP-assisted Cross-view Audio-Visual Enhancement for Unsupervised Temporal Action Localization cites this paper.

CLIP-AE: CLIP-assisted Cross-view Audio-Visual Enhancement for Unsupervised Temporal Action Localization Weakly Supervised Temporal Action Localization via Dual-Prior Collaborative Learning Guided by Multimodal Large Language Models

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-07T12:48:46.650668Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:48:46.650668Z digest=sha256:3ecebf1362c46872253f7d036488e2394dc254dfe93fd3d3b9c8122d161b5e07

Observation a3ed2881-9cd1-469e-bad0-9fa6ec7dcede · inbound

ComRoPE: Scalable and Robust Rotary Position Embedding Parameterized by Trainable Commuting Angle Matrices cites this paper.

ComRoPE: Scalable and Robust Rotary Position Embedding Parameterized by Trainable Commuting Angle Matrices Weakly Supervised Temporal Action Localization via Dual-Prior Collaborative Learning Guided by Multimodal Large Language Models

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-07T11:03:17.468809Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:03:17.468809Z digest=sha256:b47a32a52987f95f68a42aa35655ded451b6da4a53ff2d808f3974c7c441c89f

Observation 25f38f2e-2f51-40c5-92f9-cb4d5f11d566 · inbound

EASE: Embodied Active Event Perception via Self-Supervised Energy Minimization cites this paper.

EASE: Embodied Active Event Perception via Self-Supervised Energy Minimization Weakly Supervised Temporal Action Localization via Dual-Prior Collaborative Learning Guided by Multimodal Large Language Models

Reference 22

Resolution
verified exact
local_arxiv, observed 2026-08-06T23:35:05.494183Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-06T23:35:02.407618Z digest=sha256:4dc5c776ddbb9ae094ecfbed1f2e6190fab285f351f1c84351dd38c9d4ba76c3