Pith. sign in

Paper Citation Record · LEDGER

MM-Prompt: Cross-Modal Prompt Tuning for Continual Visual Question Answering

As of 17 August 2026, this Paper Citation Record lists 47 of 47 outbound references and 1 inbound Pith citation observation for arXiv:2505.19455.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2505.19455 v2

Coverage vector

measured 47 of 47 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T14:18:45.684227Z

measured 48 of 48 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-16T06:30:59.297886+00:00

measured 1 of 1 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-02T11:29:29.185783Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

47 of 47 outbound references displayed

  • verified exact1
  • verified fuzzy41
  • unresolved5
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation e6ae922a-bab8-4215-b0b6-9dfe69934b24 · outbound

This paper cites VLC-BERT: Visual question answering with contextualized commonsense knowledge.

MM-Prompt: Cross-Modal Prompt Tuning for Continual Visual Question Answering VLC-BERT: Visual question answering with contextualized commonsense knowledge

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:18:58.338712Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-07T14:18:39.472572Z digest=sha256:01dca3a9081296b71fe54346a81c0904ee5a161809f367f6a05d102ff2b92f72

Observation ac9a5d60-9325-4594-9b14-5daab540db17 · outbound

This paper cites Align before fuse: Vision and language representation learning with momen- tum distillation.

MM-Prompt: Cross-Modal Prompt Tuning for Continual Visual Question Answering Align before fuse: Vision and language representation learning with momen- tum distillation

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:18:58.045224Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-07T14:18:39.629437Z digest=sha256:1e1b5db3c57312128dfa93b0398d95de66fed6b78abbc1cf599d0ec921aadc5d

Observation 6fe75be5-1eb4-48a5-84e7-29fca135b16d · outbound

This paper cites Vqacl: A novel visual question answering continual learning setting.

MM-Prompt: Cross-Modal Prompt Tuning for Continual Visual Question Answering Vqacl: A novel visual question answering continual learning setting

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:18:57.760674Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-07T14:18:39.809754Z digest=sha256:c222f1b0fd32327c843e6b46b8bafcd34d6ad938d31a5242525c8dfcdab1650b

Observation 36d15d62-f440-4a59-bac7-880688c382d0 · outbound

This paper cites Enhancing Continual Learning in Visual Question Answering with Modality-Aware Feature Distillation.

MM-Prompt: Cross-Modal Prompt Tuning for Continual Visual Question Answering Enhancing Continual Learning in Visual Question Answering with Modality-Aware Feature Distillation

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-07T14:18:39.958714Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:18:39.958714Z digest=sha256:268ca46607b9c4fa63b597245a7c9fb1f81378bc57c82e9dbdbf58d68693afee

Observation 9d597ad1-f91c-424b-8710-be8908819f64 · outbound

This paper cites Decouple before interact: Multi-modal prompt learning for continual visual question answering.

MM-Prompt: Cross-Modal Prompt Tuning for Continual Visual Question Answering Decouple before interact: Multi-modal prompt learning for continual visual question answering

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:18:57.445456Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-07T14:18:40.068963Z digest=sha256:0303e1b10195d8f3f4fb1da0b9a652c10cd10877dadc8d5d282b906899bea2f0

Observation cd25b4d6-6e6c-471a-b13d-58b5c22cbb64 · outbound

This paper cites RE- VIVE: Regional visual representation matters in knowledge-based visual question answering.

MM-Prompt: Cross-Modal Prompt Tuning for Continual Visual Question Answering RE- VIVE: Regional visual representation matters in knowledge-based visual question answering

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:18:58.667094Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-07T14:18:40.219348Z digest=sha256:ebc423fcbef6c2dce33b0a8bfc44819ade26f8dc3acf81a74bb55324e3e8b5bc

Observation 8c5b0a7f-e4cd-4b7c-999a-0345ea4cb14d · outbound

This paper cites Symbolic replay: Scene graph as prompt for continual learning on VQA task.

MM-Prompt: Cross-Modal Prompt Tuning for Continual Visual Question Answering Symbolic replay: Scene graph as prompt for continual learning on VQA task

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:18:57.179424Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-07T14:18:40.380500Z digest=sha256:c09b72f1422486ad5096e3ad93633d767aaddf83f0d53d7c84118804cfd5d6cc

Observation 9913f3ed-6ddc-4afe-82ca-58238da31f9c · outbound

This paper cites CluMo: Cluster-based Modality Fusion Prompt for Continual Learning in Visual Question Answering.

MM-Prompt: Cross-Modal Prompt Tuning for Continual Visual Question Answering CluMo: Cluster-based Modality Fusion Prompt for Continual Learning in Visual Question Answering

Reference 9

Resolution
verified exact
local_arxiv, observed 2026-08-07T14:18:46.155924Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-07T14:18:40.495264Z digest=sha256:b6925ba011da475310a7e6021fbdeae84a3b56e203c61cd862587daa39c72f45

Observation 3f726f8c-57e2-4b33-8db8-3ed788348b82 · outbound

This paper cites DualPrompt: Complementary prompting for rehearsal-free continual learning.

MM-Prompt: Cross-Modal Prompt Tuning for Continual Visual Question Answering DualPrompt: Complementary prompting for rehearsal-free continual learning

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:18:56.940890Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-07T14:18:40.652713Z digest=sha256:55108c34da69a3668edb028e5a7fe7f267815d867cc3ef8ad3b5a884fb56fcde

Observation 43e4801c-2209-4de4-8d8f-b00eaea56e1f · outbound

This paper cites Learning to prompt for continual learning.

MM-Prompt: Cross-Modal Prompt Tuning for Continual Visual Question Answering Learning to prompt for continual learning

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:18:56.599667Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-07T14:18:40.777953Z digest=sha256:bb09bf860f166d3fb0637c381b055f3df5f9a980fda8df875b019c9047628b07

Observation 6f9b46ce-0b9f-4b0c-89ae-6c2668842733 · outbound

This paper cites CODA-Prompt: Contin- ual decomposed attention-based prompting for rehearsal-free continual learning.

MM-Prompt: Cross-Modal Prompt Tuning for Continual Visual Question Answering CODA-Prompt: Contin- ual decomposed attention-based prompting for rehearsal-free continual learning

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:18:56.343115Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-07T14:18:40.943821Z digest=sha256:c4650bc0f2ed30ea777d2a8711bf4bf2acd629a803924793849ae5b0d3141a0a

Observation 22dffaf8-fe08-4f55-aee0-fb0b0bd2dddf · outbound

This paper cites Semantic residual prompts for continual learning.

MM-Prompt: Cross-Modal Prompt Tuning for Continual Visual Question Answering Semantic residual prompts for continual learning

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:18:56.052264Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-07T14:18:41.055708Z digest=sha256:fd3d58a71dca0e3131027a714e5f1d27ae86f8a012b025acc1a3d4185b96641d

Observation 6e32edd3-7d65-4cff-867b-0aad9e513f3d · outbound

This paper cites Multi-domain multi- task rehearsal for lifelong learning.

MM-Prompt: Cross-Modal Prompt Tuning for Continual Visual Question Answering Multi-domain multi- task rehearsal for lifelong learning

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:18:55.799111Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-07T14:18:41.235906Z digest=sha256:daa1063780c4dffd1a8fc9aab0db69f501729fd0fa8c93f2794ab18c747893db

Observation ab5d9594-1fa3-4350-9a12-5db1c8d4b6bc · outbound

This paper cites Exploring example influence in continual learning.

MM-Prompt: Cross-Modal Prompt Tuning for Continual Visual Question Answering Exploring example influence in continual learning

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:18:55.291288Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-07T14:18:41.384164Z digest=sha256:c1a2618802d2e558ba127fa05cb895e08f5f0d0f563ec120f4c31830fc6323ff

Observation 9b251139-b4cc-4378-9ea7-1bfc4beb2aac · outbound

This paper cites Measuring asymmetric gradient discrepancy in parallel continual learning.

MM-Prompt: Cross-Modal Prompt Tuning for Continual Visual Question Answering Measuring asymmetric gradient discrepancy in parallel continual learning

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:18:55.030671Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-07T14:18:41.574282Z digest=sha256:d2586b60a58671721c2e8c07975f722ec1cdafa29309c24ed254a8888fbf3bda

Observation 5db16f86-a349-4ba3-83fa-eb82eecf4be6 · outbound

This paper cites MAPLE: Multi-modal prompt learning.

MM-Prompt: Cross-Modal Prompt Tuning for Continual Visual Question Answering MAPLE: Multi-modal prompt learning

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:18:54.732552Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-07T14:18:41.733658Z digest=sha256:a651c47623eeada3eb04b78ff3174b2e3f2a150e1ba04b9fa6f749e497fa8fd5

Observation 43ad371d-77c6-4d4f-a749-2b9d63cc585f · outbound

This paper cites Overcoming language priors in visual question answering with adversarial regularization.

MM-Prompt: Cross-Modal Prompt Tuning for Continual Visual Question Answering Overcoming language priors in visual question answering with adversarial regularization

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:18:54.412802Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-07T14:18:41.933587Z digest=sha256:7792a3185edf218a5cc20b1354de4382fd61a3bbeb416857aa957e14a32b43a3

Observation 237b70e3-dc83-488d-a892-9b2c00d311a7 · outbound

This paper cites Making the v in VQA matter: Elevating the role of image understanding in visual question answering.

MM-Prompt: Cross-Modal Prompt Tuning for Continual Visual Question Answering Making the v in VQA matter: Elevating the role of image understanding in visual question answering

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:18:54.060826Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-07T14:18:42.083469Z digest=sha256:12547ddc96a76a81af614f43a91a9d4dfdeb09707c2976971040202c8a168008

Observation 70209ef8-75c9-493f-a996-d6ae11ef9323 · outbound

This paper cites Lawrence Zitnick, and Devi Parikh.

MM-Prompt: Cross-Modal Prompt Tuning for Continual Visual Question Answering Lawrence Zitnick, and Devi Parikh

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:18:53.736161Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-07T14:18:42.187884Z digest=sha256:c55077ca232eab9f64f5653eac9dcdebb263a3ec83a21295f4cc0051eb17a905

Observation c80d7d0b-7d58-4398-8070-651ddba2b3c0 · outbound

This paper cites A lifelong learning perspective for mobile robot control.

MM-Prompt: Cross-Modal Prompt Tuning for Continual Visual Question Answering A lifelong learning perspective for mobile robot control

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:18:53.444496Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-07T14:18:42.329149Z digest=sha256:649297fa4635eb1ffddd3bef4e451144af14fc306af61ab08d3ddddeff374bb7

Observation 5c70ba20-7bea-45a9-8d9e-2e75add3a03d · outbound

This paper cites Pre- venting zero-shot transfer degradation in continual learning of vision-language models.

MM-Prompt: Cross-Modal Prompt Tuning for Continual Visual Question Answering Pre- venting zero-shot transfer degradation in continual learning of vision-language models

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:18:53.172915Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-07T14:18:42.506071Z digest=sha256:d550836daa033a58048d30731a9fa26752aed4251ed92c5c5e937ac02e3c5909

Observation 21d07dd1-800e-4786-85b1-bb83db898b79 · outbound

This paper cites Bakker, Nicu Sebe, and Michael S.

MM-Prompt: Cross-Modal Prompt Tuning for Continual Visual Question Answering Bakker, Nicu Sebe, and Michael S

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:18:52.778162Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-07T14:18:42.625308Z digest=sha256:b8c804715969780936b62efa43c62f1bbfd1960f8c8537af0c8eec130b95e9aa

Observation 6cfdf7fe-11d6-4319-99c8-b6162bfe8e5a · outbound

This paper cites Boosting continual learning of vision-language models via mixture-of-experts adapters.

MM-Prompt: Cross-Modal Prompt Tuning for Continual Visual Question Answering Boosting continual learning of vision-language models via mixture-of-experts adapters

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:18:52.553431Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-07T14:18:42.809398Z digest=sha256:c8c6f8b6197b0c4a0f0903d7e1e88336c0c82b06f03e212b284dd6515fe9913e

Observation 3e396b1e-cf57-454d-9641-ba791695c299 · outbound

This paper cites Learn to grow: A continual structure learning framework for overcoming catastrophic forgetting.

MM-Prompt: Cross-Modal Prompt Tuning for Continual Visual Question Answering Learn to grow: A continual structure learning framework for overcoming catastrophic forgetting

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:18:52.265906Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-07T14:18:42.913039Z digest=sha256:1e75a0db32e8c19344f59f1ed81c91ae6e45848913c968b9150056c9e876d2d3

Observation 631255ab-0ad6-4160-b3c4-f0f9169afc88 · outbound

This paper cites Ex- perience replay for continual learning.

MM-Prompt: Cross-Modal Prompt Tuning for Continual Visual Question Answering Ex- perience replay for continual learning

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:18:51.993430Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-07T14:18:43.051597Z digest=sha256:f4401b39a779d6a7a65daad306f9c6ed499f8a7a581aa952ac4f7495b41fbaa7

Observation 6f3b90d0-62f3-4fae-8bd0-f18f9d6b6bbc · outbound

This paper cites Dark experience for general continual learning: A strong, simple baseline.

MM-Prompt: Cross-Modal Prompt Tuning for Continual Visual Question Answering Dark experience for general continual learning: A strong, simple baseline

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:18:51.689768Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-07T14:18:43.197659Z digest=sha256:67711b0f157af58fac61ad85a28251f524ad025fe0ddf7ac0c438948f1a22f89

Observation 803b3c9d-a998-4371-8ec5-bca40dc484e1 · outbound

This paper cites A continual learning survey: Defying forgetting in classification tasks.IEEE Transactions on Pattern Analysis and Machine Intelligence, pages 3366–3385, 2021.

MM-Prompt: Cross-Modal Prompt Tuning for Continual Visual Question Answering A continual learning survey: Defying forgetting in classification tasks.IEEE Transactions on Pattern Analysis and Machine Intelligence, pages 3366–3385, 2021

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:18:51.397181Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-07T14:18:43.327033Z digest=sha256:1fa007456d61657819c0ea24a100958da557f65f35f9ce3efb33aa02fde34d11

Observation c174679f-b047-4d86-921b-77da94b15948 · outbound

This paper cites Yu, and Irwin King.

MM-Prompt: Cross-Modal Prompt Tuning for Continual Visual Question Answering Yu, and Irwin King

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-07T14:18:43.480038Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:18:43.480038Z digest=sha256:7fdb69ba51765fc5e72bce64911712f417affd3e39fb44d20ea842188f6a41b6

Observation 743388c0-d14e-45d7-ae35-51a1893b90d8 · outbound

This paper cites Balanced multimodal learning via on-the-fly gradient modulation.

MM-Prompt: Cross-Modal Prompt Tuning for Continual Visual Question Answering Balanced multimodal learning via on-the-fly gradient modulation

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:18:51.120558Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-07T14:18:43.549152Z digest=sha256:56c1ebf5731f43d737fb73e593e53ac421a9572f2d9b9bf3200f46e00c037edf

Observation cdb1a1a3-d0bf-42d0-a702-8bc1cae5276f · outbound

This paper cites Pre-trained models: Past, present and future.AI Open, pages 225–250, 2021.

MM-Prompt: Cross-Modal Prompt Tuning for Continual Visual Question Answering Pre-trained models: Past, present and future.AI Open, pages 225–250, 2021

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:18:50.865018Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-07T14:18:43.646275Z digest=sha256:b3cb51a6ed80f601c2ee348b56e90899b14ffe695b1e77b44b4053dd6c4eccc7

Observation 86c661f7-5eef-4423-b75d-78c4f288c359 · outbound

This paper cites an unresolved cited work.

MM-Prompt: Cross-Modal Prompt Tuning for Continual Visual Question Answering Unresolved cited work

Reference 32

Resolution
unresolved
raw_fallback, observed 2026-08-07T14:18:50.550263Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-07T14:18:43.718731Z digest=sha256:08e0d526c260f3e6f60102fb7743ff4813aff0ba4bbf3f3b46dbffa13ca65a04

Observation e3966b13-3e4a-4d4c-850c-bd9290508bb3 · outbound

This paper cites LXMERT: Learning Cross-Modality Encoder Representations from Transformers.

MM-Prompt: Cross-Modal Prompt Tuning for Continual Visual Question Answering LXMERT: Learning Cross-Modality Encoder Representations from Transformers

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-07T14:18:43.801881Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:18:43.801881Z digest=sha256:b8a70d80e3ad453a21768e599eb69df3f0aab4944871f309ce67699bbff8cf94

Observation 955dbc63-ca53-4d0c-aa9a-67732d5d63cb · outbound

This paper cites DyTox: Trans- formers for continual learning with dynamic token expansion.

MM-Prompt: Cross-Modal Prompt Tuning for Continual Visual Question Answering DyTox: Trans- formers for continual learning with dynamic token expansion

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:18:50.197997Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-07T14:18:43.925124Z digest=sha256:f6286a7a460a1136d8702fdf494aa1abc581cc085f5a98d26b49ed0a588de5bf

Observation 4bff90a1-d510-4dca-94ec-7801bc613390 · outbound

This paper cites Learning transferable visual models from natural language supervision.

MM-Prompt: Cross-Modal Prompt Tuning for Continual Visual Question Answering Learning transferable visual models from natural language supervision

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:18:49.892421Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-07T14:18:44.029603Z digest=sha256:1bbb17a35997dd24c92e53b91cf8e515a0983ec4787c3d79ae15bc655baad51b

Observation 62044c85-dfe2-4f54-b81a-d161c3060f30 · outbound

This paper cites Understanding driving risks via prompt learning.

MM-Prompt: Cross-Modal Prompt Tuning for Continual Visual Question Answering Understanding driving risks via prompt learning

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:18:49.565975Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-07T14:18:44.129858Z digest=sha256:5ea8ba0027d3ba86ccf4a41478150b0ee0979db66f60625803b9e3c4d9c6f3f4

Observation 8bd7b5b4-bdea-432a-861b-01db032a8db2 · outbound

This paper cites Attention bottlenecks for multimodal fusion.

MM-Prompt: Cross-Modal Prompt Tuning for Continual Visual Question Answering Attention bottlenecks for multimodal fusion

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:18:49.297569Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-07T14:18:44.250330Z digest=sha256:8e47cf76d4106a454a0bf9ac51d534e0ff4be46026460574e807157ea9c83f5a

Observation 70997690-a4ce-45f6-9772-630cf5f68273 · outbound

This paper cites Difnet: Boosting visual information flow for image captioning.

MM-Prompt: Cross-Modal Prompt Tuning for Continual Visual Question Answering Difnet: Boosting visual information flow for image captioning

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:18:48.980507Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-07T14:18:44.369136Z digest=sha256:5cd77dec347d6d11899e290f2b728aeada31fbe6825cdb4f39b1824c75a7cb0e

Observation d494e25d-0efa-4917-8cee-caa28cd4a5f1 · outbound

This paper cites Aligning visual regions and textual concepts for semantic-grounded image representations.

MM-Prompt: Cross-Modal Prompt Tuning for Continual Visual Question Answering Aligning visual regions and textual concepts for semantic-grounded image representations

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:18:48.663665Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-07T14:18:44.478499Z digest=sha256:4867d50f596916cfc48e5623aede9aae552ebc03358b85094df9a8170b15edef

Observation 7bf90daf-328b-4d20-9e9a-b12e51906647 · outbound

This paper cites Masked autoencoders are scalable vision learners.

MM-Prompt: Cross-Modal Prompt Tuning for Continual Visual Question Answering Masked autoencoders are scalable vision learners

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:18:48.441359Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-07T14:18:44.622402Z digest=sha256:76d42cbfb37eabe740b58c6045d208ebfee762c891afcb8f326c87bcffbb7f3e

Observation cb88f53b-56bb-45d3-a44b-db000d993c3a · outbound

This paper cites A Survey on Masked Autoencoder for Self-supervised Learning in Vision and Beyond.

MM-Prompt: Cross-Modal Prompt Tuning for Continual Visual Question Answering A Survey on Masked Autoencoder for Self-supervised Learning in Vision and Beyond

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-07T14:18:44.745026Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:18:44.745026Z digest=sha256:46dd3fd41b7a05c36c52444b2f39b91c1acdd9f024c9b9565ee140a3a07611d2

Observation ccaa9d9f-cdee-4e79-bc59-91990f4f9e7f · outbound

This paper cites Can we gain more from orthogonality regularizations in training deep networks? InAdvances in Neural Information Processing Systems (NeurIPS), 2018.

MM-Prompt: Cross-Modal Prompt Tuning for Continual Visual Question Answering Can we gain more from orthogonality regularizations in training deep networks? InAdvances in Neural Information Processing Systems (NeurIPS), 2018

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:18:48.119044Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-07T14:18:44.854051Z digest=sha256:acbb54af3a679a19a05cbb36df9bd904bf7fc8cf78014e4873baec4ddb8182de

Observation 96122ca3-70b4-4fb7-9c68-0b641f3927a0 · outbound

This paper cites NExT-QA: Next phase of question answering to explaining temporal actions.

MM-Prompt: Cross-Modal Prompt Tuning for Continual Visual Question Answering NExT-QA: Next phase of question answering to explaining temporal actions

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:18:47.810672Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-07T14:18:44.946111Z digest=sha256:89f944235b2fc7de889ced60b3171e4a0e65a6ab5688b59d993257891a95bf37

Observation 72d88bc0-f0aa-417a-91b8-037b2f7322c4 · outbound

This paper cites A simple weight decay can improve generalization.

MM-Prompt: Cross-Modal Prompt Tuning for Continual Visual Question Answering A simple weight decay can improve generalization

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:18:47.551554Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-07T14:18:45.088368Z digest=sha256:5d108f9034b192724400e43af1fc7109a237c45341f7d2464d47c55b0a3e2a18

Observation 2e79ac51-df0b-40c6-91ca-cc1e558b8d17 · outbound

This paper cites Bottom-up and top-down attention for image captioning and visual question answering.

MM-Prompt: Cross-Modal Prompt Tuning for Continual Visual Question Answering Bottom-up and top-down attention for image captioning and visual question answering

Reference 45

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:18:47.282811Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-07T14:18:45.264873Z digest=sha256:d404e89249069580b94c09de4af911928efdf001c539dafef628c6831cf8a217

Observation 0aa209b1-8d7c-4312-9310-a85912551606 · outbound

This paper cites Visualizing data using t-SNE.Journal of Machine Learning Research, pages 2579–2605, 2008.

MM-Prompt: Cross-Modal Prompt Tuning for Continual Visual Question Answering Visualizing data using t-SNE.Journal of Machine Learning Research, pages 2579–2605, 2008

Reference 46

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:18:46.956737Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-07T14:18:45.420900Z digest=sha256:44df5ad734fc3b999091e63de30ac08cc29d90a34aaf88b5d16ff3ce1ed386e8

Observation 3eb9922d-d7b6-4c2f-9173-1f5e80ce4b86 · outbound

This paper cites Scaling instruction-finetuned language models.Journal of Machine Learning Research, 2024.

MM-Prompt: Cross-Modal Prompt Tuning for Continual Visual Question Answering Scaling instruction-finetuned language models.Journal of Machine Learning Research, 2024

Reference 47

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:18:46.672087Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-07T14:18:45.539148Z digest=sha256:5bfa7d95e5f8a40371cd654b71f8e3ea32cb5a2c3a5b87e43659ac8969459835

Observation f0ed1da2-ed8a-4d9b-a2b2-41744b569d0b · outbound

This paper cites Plus”, “Mean Pooling.

MM-Prompt: Cross-Modal Prompt Tuning for Continual Visual Question Answering Plus”, “Mean Pooling

Reference 48

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:18:46.411761Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-07T14:18:45.684227Z digest=sha256:7863228a751ab652fbfaf50aa2d28c7255da718839c309cc202e6d2e28350e5e

Pith citing papers

Observation 2d7753f2-fd46-4f6c-a6d5-04b682a174bb · inbound

Group Preference Collapse in Personalized Multimodal Large Language Models cites this paper.

Group Preference Collapse in Personalized Multimodal Large Language Models MM-Prompt: Cross-Modal Prompt Tuning for Continual Visual Question Answering

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-02T11:29:29.185783Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T11:29:29.185783Z digest=sha256:a07d1aee33000886680d11e2a8b8f7d1e14d172e9754b3c991feccb50e580c3f