Pith. sign in

Paper Citation Record · LEDGER

MambaTrans: Multimodal Fusion Image Translation via Large Language Model Priors for Downstream Visual Tasks

As of 11 August 2026, this Paper Citation Record lists 47 of 47 outbound references and 1 inbound Pith citation observation for arXiv:2508.07803.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2508.07803 v1

Coverage vector

measured 47 of 47 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-05T21:52:25.017468Z

measured 48 of 48 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-10T06:31:04.303077+00:00

measured 1 of 1 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-07-11T23:54:32.839280Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

47 of 47 outbound references displayed

  • verified exact2
  • verified fuzzy38
  • unresolved7
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 168e4721-7e28-416b-8eea-e6e21ab8983a · outbound

This paper cites A deep learning framework for infrared and visible image fusion without strict registration.

MambaTrans: Multimodal Fusion Image Translation via Large Language Model Priors for Downstream Visual Tasks A deep learning framework for infrared and visible image fusion without strict registration

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T21:52:33.914320Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-05T21:52:20.843522Z digest=sha256:90c04c4c9dcf2ece269d60afbc449db910ae1227a686ef3f7ff2bd4e3c6fc0a2

Observation a5777d18-1a03-47f1-8b16-353c70a1a5b0 · outbound

This paper cites All-weather multi-modality image fusion: Unified framework and 100k benchmark.

MambaTrans: Multimodal Fusion Image Translation via Large Language Model Priors for Downstream Visual Tasks All-weather multi-modality image fusion: Unified framework and 100k benchmark

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-05T21:52:20.944258Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T21:52:20.944258Z digest=sha256:a7d1d8a8e9509b2cd813944dc5d0a872e1a2d4dbe5a8165cf39b3102fc285bba

Observation 50ddb353-3e72-4768-adcb-9ec1833028c3 · outbound

This paper cites Infrared and visible image fusion based on domain transform filtering and sparse representation.

MambaTrans: Multimodal Fusion Image Translation via Large Language Model Priors for Downstream Visual Tasks Infrared and visible image fusion based on domain transform filtering and sparse representation

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T21:52:33.758344Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-05T21:52:21.052336Z digest=sha256:b28dacce6bed3cf5719ec0e042badae333ef209a8d8aee1189d280f6be9a4d57

Observation f646ed0c-ad4d-4340-860a-01556952b60f · outbound

This paper cites Infrared and visible image fusion: From data compatibility to task adaption.

MambaTrans: Multimodal Fusion Image Translation via Large Language Model Priors for Downstream Visual Tasks Infrared and visible image fusion: From data compatibility to task adaption

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T21:52:33.536389Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-05T21:52:21.160857Z digest=sha256:7a3f5f1fd0b0cebb6d6dd9bf83581a64c03ab04305041aa0e3b146772537a6d5

Observation 9e8b1584-9e0b-4dbf-8c83-bc5edf9765eb · outbound

This paper cites Simultaneous tri-modal medical image fusion and super-resolution using conditional diffusion model.

MambaTrans: Multimodal Fusion Image Translation via Large Language Model Priors for Downstream Visual Tasks Simultaneous tri-modal medical image fusion and super-resolution using conditional diffusion model

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T21:52:33.295541Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-05T21:52:21.259991Z digest=sha256:27253d0c5ab7136793cbd7559fc82f6db5f013277648766096bc572c5e9d0fc7

Observation f2d9ac42-ac80-4ce0-bf67-5b308133d718 · outbound

This paper cites Fast r-cnn.

MambaTrans: Multimodal Fusion Image Translation via Large Language Model Priors for Downstream Visual Tasks Fast r-cnn

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T21:52:33.066319Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-05T21:52:21.481919Z digest=sha256:a3c391f975dda2045d098e9373de9885c397ca0a7528befc67ba7a552f1129d1

Observation f8ae765a-2b45-451b-a981-b47b86160e2c · outbound

This paper cites Microsoft coco: Common objects in context.

MambaTrans: Multimodal Fusion Image Translation via Large Language Model Priors for Downstream Visual Tasks Microsoft coco: Common objects in context

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T21:52:32.837681Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-05T21:52:21.572846Z digest=sha256:d154429a76f0526aed22129a98708d902d7f1e5e305b0eeaf8751ad32f300a63

Observation 95698043-7484-408d-aca9-888a1a635f15 · outbound

This paper cites Visualizing data using t-sne.

MambaTrans: Multimodal Fusion Image Translation via Large Language Model Priors for Downstream Visual Tasks Visualizing data using t-sne

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-05T21:52:21.713153Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T21:52:21.713153Z digest=sha256:272a9acdf88234de19812066e8962bc7643cdbda3fadffa1855582b63ac71647

Observation 3dc53547-8931-4a95-a8f5-f46167226cb4 · outbound

This paper cites Tsjnet: A multi-modality target and semantic awareness joint-driven image fusion network.

MambaTrans: Multimodal Fusion Image Translation via Large Language Model Priors for Downstream Visual Tasks Tsjnet: A multi-modality target and semantic awareness joint-driven image fusion network

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-05T21:52:21.795023Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T21:52:21.795023Z digest=sha256:5d4b41da68f76e123024762f56adaa4e42ab3d45170c8d0f4c0b199b7b066103

Observation 5957c9c7-fda6-4c8a-8295-3b4d8b2673e0 · outbound

This paper cites Target-aware dual adversarial learning and a multi-scenario multi-modality benchmark to fuse infrared and visible for object detection.

MambaTrans: Multimodal Fusion Image Translation via Large Language Model Priors for Downstream Visual Tasks Target-aware dual adversarial learning and a multi-scenario multi-modality benchmark to fuse infrared and visible for object detection

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T21:52:32.653248Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-05T21:52:21.874881Z digest=sha256:79beee1243fb6c9b1dc199ef0a36bd0a6444d858fbd202ffc369d9e474d65871

Observation 0baafcb8-965e-4a87-9a73-8d1f3a5b69ed · outbound

This paper cites Fs-diff: Semantic guidance and clarity-aware simultaneous multimodal image fusion and super-resolution.

MambaTrans: Multimodal Fusion Image Translation via Large Language Model Priors for Downstream Visual Tasks Fs-diff: Semantic guidance and clarity-aware simultaneous multimodal image fusion and super-resolution

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T21:52:32.377319Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-05T21:52:21.983819Z digest=sha256:0c0ad8462efced18af771f9338ec28c2b51d035c6d7a51faf06e875257f25768

Observation f2a1e272-e385-4dd3-bc79-ab46406dff3e · outbound

This paper cites A task-guided, implicitly-searched and metainitialized deep model for image fusion.

MambaTrans: Multimodal Fusion Image Translation via Large Language Model Priors for Downstream Visual Tasks A task-guided, implicitly-searched and metainitialized deep model for image fusion

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T21:52:32.242375Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-05T21:52:22.083478Z digest=sha256:a3ee9d69014df5968dc74b0887f017ddfcff9a0eccdc07c1a300e39396dc6323

Observation 330debc3-4de8-4489-87c9-cdf7d6008795 · outbound

This paper cites Image fusion in the loop of high-level vision tasks: A semantic-aware real-time infrared and visible image fusion network.

MambaTrans: Multimodal Fusion Image Translation via Large Language Model Priors for Downstream Visual Tasks Image fusion in the loop of high-level vision tasks: A semantic-aware real-time infrared and visible image fusion network

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T21:52:32.072413Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-05T21:52:22.190602Z digest=sha256:ddd5c33e8d327e2759c8f2a9a6c81d1e2676423dc316cec719c4f9578acc2502

Observation e1cbd90c-fd38-4701-8161-0ce3eb1aa078 · outbound

This paper cites Mrfs: Mutually reinforcing image fusion and segmentation.

MambaTrans: Multimodal Fusion Image Translation via Large Language Model Priors for Downstream Visual Tasks Mrfs: Mutually reinforcing image fusion and segmentation

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T21:52:31.755632Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-05T21:52:22.300125Z digest=sha256:7d332c6d70ebd50335eadd47e61382bc2a555b91fe880684c5c1cfd743985985

Observation 3fa7d815-f3a3-414c-9f10-8e61c88de446 · outbound

This paper cites Image-to-image translation: Methods and applications.

MambaTrans: Multimodal Fusion Image Translation via Large Language Model Priors for Downstream Visual Tasks Image-to-image translation: Methods and applications

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T21:52:31.569164Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-05T21:52:22.415638Z digest=sha256:c518e94db5644176a3044c98738aa1c5cd25e416c7393a3aafc8ca91d4bd674f

Observation 628f808a-9b94-4aa0-9ef2-77bf7ac7885a · outbound

This paper cites Source-free open compound domain adaptation in semantic segmentation.

MambaTrans: Multimodal Fusion Image Translation via Large Language Model Priors for Downstream Visual Tasks Source-free open compound domain adaptation in semantic segmentation

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T21:52:31.357536Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-05T21:52:22.499013Z digest=sha256:1d679f5955007c1b4efa91cfcac3c804c6305478c5de49b64d512857083a0b56

Observation 0d2eabd2-5c76-49ee-ac91-80f5a3b7dafb · outbound

This paper cites Infragan: A gan architecture to transfer visible images to infrared domain.

MambaTrans: Multimodal Fusion Image Translation via Large Language Model Priors for Downstream Visual Tasks Infragan: A gan architecture to transfer visible images to infrared domain

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T21:52:31.076644Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-05T21:52:22.571106Z digest=sha256:c9fe3f3c47b03f50ba203b9c5660c8d8b49189fcbf07507589637e7f0ec10505

Observation 37b2628c-cfeb-4652-a9a3-197fa0478a0b · outbound

This paper cites Cnn-based thermal infrared person detection by domain adaptation.

MambaTrans: Multimodal Fusion Image Translation via Large Language Model Priors for Downstream Visual Tasks Cnn-based thermal infrared person detection by domain adaptation

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T21:52:30.887222Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-05T21:52:22.633845Z digest=sha256:d790488c3dbeda953d01a17ed60fefdaefd3296d30930c85684c3ec105a74037

Observation 56bfd79f-6e98-45f7-974b-c32740ef96a2 · outbound

This paper cites Hallucidet: hallucinating rgb modality for person detection through privileged information.

MambaTrans: Multimodal Fusion Image Translation via Large Language Model Priors for Downstream Visual Tasks Hallucidet: hallucinating rgb modality for person detection through privileged information

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T21:52:30.699864Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-05T21:52:22.737979Z digest=sha256:3880f667ced6d1526a36659a1bf8bcee17bb7ea468180c6f02c82a2c1ed4e77d

Observation 7a470abe-68f6-4c2d-ac65-ed82df8276a5 · outbound

This paper cites Modality translation for object detection adaptation without forgetting prior knowledge.

MambaTrans: Multimodal Fusion Image Translation via Large Language Model Priors for Downstream Visual Tasks Modality translation for object detection adaptation without forgetting prior knowledge

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T21:52:30.456942Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-05T21:52:22.812104Z digest=sha256:e80611b4adfb4f7338275f980196549e02b6d1b547d48452ec1e0b6cc870f948

Observation 42e047fa-8044-48e9-8f00-d3fc651b7201 · outbound

This paper cites Mamba: Linear-Time Sequence Modeling with Selective State Spaces.

MambaTrans: Multimodal Fusion Image Translation via Large Language Model Priors for Downstream Visual Tasks Mamba: Linear-Time Sequence Modeling with Selective State Spaces

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-05T21:52:22.912267Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T21:52:22.912267Z digest=sha256:7a1cd012dd46dbffe0d95eec7c1c19c92869802d481ec93c5672f35fe24749e5

Observation a5d5d2c5-e078-4562-b432-69a715a0aae2 · outbound

This paper cites Vmamba: Visual state space model.

MambaTrans: Multimodal Fusion Image Translation via Large Language Model Priors for Downstream Visual Tasks Vmamba: Visual state space model

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T21:52:30.269904Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-05T21:52:22.998893Z digest=sha256:9e689219be6786db623a46c2dfc3fa9bcae5012563d6d6a23007409d3ab78594

Observation e0502ec9-18c2-4fc5-8539-79198965130f · outbound

This paper cites Mambair: A simple baseline for image restoration with state-space model.

MambaTrans: Multimodal Fusion Image Translation via Large Language Model Priors for Downstream Visual Tasks Mambair: A simple baseline for image restoration with state-space model

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T21:52:30.002559Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-05T21:52:23.077544Z digest=sha256:32838b5e44286d0fd08e3c853561abcade04c17d34d913ad771a4a8417e109b1

Observation 324bcfbb-487b-47e7-ae1e-64013fd9f8a8 · outbound

This paper cites Vision mamba: efficient visual representation learning with bidirectional state space model.

MambaTrans: Multimodal Fusion Image Translation via Large Language Model Priors for Downstream Visual Tasks Vision mamba: efficient visual representation learning with bidirectional state space model

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T21:52:29.682263Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-05T21:52:23.182088Z digest=sha256:9dc4fa74d14c60e8910b85e01e9f1b53586dd9ef4d6fb360162276fc364dfdff

Observation 6e766b80-804f-4162-839a-63acd56fd705 · outbound

This paper cites GLU Variants Improve Transformer.

MambaTrans: Multimodal Fusion Image Translation via Large Language Model Priors for Downstream Visual Tasks GLU Variants Improve Transformer

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-05T21:52:23.234051Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T21:52:23.234051Z digest=sha256:9079f2ec65642f5090fdc8e7a1337bda72a5ac018cf4aea9d7d887c377d426d1

Observation 0a9be89a-6367-4f9a-82d8-b1f73f844dfc · outbound

This paper cites Faster r-cnn: Towards real-time object detection with region proposal networks.

MambaTrans: Multimodal Fusion Image Translation via Large Language Model Priors for Downstream Visual Tasks Faster r-cnn: Towards real-time object detection with region proposal networks

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-05T21:52:23.332065Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T21:52:23.332065Z digest=sha256:7fe4c5a7666ef424221501df43d99fa59ff680bb16f93a7f3175789c1bb35fae

Observation d099b7d5-f44e-47f0-8b07-4acc0cfe7479 · outbound

This paper cites Piafusion: A progressive infrared and visible image fusion network based on illumination aware.

MambaTrans: Multimodal Fusion Image Translation via Large Language Model Priors for Downstream Visual Tasks Piafusion: A progressive infrared and visible image fusion network based on illumination aware

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T21:52:29.362717Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-05T21:52:23.431911Z digest=sha256:1d2b87348f2260d5dbb772011c538c1945d7e69a33058796c3ab45f091e90642

Observation 5fa405c6-2f88-4b12-aa1f-b3ab263b726c · outbound

This paper cites Every SAM Drop Counts: Embracing Semantic Priors for Multi-Modality Image Fusion and Beyond.

MambaTrans: Multimodal Fusion Image Translation via Large Language Model Priors for Downstream Visual Tasks Every SAM Drop Counts: Embracing Semantic Priors for Multi-Modality Image Fusion and Beyond

Reference 29

Resolution
verified exact
local_arxiv, observed 2026-08-05T21:52:25.307217Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-05T21:52:23.523944Z digest=sha256:549052931045773622d36e4ec00de0770257faf5eec8142b30860da9dace5712

Observation c6a96941-b04a-45d1-b68f-934966f34a9e · outbound

This paper cites DCEvo: Discriminative Cross-Dimensional Evolutionary Learning for Infrared and Visible Image Fusion.

MambaTrans: Multimodal Fusion Image Translation via Large Language Model Priors for Downstream Visual Tasks DCEvo: Discriminative Cross-Dimensional Evolutionary Learning for Infrared and Visible Image Fusion

Reference 30

Resolution
verified exact
local_arxiv, observed 2026-08-05T21:52:25.159340Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-05T21:52:23.586355Z digest=sha256:2a9ca278b35179388166440cafa40192620fa4011218b24eebd813331603f905

Observation a4137315-f119-4e1f-854a-992dc18a4587 · outbound

This paper cites Where elegance meets precision: towards a compact, automatic, and flexible framework for multi-modality image fusion and applications.

MambaTrans: Multimodal Fusion Image Translation via Large Language Model Priors for Downstream Visual Tasks Where elegance meets precision: towards a compact, automatic, and flexible framework for multi-modality image fusion and applications

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T21:52:29.079716Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-05T21:52:23.647965Z digest=sha256:473377e8ad62d81d272fa9a2f426ff0c1c0021f676b555374f3f8ea777424a95

Observation 5dc11091-632f-42ca-ac3b-9c475a758d04 · outbound

This paper cites Equivariant multi-modality image fusion.

MambaTrans: Multimodal Fusion Image Translation via Large Language Model Priors for Downstream Visual Tasks Equivariant multi-modality image fusion

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T21:52:28.797720Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-05T21:52:23.710572Z digest=sha256:54a7deb206053b9363b84e5da488983a3482d575c24ffc6a55b6931e8005c81c

Observation 255a2d7b-c502-452d-b0b0-523e20e56bb5 · outbound

This paper cites Cddfuse: Correlation-driven dual-branch feature decomposition for multi-modality image fusion.

MambaTrans: Multimodal Fusion Image Translation via Large Language Model Priors for Downstream Visual Tasks Cddfuse: Correlation-driven dual-branch feature decomposition for multi-modality image fusion

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T21:52:28.529331Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-05T21:52:23.786134Z digest=sha256:89544df59266bac5700af583ebac130644f1cc6d808dd326b279fdecae854140

Observation 6273535e-be8e-4d07-85b2-fb4cdc44f161 · outbound

This paper cites Probing synergistic high-order interaction in infrared and visible image fusion.

MambaTrans: Multimodal Fusion Image Translation via Large Language Model Priors for Downstream Visual Tasks Probing synergistic high-order interaction in infrared and visible image fusion

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T21:52:28.297204Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-05T21:52:23.842113Z digest=sha256:99f2f0c1b3866e8cc8d68f40556779b6a3826a95bfdf7809e8f9153711666ca2

Observation b702cb31-6238-41f4-b971-534b2127bb34 · outbound

This paper cites Coconet: Coupled contrastive learning network with multi-level feature ensemble for multi-modality image fusion.

MambaTrans: Multimodal Fusion Image Translation via Large Language Model Priors for Downstream Visual Tasks Coconet: Coupled contrastive learning network with multi-level feature ensemble for multi-modality image fusion

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T21:52:28.094732Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-05T21:52:23.934211Z digest=sha256:b8a620787a736809384835e2d28c6a75b801dbf80beed6c977c745880bd61399

Observation 572f25c5-9d32-4158-82de-22b71dfa0ef8 · outbound

This paper cites Contrastive learning for unpaired image-to-image translation.

MambaTrans: Multimodal Fusion Image Translation via Large Language Model Priors for Downstream Visual Tasks Contrastive learning for unpaired image-to-image translation

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T21:52:27.905015Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-05T21:52:23.996337Z digest=sha256:3f7aaa2a4cd4468773e1908dbbba5b3a35f46ce69a9d227de5638af9b4edd6d0

Observation e7609665-980b-4198-b359-3000148efe3b · outbound

This paper cites The pascal visual object classes challenge: A retrospective.

MambaTrans: Multimodal Fusion Image Translation via Large Language Model Priors for Downstream Visual Tasks The pascal visual object classes challenge: A retrospective

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T21:52:27.731790Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-05T21:52:24.109184Z digest=sha256:71a384e77ff6c007766778142f7493b48aa171d4d8322a53775bdc8e55ddfbf9

Observation b6d9f432-c6e3-46b5-8ed4-e3975836ac9b · outbound

This paper cites Assessment of image fusion procedures using entropy, image quality, and multispectral classification.

MambaTrans: Multimodal Fusion Image Translation via Large Language Model Priors for Downstream Visual Tasks Assessment of image fusion procedures using entropy, image quality, and multispectral classification

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T21:52:27.510761Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-05T21:52:24.196100Z digest=sha256:2f396d8f3b04c0788b2e9062eea4ba7fa10322c1b27eb8dec17fb80520c67973

Observation cc129a2d-01ec-4c29-800e-51208e30d266 · outbound

This paper cites Detail preserved fusion of visible and infrared images using regional saliency extraction and multi-scale image decomposition.

MambaTrans: Multimodal Fusion Image Translation via Large Language Model Priors for Downstream Visual Tasks Detail preserved fusion of visible and infrared images using regional saliency extraction and multi-scale image decomposition

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T21:52:27.340977Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-05T21:52:24.285735Z digest=sha256:ef6e740aa1c2ed08f5fd9d17a2a642ad3e3b82949e82cc7d7be08ed0463c6972

Observation b8565c7b-b79f-498e-b190-0ad0f5caacb8 · outbound

This paper cites Fsim: A feature similarity index for image quality assessment.

MambaTrans: Multimodal Fusion Image Translation via Large Language Model Priors for Downstream Visual Tasks Fsim: A feature similarity index for image quality assessment

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T21:52:27.130819Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-05T21:52:24.378057Z digest=sha256:a4bd68d4c38c08fafdb2b61a37dec75162b53cffa45a62ffe40898aa4ddc2a1b

Observation b5b73b88-d3eb-4bbd-91b5-d3110176138c · outbound

This paper cites Image fusion metric based on mutual information and tsallis entropy.

MambaTrans: Multimodal Fusion Image Translation via Large Language Model Priors for Downstream Visual Tasks Image fusion metric based on mutual information and tsallis entropy

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T21:52:26.953724Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-05T21:52:24.458774Z digest=sha256:0095acf975388c588080cf13bb1ac7229d063947bd0057f8d2667016d4245686

Observation b1156e03-b3a0-4d7c-a857-8c796da41b4c · outbound

This paper cites Image quality measures and their performance.

MambaTrans: Multimodal Fusion Image Translation via Large Language Model Priors for Downstream Visual Tasks Image quality measures and their performance

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T21:52:26.804297Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-05T21:52:24.550186Z digest=sha256:78aa53d2b7764fb6e6a25c83d237166e141d309485450c9e045b5eac54d1245c

Observation 5595f56a-d325-46f5-a2cc-d8ba3b43b6f7 · outbound

This paper cites Contrastive learning for unpaired image-to-image translation.

MambaTrans: Multimodal Fusion Image Translation via Large Language Model Priors for Downstream Visual Tasks Contrastive learning for unpaired image-to-image translation

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T21:52:26.584059Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-05T21:52:24.609982Z digest=sha256:65bbb39f674c67c938a36d9a7c058261aaec33d5143eae444769a7e3fb7bc41b

Observation 7d03129e-0337-429b-a8e3-b68856ba74bb · outbound

This paper cites Doubao vision pro 32k.

MambaTrans: Multimodal Fusion Image Translation via Large Language Model Priors for Downstream Visual Tasks Doubao vision pro 32k

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T21:52:26.427719Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-05T21:52:24.695539Z digest=sha256:9a4ee337f9b602dfdfae6429f63f2829b042ef1da23aa2bceccd9d5fce3f6a1d

Observation 0a875a7b-ce21-4c50-a7c2-56eff4607016 · outbound

This paper cites Fastinst: A simple query-based model for real-time instance segmentation.

MambaTrans: Multimodal Fusion Image Translation via Large Language Model Priors for Downstream Visual Tasks Fastinst: A simple query-based model for real-time instance segmentation

Reference 45

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T21:52:26.250516Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-05T21:52:24.779892Z digest=sha256:3a56a02ad27a76d1a782e8fbd2b656cf599ed6b75be5eb635eb8924f623f0c6f

Observation 85e2643c-66e2-4d01-9075-53beed274f0e · outbound

This paper cites Cascade r-cnn: Delving into high quality object detection.

MambaTrans: Multimodal Fusion Image Translation via Large Language Model Priors for Downstream Visual Tasks Cascade r-cnn: Delving into high quality object detection

Reference 46

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T21:52:26.089616Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-05T21:52:24.838186Z digest=sha256:db795895fab0ce6ebc19134c3fc4370aba6d86c9fef9786af9abfb3b7518aae7

Observation e28765f1-bead-406f-80c3-2e89ef03ec38 · outbound

This paper cites Mask dino: Towards a unified transformer-based framework for object detection and segmentation.

MambaTrans: Multimodal Fusion Image Translation via Large Language Model Priors for Downstream Visual Tasks Mask dino: Towards a unified transformer-based framework for object detection and segmentation

Reference 47

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T21:52:25.921270Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-05T21:52:24.927151Z digest=sha256:651322bfb333421e94e4773aef24f4475e0a09d3f6182fccd19c2157f2aebfc9

Observation 694e4564-e847-47f0-a216-ccae6b2ecc79 · outbound

This paper cites YOLOv11: An Overview of the Key Architectural Enhancements.

MambaTrans: Multimodal Fusion Image Translation via Large Language Model Priors for Downstream Visual Tasks YOLOv11: An Overview of the Key Architectural Enhancements

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-05T21:52:25.017468Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T21:52:25.017468Z digest=sha256:a83e1b20bff271c2bc75b193b9e666b5fc435c645ed86032121045c8ecff737b

Pith citing papers

Observation 8186a886-802d-47e0-bb27-ebe0460eea66 · inbound

InfraNet: Quality-Aware RGB Guidance for Efficient Infrared Object Detection cites this paper.

InfraNet: Quality-Aware RGB Guidance for Efficient Infrared Object Detection MambaTrans: Multimodal Fusion Image Translation via Large Language Model Priors for Downstream Visual Tasks

Reference 58

Resolution
unresolved
no resolver link, observed 2026-07-11T23:54:32.839280Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T23:54:32.839280Z digest=sha256:ce5b1f9c9ae41137ff1e3a3b445bb2284fd09346f4229390cc51fc71d431f2fe