Pith. sign in

Paper Citation Record · LEDGER

MambaTrans: Multimodal Fusion Image Translation via Large Language Model Priors for Downstream Visual Tasks

As of 18 August 2026, this Paper Citation Record lists 47 of 47 outbound references and 1 inbound Pith citation observation for arXiv:2508.07803.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2508.07803 v1

Coverage vector

measured 47 of 47 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-05T21:52:25.017468Z

measured 48 of 48 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-18T06:34:40.430872+00:00

measured 1 of 1 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-07-11T23:54:32.839280Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

47 of 47 outbound references displayed

  • verified exact2
  • verified fuzzy38
  • unresolved7
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 168e4721-7e28-416b-8eea-e6e21ab8983a · outbound

This paper cites A deep learning framework for infrared and visible image fusion without strict registration.

MambaTrans: Multimodal Fusion Image Translation via Large Language Model Priors for Downstream Visual Tasks A deep learning framework for infrared and visible image fusion without strict registration

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T21:52:33.914320Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-05T21:52:20.843522Z digest=sha256:5b748e75fe7b0facb12cad7928c71ffa309d033c8ec1df80a6954e0f2a5ff292

Observation a5777d18-1a03-47f1-8b16-353c70a1a5b0 · outbound

This paper cites All-weather multi-modality image fusion: Unified framework and 100k benchmark.

MambaTrans: Multimodal Fusion Image Translation via Large Language Model Priors for Downstream Visual Tasks All-weather multi-modality image fusion: Unified framework and 100k benchmark

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-05T21:52:20.944258Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T21:52:20.944258Z digest=sha256:a7d1d8a8e9509b2cd813944dc5d0a872e1a2d4dbe5a8165cf39b3102fc285bba

Observation 50ddb353-3e72-4768-adcb-9ec1833028c3 · outbound

This paper cites Infrared and visible image fusion based on domain transform filtering and sparse representation.

MambaTrans: Multimodal Fusion Image Translation via Large Language Model Priors for Downstream Visual Tasks Infrared and visible image fusion based on domain transform filtering and sparse representation

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T21:52:33.758344Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-05T21:52:21.052336Z digest=sha256:18402017c29859372cc70178aad4835c9ad5f640c17694bff918f4d3fb79ace3

Observation f646ed0c-ad4d-4340-860a-01556952b60f · outbound

This paper cites Infrared and visible image fusion: From data compatibility to task adaption.

MambaTrans: Multimodal Fusion Image Translation via Large Language Model Priors for Downstream Visual Tasks Infrared and visible image fusion: From data compatibility to task adaption

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T21:52:33.536389Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-05T21:52:21.160857Z digest=sha256:3c0a7d2b3f02231b94bc406336ad5e8b681f03ae7dc614e543a9aa1d88c71a0e

Observation 9e8b1584-9e0b-4dbf-8c83-bc5edf9765eb · outbound

This paper cites Simultaneous tri-modal medical image fusion and super-resolution using conditional diffusion model.

MambaTrans: Multimodal Fusion Image Translation via Large Language Model Priors for Downstream Visual Tasks Simultaneous tri-modal medical image fusion and super-resolution using conditional diffusion model

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T21:52:33.295541Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-05T21:52:21.259991Z digest=sha256:f9afa66db895b35133ebfd1e3b5aff8dbfc45a18752ae39ea531ab3d0b43586b

Observation f2d9ac42-ac80-4ce0-bf67-5b308133d718 · outbound

This paper cites Fast r-cnn.

MambaTrans: Multimodal Fusion Image Translation via Large Language Model Priors for Downstream Visual Tasks Fast r-cnn

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T21:52:33.066319Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-05T21:52:21.481919Z digest=sha256:bf38d59ad7296bd2a9e0c1309eefbe5cbb0ccfba03795d6d0b14cf925a027f91

Observation f8ae765a-2b45-451b-a981-b47b86160e2c · outbound

This paper cites Microsoft coco: Common objects in context.

MambaTrans: Multimodal Fusion Image Translation via Large Language Model Priors for Downstream Visual Tasks Microsoft coco: Common objects in context

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T21:52:32.837681Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-05T21:52:21.572846Z digest=sha256:edc9acf51ee44e7f9fea3ca127a25220605674170541dcae040da2416bbe311d

Observation 95698043-7484-408d-aca9-888a1a635f15 · outbound

This paper cites Visualizing data using t-sne.

MambaTrans: Multimodal Fusion Image Translation via Large Language Model Priors for Downstream Visual Tasks Visualizing data using t-sne

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-05T21:52:21.713153Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T21:52:21.713153Z digest=sha256:272a9acdf88234de19812066e8962bc7643cdbda3fadffa1855582b63ac71647

Observation 3dc53547-8931-4a95-a8f5-f46167226cb4 · outbound

This paper cites Tsjnet: A multi-modality target and semantic awareness joint-driven image fusion network.

MambaTrans: Multimodal Fusion Image Translation via Large Language Model Priors for Downstream Visual Tasks Tsjnet: A multi-modality target and semantic awareness joint-driven image fusion network

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-05T21:52:21.795023Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T21:52:21.795023Z digest=sha256:5d4b41da68f76e123024762f56adaa4e42ab3d45170c8d0f4c0b199b7b066103

Observation 5957c9c7-fda6-4c8a-8295-3b4d8b2673e0 · outbound

This paper cites Target-aware dual adversarial learning and a multi-scenario multi-modality benchmark to fuse infrared and visible for object detection.

MambaTrans: Multimodal Fusion Image Translation via Large Language Model Priors for Downstream Visual Tasks Target-aware dual adversarial learning and a multi-scenario multi-modality benchmark to fuse infrared and visible for object detection

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T21:52:32.653248Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-05T21:52:21.874881Z digest=sha256:ad9e49364cf94ac63ca98fe4a02ad80956820ac2195d44c6a3df441c5e22b285

Observation 0baafcb8-965e-4a87-9a73-8d1f3a5b69ed · outbound

This paper cites Fs-diff: Semantic guidance and clarity-aware simultaneous multimodal image fusion and super-resolution.

MambaTrans: Multimodal Fusion Image Translation via Large Language Model Priors for Downstream Visual Tasks Fs-diff: Semantic guidance and clarity-aware simultaneous multimodal image fusion and super-resolution

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T21:52:32.377319Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-05T21:52:21.983819Z digest=sha256:66c598dd5c0d192addfd8a25dc77549e1983714c81f945bf5457353720a07839

Observation f2a1e272-e385-4dd3-bc79-ab46406dff3e · outbound

This paper cites A task-guided, implicitly-searched and metainitialized deep model for image fusion.

MambaTrans: Multimodal Fusion Image Translation via Large Language Model Priors for Downstream Visual Tasks A task-guided, implicitly-searched and metainitialized deep model for image fusion

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T21:52:32.242375Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-05T21:52:22.083478Z digest=sha256:9086c71db5f98b000ae59662f0929ba476138ecbc1b1ae624e5ce76515b21eed

Observation 330debc3-4de8-4489-87c9-cdf7d6008795 · outbound

This paper cites Image fusion in the loop of high-level vision tasks: A semantic-aware real-time infrared and visible image fusion network.

MambaTrans: Multimodal Fusion Image Translation via Large Language Model Priors for Downstream Visual Tasks Image fusion in the loop of high-level vision tasks: A semantic-aware real-time infrared and visible image fusion network

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T21:52:32.072413Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-05T21:52:22.190602Z digest=sha256:94fced56ba380f4cb21e562fe2e7ecba462fa9fe03263a03d3c75dd9bb309147

Observation e1cbd90c-fd38-4701-8161-0ce3eb1aa078 · outbound

This paper cites Mrfs: Mutually reinforcing image fusion and segmentation.

MambaTrans: Multimodal Fusion Image Translation via Large Language Model Priors for Downstream Visual Tasks Mrfs: Mutually reinforcing image fusion and segmentation

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T21:52:31.755632Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-05T21:52:22.300125Z digest=sha256:2b52db4019e15a0e27e46236af66065e4768958b7c33b13b5b2a0858f78b78d7

Observation 3fa7d815-f3a3-414c-9f10-8e61c88de446 · outbound

This paper cites Image-to-image translation: Methods and applications.

MambaTrans: Multimodal Fusion Image Translation via Large Language Model Priors for Downstream Visual Tasks Image-to-image translation: Methods and applications

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T21:52:31.569164Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-05T21:52:22.415638Z digest=sha256:5c076d7f81f4d0be8706ca0ff90c6381ee50d3c59c588adce6a41f55c0e0873b

Observation 628f808a-9b94-4aa0-9ef2-77bf7ac7885a · outbound

This paper cites Source-free open compound domain adaptation in semantic segmentation.

MambaTrans: Multimodal Fusion Image Translation via Large Language Model Priors for Downstream Visual Tasks Source-free open compound domain adaptation in semantic segmentation

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T21:52:31.357536Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-05T21:52:22.499013Z digest=sha256:03ef019c896f58b1d6556165a93bb85c1f0278365b79c0b86b4df4772ba7e54c

Observation 0d2eabd2-5c76-49ee-ac91-80f5a3b7dafb · outbound

This paper cites Infragan: A gan architecture to transfer visible images to infrared domain.

MambaTrans: Multimodal Fusion Image Translation via Large Language Model Priors for Downstream Visual Tasks Infragan: A gan architecture to transfer visible images to infrared domain

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T21:52:31.076644Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-05T21:52:22.571106Z digest=sha256:1e1c96658662ab0081b8d9150c15796946fba0505f524805796fbfad22bff8d6

Observation 37b2628c-cfeb-4652-a9a3-197fa0478a0b · outbound

This paper cites Cnn-based thermal infrared person detection by domain adaptation.

MambaTrans: Multimodal Fusion Image Translation via Large Language Model Priors for Downstream Visual Tasks Cnn-based thermal infrared person detection by domain adaptation

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T21:52:30.887222Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-05T21:52:22.633845Z digest=sha256:7da66e984469cc1cf56885a9ab13b112484c1f374c46564ee493da3e4ad73a28

Observation 56bfd79f-6e98-45f7-974b-c32740ef96a2 · outbound

This paper cites Hallucidet: hallucinating rgb modality for person detection through privileged information.

MambaTrans: Multimodal Fusion Image Translation via Large Language Model Priors for Downstream Visual Tasks Hallucidet: hallucinating rgb modality for person detection through privileged information

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T21:52:30.699864Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-05T21:52:22.737979Z digest=sha256:291e00c2811c11870b452af7265429f3b7531f4a7331f9c3755e51fdf1195c2b

Observation 7a470abe-68f6-4c2d-ac65-ed82df8276a5 · outbound

This paper cites Modality translation for object detection adaptation without forgetting prior knowledge.

MambaTrans: Multimodal Fusion Image Translation via Large Language Model Priors for Downstream Visual Tasks Modality translation for object detection adaptation without forgetting prior knowledge

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T21:52:30.456942Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-05T21:52:22.812104Z digest=sha256:4a425981666d716c26c0f94bc470900415036ffd07f3215a97c005758408f56e

Observation 42e047fa-8044-48e9-8f00-d3fc651b7201 · outbound

This paper cites Mamba: Linear-Time Sequence Modeling with Selective State Spaces.

MambaTrans: Multimodal Fusion Image Translation via Large Language Model Priors for Downstream Visual Tasks Mamba: Linear-Time Sequence Modeling with Selective State Spaces

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-05T21:52:22.912267Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T21:52:22.912267Z digest=sha256:ba9c53f8a068258bd3049043b30bcc49b1aa80803ccebcaebbadc4332e881bd6

Observation a5d5d2c5-e078-4562-b432-69a715a0aae2 · outbound

This paper cites Vmamba: Visual state space model.

MambaTrans: Multimodal Fusion Image Translation via Large Language Model Priors for Downstream Visual Tasks Vmamba: Visual state space model

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T21:52:30.269904Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-05T21:52:22.998893Z digest=sha256:721b70f829490212a2419a9691e013226a443c9f5d3b8aa8af1f7f99f0f2d3f1

Observation e0502ec9-18c2-4fc5-8539-79198965130f · outbound

This paper cites Mambair: A simple baseline for image restoration with state-space model.

MambaTrans: Multimodal Fusion Image Translation via Large Language Model Priors for Downstream Visual Tasks Mambair: A simple baseline for image restoration with state-space model

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T21:52:30.002559Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-05T21:52:23.077544Z digest=sha256:ee184d980358410febc575483be1855dbe5d47bb6fc1e22226ef2a34bedac41e

Observation 324bcfbb-487b-47e7-ae1e-64013fd9f8a8 · outbound

This paper cites Vision mamba: efficient visual representation learning with bidirectional state space model.

MambaTrans: Multimodal Fusion Image Translation via Large Language Model Priors for Downstream Visual Tasks Vision mamba: efficient visual representation learning with bidirectional state space model

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T21:52:29.682263Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-05T21:52:23.182088Z digest=sha256:6c79b1685f51f6209a441a0bd301bbee30e67bfc481f1c7758e1b3c7deb6253a

Observation 6e766b80-804f-4162-839a-63acd56fd705 · outbound

This paper cites GLU Variants Improve Transformer.

MambaTrans: Multimodal Fusion Image Translation via Large Language Model Priors for Downstream Visual Tasks GLU Variants Improve Transformer

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-05T21:52:23.234051Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T21:52:23.234051Z digest=sha256:8b90697b58f5af83748d9c58dbb18f720abd86d045e9f97a1353aa9ed5b49bbc

Observation 0a9be89a-6367-4f9a-82d8-b1f73f844dfc · outbound

This paper cites Faster r-cnn: Towards real-time object detection with region proposal networks.

MambaTrans: Multimodal Fusion Image Translation via Large Language Model Priors for Downstream Visual Tasks Faster r-cnn: Towards real-time object detection with region proposal networks

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-05T21:52:23.332065Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T21:52:23.332065Z digest=sha256:7fe4c5a7666ef424221501df43d99fa59ff680bb16f93a7f3175789c1bb35fae

Observation d099b7d5-f44e-47f0-8b07-4acc0cfe7479 · outbound

This paper cites Piafusion: A progressive infrared and visible image fusion network based on illumination aware.

MambaTrans: Multimodal Fusion Image Translation via Large Language Model Priors for Downstream Visual Tasks Piafusion: A progressive infrared and visible image fusion network based on illumination aware

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T21:52:29.362717Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-05T21:52:23.431911Z digest=sha256:c6fbbee9842fdb1ed739f0e11278d5b80fa6ae16f206c6d4f7dbbd4e90f0fbc9

Observation 5fa405c6-2f88-4b12-aa1f-b3ab263b726c · outbound

This paper cites Every SAM Drop Counts: Embracing Semantic Priors for Multi-Modality Image Fusion and Beyond.

MambaTrans: Multimodal Fusion Image Translation via Large Language Model Priors for Downstream Visual Tasks Every SAM Drop Counts: Embracing Semantic Priors for Multi-Modality Image Fusion and Beyond

Reference 29

Resolution
verified exact
local_arxiv, observed 2026-08-05T21:52:25.307217Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-05T21:52:23.523944Z digest=sha256:e8c095e71e3ef7b9fce6e53f6a712bc8ebf7f55d7dfc30d7cf3cb423d144e5d7

Observation c6a96941-b04a-45d1-b68f-934966f34a9e · outbound

This paper cites DCEvo: Discriminative Cross-Dimensional Evolutionary Learning for Infrared and Visible Image Fusion.

MambaTrans: Multimodal Fusion Image Translation via Large Language Model Priors for Downstream Visual Tasks DCEvo: Discriminative Cross-Dimensional Evolutionary Learning for Infrared and Visible Image Fusion

Reference 30

Resolution
verified exact
local_arxiv, observed 2026-08-05T21:52:25.159340Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-05T21:52:23.586355Z digest=sha256:6f8970980002b004f08e61a95d094d8294cb51cd8f7ae3f1861afd8bb3981edf

Observation a4137315-f119-4e1f-854a-992dc18a4587 · outbound

This paper cites Where elegance meets precision: towards a compact, automatic, and flexible framework for multi-modality image fusion and applications.

MambaTrans: Multimodal Fusion Image Translation via Large Language Model Priors for Downstream Visual Tasks Where elegance meets precision: towards a compact, automatic, and flexible framework for multi-modality image fusion and applications

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T21:52:29.079716Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-05T21:52:23.647965Z digest=sha256:b8346f4fcd6cb74002143225d272599e25afb8732ccd4ac4cd53c2b01569ce11

Observation 5dc11091-632f-42ca-ac3b-9c475a758d04 · outbound

This paper cites Equivariant multi-modality image fusion.

MambaTrans: Multimodal Fusion Image Translation via Large Language Model Priors for Downstream Visual Tasks Equivariant multi-modality image fusion

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T21:52:28.797720Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-05T21:52:23.710572Z digest=sha256:3c9cdac3c6267e2b40a79c633b3ee25721b06bd77c3b270ad92c2dcc45cdf58a

Observation 255a2d7b-c502-452d-b0b0-523e20e56bb5 · outbound

This paper cites Cddfuse: Correlation-driven dual-branch feature decomposition for multi-modality image fusion.

MambaTrans: Multimodal Fusion Image Translation via Large Language Model Priors for Downstream Visual Tasks Cddfuse: Correlation-driven dual-branch feature decomposition for multi-modality image fusion

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T21:52:28.529331Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-05T21:52:23.786134Z digest=sha256:394fd36f3f2f4985a133fbd275d4d0f74a3cfb958d39032e10a10698f4ac10e5

Observation 6273535e-be8e-4d07-85b2-fb4cdc44f161 · outbound

This paper cites Probing synergistic high-order interaction in infrared and visible image fusion.

MambaTrans: Multimodal Fusion Image Translation via Large Language Model Priors for Downstream Visual Tasks Probing synergistic high-order interaction in infrared and visible image fusion

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T21:52:28.297204Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-05T21:52:23.842113Z digest=sha256:e7742197cadc9e2b9633da0b8ce9ab4c40e5ec3d7c015e8d25c77f9e6e32f267

Observation b702cb31-6238-41f4-b971-534b2127bb34 · outbound

This paper cites Coconet: Coupled contrastive learning network with multi-level feature ensemble for multi-modality image fusion.

MambaTrans: Multimodal Fusion Image Translation via Large Language Model Priors for Downstream Visual Tasks Coconet: Coupled contrastive learning network with multi-level feature ensemble for multi-modality image fusion

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T21:52:28.094732Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-05T21:52:23.934211Z digest=sha256:43d6f4191913a99979231357222ba4365bd2dd73b45f6faed10cb26012c5ff05

Observation 572f25c5-9d32-4158-82de-22b71dfa0ef8 · outbound

This paper cites Contrastive learning for unpaired image-to-image translation.

MambaTrans: Multimodal Fusion Image Translation via Large Language Model Priors for Downstream Visual Tasks Contrastive learning for unpaired image-to-image translation

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T21:52:27.905015Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-05T21:52:23.996337Z digest=sha256:c6743a75515a1dade4565a476c1608173cd91a8dc9aad4f0fff895a2e1d01bd2

Observation e7609665-980b-4198-b359-3000148efe3b · outbound

This paper cites The pascal visual object classes challenge: A retrospective.

MambaTrans: Multimodal Fusion Image Translation via Large Language Model Priors for Downstream Visual Tasks The pascal visual object classes challenge: A retrospective

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T21:52:27.731790Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-05T21:52:24.109184Z digest=sha256:c4909e2abfa717b3aa9a2b76ee4b984fb707b09e5e7f23a1d2e92df67933e141

Observation b6d9f432-c6e3-46b5-8ed4-e3975836ac9b · outbound

This paper cites Assessment of image fusion procedures using entropy, image quality, and multispectral classification.

MambaTrans: Multimodal Fusion Image Translation via Large Language Model Priors for Downstream Visual Tasks Assessment of image fusion procedures using entropy, image quality, and multispectral classification

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T21:52:27.510761Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-05T21:52:24.196100Z digest=sha256:6e5c97174111aed42f57b6890ba918d0de7470abc1f960b41c1df26e0965806a

Observation cc129a2d-01ec-4c29-800e-51208e30d266 · outbound

This paper cites Detail preserved fusion of visible and infrared images using regional saliency extraction and multi-scale image decomposition.

MambaTrans: Multimodal Fusion Image Translation via Large Language Model Priors for Downstream Visual Tasks Detail preserved fusion of visible and infrared images using regional saliency extraction and multi-scale image decomposition

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T21:52:27.340977Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-05T21:52:24.285735Z digest=sha256:3e19209feb1c54f1f17a3586760934428e12a3bba9ce2bbffb1f56c82731bc07

Observation b8565c7b-b79f-498e-b190-0ad0f5caacb8 · outbound

This paper cites Fsim: A feature similarity index for image quality assessment.

MambaTrans: Multimodal Fusion Image Translation via Large Language Model Priors for Downstream Visual Tasks Fsim: A feature similarity index for image quality assessment

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T21:52:27.130819Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-05T21:52:24.378057Z digest=sha256:d93b0f27e669255239e24f6735e9d2c533d7238e2db057c477658dab844c135c

Observation b5b73b88-d3eb-4bbd-91b5-d3110176138c · outbound

This paper cites Image fusion metric based on mutual information and tsallis entropy.

MambaTrans: Multimodal Fusion Image Translation via Large Language Model Priors for Downstream Visual Tasks Image fusion metric based on mutual information and tsallis entropy

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T21:52:26.953724Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-05T21:52:24.458774Z digest=sha256:bbe63f81ec2fb5b23d50d5bf4df88b031394c47034d56f92ffd36012ac2d142b

Observation b1156e03-b3a0-4d7c-a857-8c796da41b4c · outbound

This paper cites Image quality measures and their performance.

MambaTrans: Multimodal Fusion Image Translation via Large Language Model Priors for Downstream Visual Tasks Image quality measures and their performance

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T21:52:26.804297Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-05T21:52:24.550186Z digest=sha256:bc7fbdede70d656d31061ec74c018a5db2779f3fdd5b12e7b86ac6d0d72d59a0

Observation 5595f56a-d325-46f5-a2cc-d8ba3b43b6f7 · outbound

This paper cites Contrastive learning for unpaired image-to-image translation.

MambaTrans: Multimodal Fusion Image Translation via Large Language Model Priors for Downstream Visual Tasks Contrastive learning for unpaired image-to-image translation

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T21:52:26.584059Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-05T21:52:24.609982Z digest=sha256:f10e3187425f7344ba62e8c41c30ef363caf02331a5b99795f33d56f97575926

Observation 7d03129e-0337-429b-a8e3-b68856ba74bb · outbound

This paper cites Doubao vision pro 32k.

MambaTrans: Multimodal Fusion Image Translation via Large Language Model Priors for Downstream Visual Tasks Doubao vision pro 32k

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T21:52:26.427719Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-05T21:52:24.695539Z digest=sha256:2515bab02c5d75553e8994d91b19847ed16bcb306164f2cd7463093b520106e8

Observation 0a875a7b-ce21-4c50-a7c2-56eff4607016 · outbound

This paper cites Fastinst: A simple query-based model for real-time instance segmentation.

MambaTrans: Multimodal Fusion Image Translation via Large Language Model Priors for Downstream Visual Tasks Fastinst: A simple query-based model for real-time instance segmentation

Reference 45

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T21:52:26.250516Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-05T21:52:24.779892Z digest=sha256:fe1e83d7ece9f3490206b12eb6a0e7f9d23b740a89224befda797fc9415d42d8

Observation 85e2643c-66e2-4d01-9075-53beed274f0e · outbound

This paper cites Cascade r-cnn: Delving into high quality object detection.

MambaTrans: Multimodal Fusion Image Translation via Large Language Model Priors for Downstream Visual Tasks Cascade r-cnn: Delving into high quality object detection

Reference 46

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T21:52:26.089616Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-05T21:52:24.838186Z digest=sha256:b5859231c94b55793745b44db11dea84d392ada5e8a9eeb4da84f9cdda676d39

Observation e28765f1-bead-406f-80c3-2e89ef03ec38 · outbound

This paper cites Mask dino: Towards a unified transformer-based framework for object detection and segmentation.

MambaTrans: Multimodal Fusion Image Translation via Large Language Model Priors for Downstream Visual Tasks Mask dino: Towards a unified transformer-based framework for object detection and segmentation

Reference 47

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T21:52:25.921270Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-05T21:52:24.927151Z digest=sha256:d6b8effe3a489325edd41dee5f54db497f5824f4bd8bba67360010e1ad91b75f

Observation 694e4564-e847-47f0-a216-ccae6b2ecc79 · outbound

This paper cites YOLOv11: An Overview of the Key Architectural Enhancements.

MambaTrans: Multimodal Fusion Image Translation via Large Language Model Priors for Downstream Visual Tasks YOLOv11: An Overview of the Key Architectural Enhancements

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-05T21:52:25.017468Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T21:52:25.017468Z digest=sha256:a176f8525c637df046eb5824fb868b101a0d511071a529189357c084e1903809

Pith citing papers

Observation 8186a886-802d-47e0-bb27-ebe0460eea66 · inbound

InfraNet: Quality-Aware RGB Guidance for Efficient Infrared Object Detection cites this paper.

InfraNet: Quality-Aware RGB Guidance for Efficient Infrared Object Detection MambaTrans: Multimodal Fusion Image Translation via Large Language Model Priors for Downstream Visual Tasks

Reference 58

Resolution
unresolved
no resolver link, observed 2026-07-11T23:54:32.839280Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T23:54:32.839280Z digest=sha256:e969c59f36e0a438de64470b5cddc17acc2b7710fe159a487c98ec098877e275