Pith. sign in

Paper Citation Record · LEDGER

Bridging Domain Generalization to Multimodal Domain Generalization via Unified Representations

As of 15 August 2026, this Paper Citation Record lists 58 of 58 outbound references and 2 inbound Pith citation observations for arXiv:2507.03304.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2507.03304 v1

Coverage vector

measured 58 of 58 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-06T20:17:33.743609Z

measured 60 of 60 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-15T06:32:42.880941+00:00

measured 2 of 2 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-06T15:48:28.447490Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-08-05T21:53:33.960468Z

Reference resolution

58 of 58 outbound references displayed

  • verified exact1
  • verified fuzzy35
  • unresolved22
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 367629b1-06fd-44bf-8af8-a12ad636b06d · outbound

This paper cites Robust cross-modal representation learning with progressive self- distillation.

Bridging Domain Generalization to Multimodal Domain Generalization via Unified Representations Robust cross-modal representation learning with progressive self- distillation

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-06T20:17:30.662544Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:17:30.662544Z digest=sha256:67293c4c34218fc7352b54d26ff468517288d0b4cc032f04797198ac39a5cfce

Observation 3c7174f9-cb10-461d-8361-32a23f45ea8e · outbound

This paper cites Person30k: A dual-meta general- ization network for person re-identification.

Bridging Domain Generalization to Multimodal Domain Generalization via Unified Representations Person30k: A dual-meta general- ization network for person re-identification

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:17:34.311524Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-06T20:17:30.714796Z digest=sha256:cf2983b714e910998ad03a86d07ef3ecbccbd9e4507ad26a7d0aff24bbeb2d78

Observation 625d2206-bb91-4a9e-b088-a9f8acd006fd · outbound

This paper cites Ex- ploiting domain-specific features to enhance domain gener- alization.

Bridging Domain Generalization to Multimodal Domain Generalization via Unified Representations Ex- ploiting domain-specific features to enhance domain gener- alization

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:17:34.302108Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-06T20:17:30.817961Z digest=sha256:731e90062cef6233f9fcc18f9634a7b2681fa22abfa288f70e3b015cc8ef660a

Observation c46aa49c-7926-4040-aacf-81d74b420993 · outbound

This paper cites Domain generalization by solving jigsaw puzzles.

Bridging Domain Generalization to Multimodal Domain Generalization via Unified Representations Domain generalization by solving jigsaw puzzles

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:17:34.293311Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-06T20:17:30.939942Z digest=sha256:6a131fe1c62178dd8fcebfad76ad9abb0c642921084613092d49a3c34b1d8fe5

Observation 83fe79c4-12a6-4c3f-8875-4dce30852bc1 · outbound

This paper cites Vggsound: A large-scale audio-visual dataset.

Bridging Domain Generalization to Multimodal Domain Generalization via Unified Representations Vggsound: A large-scale audio-visual dataset

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:17:34.285032Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-06T20:17:31.021925Z digest=sha256:9c434b2f13c18d2c75159aeb4c71f837233c375cb08608af3b7435bdcabe597f

Observation ca91e6ad-7b25-4850-ba6a-1cd22ec982d5 · outbound

This paper cites Uniter: Universal image-text representation learning.

Bridging Domain Generalization to Multimodal Domain Generalization via Unified Representations Uniter: Universal image-text representation learning

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-06T20:17:31.100518Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:17:31.100518Z digest=sha256:7edaf87880f12f276be5c77a56604e4ff7792873c9b4c9c190ed8b59cdfe5943

Observation d2284a4b-3ff1-4c65-97d7-a6ee1734dc99 · outbound

This paper cites Club: A contrastive log-ratio up- per bound of mutual information.

Bridging Domain Generalization to Multimodal Domain Generalization via Unified Representations Club: A contrastive log-ratio up- per bound of mutual information

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:17:34.271871Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-06T20:17:31.162298Z digest=sha256:893c9f7f0add14b756293cb02cb7782ea142b494fb92565dc276398de21c6c6c

Observation 7235d909-7359-4f44-a604-e0d0da7610e4 · outbound

This paper cites Robustnet: Improving domain generalization in urban-scene segmentation via in- stance selective whitening.

Bridging Domain Generalization to Multimodal Domain Generalization via Unified Representations Robustnet: Improving domain generalization in urban-scene segmentation via in- stance selective whitening

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:17:34.262285Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-06T20:17:31.230593Z digest=sha256:8140861b9bc7932e2b9f2e5388c93569aa5b3e9a4fbb0fc02be4db3e0a7c0108

Observation 76568704-939d-4ba7-b555-2531e143678e · outbound

This paper cites Openmmlab’s next generation video understanding toolbox and benchmark.

Bridging Domain Generalization to Multimodal Domain Generalization via Unified Representations Openmmlab’s next generation video understanding toolbox and benchmark

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:17:34.253281Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-06T20:17:31.322944Z digest=sha256:8ca4e801ea47e13d5646d3aabecad7a6a54b7edd2a86c8fa94b24d36d74b6021

Observation 0d48e055-6574-4c8c-9a5b-3142ba062454 · outbound

This paper cites Scaling egocentric vision: The epic-kitchens dataset.

Bridging Domain Generalization to Multimodal Domain Generalization via Unified Representations Scaling egocentric vision: The epic-kitchens dataset

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:17:34.243179Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-06T20:17:31.387598Z digest=sha256:f9631078d9c63a88fcc0e85f6e46c9d71df4c0582822e8a66ebd01e212c8b6dd

Observation d1ec1c8e-4551-4c30-8080-8f31d01b84cd · outbound

This paper cites Simmmdg: A simple and effective framework for multi-modal domain generalization.

Bridging Domain Generalization to Multimodal Domain Generalization via Unified Representations Simmmdg: A simple and effective framework for multi-modal domain generalization

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:17:34.234404Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-06T20:17:31.494194Z digest=sha256:17d3324fb25375bb63ff571a785824b569049c10115b5c2ba8da4f3d706b2dc2

Observation 8136cfee-7770-41b1-be7e-ae6e20088cf3 · outbound

This paper cites Towards mul- timodal open-set domain generalization and adaptation through self-supervision.

Bridging Domain Generalization to Multimodal Domain Generalization via Unified Representations Towards mul- timodal open-set domain generalization and adaptation through self-supervision

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:17:34.224985Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-06T20:17:31.530990Z digest=sha256:94994a1f1a947ddba18290fd917a157f57a22bdd2e0ac4cd55c030aacf08e0bc

Observation 7551c2b5-df3c-4027-b81d-a1c2b3935fed · outbound

This paper cites Multi-modal align- ment using representation codebook.

Bridging Domain Generalization to Multimodal Domain Generalization via Unified Representations Multi-modal align- ment using representation codebook

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:17:34.216171Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-06T20:17:31.597009Z digest=sha256:7fb4939f4a5ebf138d07dd539492630c0f1f660a16aab75afcd7c2c0baccaeb7

Observation 054b1fa6-6cf8-47af-8382-67846ed614dc · outbound

This paper cites Cross-modal representation flattening for multi-modal do- main generalization.

Bridging Domain Generalization to Multimodal Domain Generalization via Unified Representations Cross-modal representation flattening for multi-modal do- main generalization

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:17:34.208054Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-06T20:17:31.658418Z digest=sha256:80d7fc7dabdcc936fc495253670ef326f95a7931049470c4cecf40350aac81d9

Observation 40d239d3-9cd2-4cac-a313-79ab721ea3d0 · outbound

This paper cites Ace: A generative cross-modal retrieval framework with coarse-to-fine semantic modeling.

Bridging Domain Generalization to Multimodal Domain Generalization via Unified Representations Ace: A generative cross-modal retrieval framework with coarse-to-fine semantic modeling

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-06T20:17:31.731723Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:17:31.731723Z digest=sha256:2d2d78edf613d51204b3784efaa9236a1fca6968e58454000cac4950abfa2a17

Observation 32bfe891-36d2-4b79-a35d-3dd57f575f5d · outbound

This paper cites Slowfast networks for video recognition.

Bridging Domain Generalization to Multimodal Domain Generalization via Unified Representations Slowfast networks for video recognition

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-06T20:17:31.828849Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:17:31.828849Z digest=sha256:52876155c906cab80b8adf38bd37ac0baf47ca1e22888a168b0b6989da417e86

Observation 5aab43cb-8fb3-423a-9c37-4e3974d0e37e · outbound

This paper cites Sharpness-Aware Minimization for Efficiently Improving Generalization.

Bridging Domain Generalization to Multimodal Domain Generalization via Unified Representations Sharpness-Aware Minimization for Efficiently Improving Generalization

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-06T20:17:31.994571Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:17:31.994571Z digest=sha256:63bc8240a50208c8544d1630aa3d8e906ebe1b5edad390818ae4a81aefaedc1d

Observation a4ae10b5-14ce-4b08-8abd-2159760e5ae9 · outbound

This paper cites Domain-adversarial training of neural networks.

Bridging Domain Generalization to Multimodal Domain Generalization via Unified Representations Domain-adversarial training of neural networks

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:17:34.194868Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-06T20:17:32.116201Z digest=sha256:cf45e8d610b0c80ea23307648816e34672199636f4b5a621a7a6878fa4a179f9

Observation 62b183e1-9bb4-4426-adb8-5d3a9d5dc224 · outbound

This paper cites Imagebind: One embedding space to bind them all.

Bridging Domain Generalization to Multimodal Domain Generalization via Unified Representations Imagebind: One embedding space to bind them all

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-06T20:17:32.233001Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:17:32.233001Z digest=sha256:46b7c88a6fd63994227db5ef90e0ae20fa49140dfedb81f79b52a6876d2ab68e

Observation 7f52017b-2b61-49a6-8f99-1983ff533651 · outbound

This paper cites Learning Shared Semantic Space for Speech-to-Text Translation.

Bridging Domain Generalization to Multimodal Domain Generalization via Unified Representations Learning Shared Semantic Space for Speech-to-Text Translation

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-06T20:17:32.349703Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:17:32.349703Z digest=sha256:dd58fd20c39cdf1d6802630a711da21dfcfa907b6c66c5ffd29904e8d8e35754

Observation 672a4428-19a5-42d5-80db-97b04e360d6a · outbound

This paper cites Mixgen: A new multi- modal data augmentation.

Bridging Domain Generalization to Multimodal Domain Generalization via Unified Representations Mixgen: A new multi- modal data augmentation

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:17:34.183265Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-06T20:17:32.475160Z digest=sha256:c0612f23192f41740176cca0d74a057e9441c587b8d61c72b6222ff40b3c3361

Observation 72a52d41-58aa-433a-860e-cda34ffc7a6a · outbound

This paper cites Deep residual learning for image recognition.

Bridging Domain Generalization to Multimodal Domain Generalization via Unified Representations Deep residual learning for image recognition

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-06T20:17:32.574686Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:17:32.574686Z digest=sha256:4dc1f09ceecbf0572e12cfbca74ded0abed8ec0c8af2d3d1f9b0ca6549987702

Observation d5116977-16d3-4b68-b859-77414f9c3c14 · outbound

This paper cites Enhancing Multimodal Unified Representations for Cross Modal Generalization.

Bridging Domain Generalization to Multimodal Domain Generalization via Unified Representations Enhancing Multimodal Unified Representations for Cross Modal Generalization

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-06T20:17:32.719103Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:17:32.719103Z digest=sha256:d64b6e7618b519db42aff42949ada9ae51e9c0d057a554113bd9b433b849beb6

Observation e332a5bf-2105-4644-b886-cbb5985d958b · outbound

This paper cites Semantic residual for multimodal unified discrete representation.

Bridging Domain Generalization to Multimodal Domain Generalization via Unified Representations Semantic residual for multimodal unified discrete representation

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:17:34.170036Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-06T20:17:32.834039Z digest=sha256:9ec452b79bd1c803abd484f41f7e37f3b9e2f713a975e4a2c806ecb8f7e07af0

Observation 1c6a9cf2-7ba9-45a6-b35f-a0b2098824d7 · outbound

This paper cites Overcoming both domain shift and label shift for referring video segmentation.

Bridging Domain Generalization to Multimodal Domain Generalization via Unified Representations Overcoming both domain shift and label shift for referring video segmentation

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:17:34.160480Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-06T20:17:32.986054Z digest=sha256:9992b109c7029cce77c214ad1b3726d93725b7b83bfe6194b9bc36b4754cc529

Observation 1d58d270-e56e-480d-8f6c-e2db95288fdc · outbound

This paper cites Modality competition: What makes joint training of multi-modal network fail in deep learn- ing?(provably).

Bridging Domain Generalization to Multimodal Domain Generalization via Unified Representations Modality competition: What makes joint training of multi-modal network fail in deep learn- ing?(provably)

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:17:34.150997Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-06T20:17:33.145796Z digest=sha256:96afee7a1e3be10d30c633d017faf60afdca2f8b776e72ee90e823f0b296fd05

Observation 215936f2-5b94-4727-bf94-9c0c718462e4 · outbound

This paper cites Self-challenging improves cross-domain generalization.

Bridging Domain Generalization to Multimodal Domain Generalization via Unified Representations Self-challenging improves cross-domain generalization

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:17:34.142163Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-06T20:17:33.265779Z digest=sha256:9fc23719c221ae5c66be126ff8acb2e26e78efbfe969694a47258cf7d440ae1a

Observation 8e06b915-609c-48b9-a755-b4cb14c9aebf · outbound

This paper cites The Kinetics Human Action Video Dataset.

Bridging Domain Generalization to Multimodal Domain Generalization via Unified Representations The Kinetics Human Action Video Dataset

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-06T20:17:33.384873Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:17:33.384873Z digest=sha256:478b65d409bb531ce4e086bc01b59460a0394e0866f055f3e44f0202e5f3ace7

Observation 906275a0-1dc9-43df-b581-9eff0201ce3b · outbound

This paper cites Supervised contrastive learning.

Bridging Domain Generalization to Multimodal Domain Generalization via Unified Representations Supervised contrastive learning

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-06T20:17:33.525005Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:17:33.525005Z digest=sha256:f93a7d8fa3795d5d5e0db98adf1618455b8f198d5a74c3897cd5910216bc4cc8

Observation ece49bbc-b443-4e35-a483-7030e21f7ca1 · outbound

This paper cites Learning to generalize: Meta-learning for do- main generalization.

Bridging Domain Generalization to Multimodal Domain Generalization via Unified Representations Learning to generalize: Meta-learning for do- main generalization

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:17:34.127998Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-06T20:17:33.661806Z digest=sha256:a9ef96b212fea4eafd62a63d662171f3b20150d0b0f43249df91e606268a822d

Observation de586f03-dbb5-4607-b641-327b72510c13 · outbound

This paper cites Domain generalization for med- ical imaging classification with linear-dependency regular- ization.

Bridging Domain Generalization to Multimodal Domain Generalization via Unified Representations Domain generalization for med- ical imaging classification with linear-dependency regular- ization

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:17:34.118896Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-06T20:17:33.664130Z digest=sha256:7969ab5a7547712d27595384d766e6530bf2a25011e882c3e630aa753e554724

Observation 6a796008-bc33-4c1b-bc93-88294838052a · outbound

This paper cites Cross-Modal Discrete Representation Learning.

Bridging Domain Generalization to Multimodal Domain Generalization via Unified Representations Cross-Modal Discrete Representation Learning

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-06T20:17:33.666611Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:17:33.666611Z digest=sha256:7ebe36a430eff01d4299e9df199e2e2c4391ac86b0ae20351ae2b5bb208d21c1

Observation 4c65e704-fbf4-466e-89c6-9418f63b6845 · outbound

This paper cites Feddg: Federated domain generalization on medical image segmentation via episodic learning in continuous fre- quency space.

Bridging Domain Generalization to Multimodal Domain Generalization via Unified Representations Feddg: Federated domain generalization on medical image segmentation via episodic learning in continuous fre- quency space

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:17:34.109865Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-06T20:17:33.669339Z digest=sha256:c95851c2bc03ee8ae3a3149a6fbddf190baefc917c9dc2983652d6287090cd48

Observation 804d2320-b95d-4157-a804-a18d31913ab4 · outbound

This paper cites Unified-io: A unified model for vision, language, and multi-modal tasks.

Bridging Domain Generalization to Multimodal Domain Generalization via Unified Representations Unified-io: A unified model for vision, language, and multi-modal tasks

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:17:34.101332Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-06T20:17:33.671848Z digest=sha256:d2a0456ce2fb0cbfa8ad468da277765f6240680ef2774735c6656f350814fdda

Observation 7b655176-915c-432c-a424-59cafb054e08 · outbound

This paper cites Do- main generalisation via risk distribution matching.

Bridging Domain Generalization to Multimodal Domain Generalization via Unified Representations Do- main generalisation via risk distribution matching

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:17:34.092456Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-06T20:17:33.674222Z digest=sha256:7c7bdeff0ce234cc77a9db21661c6355d240779ccda130af8ae3343cdd04f9f2

Observation ecb813f9-36db-44a4-a33c-4bd6f24d62fa · outbound

This paper cites Unsupervised learning of visual representations by solving jigsaw puzzles.

Bridging Domain Generalization to Multimodal Domain Generalization via Unified Representations Unsupervised learning of visual representations by solving jigsaw puzzles

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-06T20:17:33.677254Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:17:33.677254Z digest=sha256:306b8a90eb61d92a6dd2cfb7415fc902b12217d022bb3798bc83e6a3c6e2228d

Observation afd8d279-07a5-40f5-bf63-bf0ab16a0799 · outbound

This paper cites Causality-inspired single- source domain generalization for medical image segmenta- tion.

Bridging Domain Generalization to Multimodal Domain Generalization via Unified Representations Causality-inspired single- source domain generalization for medical image segmenta- tion

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:17:34.079762Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-06T20:17:33.680080Z digest=sha256:91a7ee532c833b0998cf2d023eee8950c6a8014bb7cae27fff66a087332ec692

Observation f62ac8d6-b12d-4f7b-9f23-028cd51ac077 · outbound

This paper cites Two at once: Enhancing learning and generalization capacities via ibn-net.

Bridging Domain Generalization to Multimodal Domain Generalization via Unified Representations Two at once: Enhancing learning and generalization capacities via ibn-net

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:17:34.071718Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-06T20:17:33.682525Z digest=sha256:97bf01185aab2296199db54e2fa58221676a0c9cce2807c6696e9fd0e20829b1

Observation 0dfb147a-8649-456e-8643-73fe8cb170ff · outbound

This paper cites Audio-visual speech recognition with a hybrid ctc/attention architecture.

Bridging Domain Generalization to Multimodal Domain Generalization via Unified Representations Audio-visual speech recognition with a hybrid ctc/attention architecture

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-06T20:17:33.685007Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:17:33.685007Z digest=sha256:daefd6744d6802b71bec304345c09d378e4e6f65b4d5e1f716e86a211c8e4246

Observation 235b3f85-4e5e-4de5-9357-d3a68211a5dc · outbound

This paper cites Domain generalization through audio- visual relative norm alignment in first person action recog- nition.

Bridging Domain Generalization to Multimodal Domain Generalization via Unified Representations Domain generalization through audio- visual relative norm alignment in first person action recog- nition

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:17:34.059536Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-06T20:17:33.687167Z digest=sha256:13aff0ad620af177822ba67d368467df0cee18955ef6e5ce1237aa79c4ef6057

Observation c8abcd8e-0385-4cf5-b14f-903b0c4cf40a · outbound

This paper cites Learn- ing transferable visual models from natural language super- vision.

Bridging Domain Generalization to Multimodal Domain Generalization via Unified Representations Learn- ing transferable visual models from natural language super- vision

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-06T20:17:33.690248Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:17:33.690248Z digest=sha256:00dd10166985fd682d3c00dbeeea0ab8b21b80c3414ca03577a8ad32cf9da34e

Observation 840e71de-0fed-49d7-8705-117c18b55733 · outbound

This paper cites Domain generalization of 3d semantic segmenta- tion in autonomous driving.

Bridging Domain Generalization to Multimodal Domain Generalization via Unified Representations Domain generalization of 3d semantic segmenta- tion in autonomous driving

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:17:34.046466Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-06T20:17:33.692980Z digest=sha256:4218e7c5c9b06adba4c1989fb0c996e6bfc66696c11fdcd501d8c646a1daa862

Observation ab3faffb-3a66-45b9-9aeb-2def61d19010 · outbound

This paper cites Xkd: Cross-modal knowl- edge distillation with domain alignment for video represen- tation learning.

Bridging Domain Generalization to Multimodal Domain Generalization via Unified Representations Xkd: Cross-modal knowl- edge distillation with domain alignment for video represen- tation learning

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:17:34.037073Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-06T20:17:33.695469Z digest=sha256:2bd4be1909b37f2b5b6d9c57de5c114b4b7273d66269825274ddee0f1a76fdf5

Observation 57f60cfa-bad4-45c9-9323-aeb604e93460 · outbound

This paper cites Domain randomization for transferring deep neural networks from simulation to the real world.

Bridging Domain Generalization to Multimodal Domain Generalization via Unified Representations Domain randomization for transferring deep neural networks from simulation to the real world

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-06T20:17:33.704652Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:17:33.704652Z digest=sha256:3ed81f9e56e77130629fd049bfcc4873578f4313742634652d9f370722209af7

Observation 35372632-7d31-427b-963e-610a229af351 · outbound

This paper cites Deep Domain Confusion: Maximizing for Domain Invariance.

Bridging Domain Generalization to Multimodal Domain Generalization via Unified Representations Deep Domain Confusion: Maximizing for Domain Invariance

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-06T20:17:33.706889Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:17:33.706889Z digest=sha256:abf7207cdf7b52be07069527f43efc13408cd80cd2a76c6dbb8681abbf9dbe5d

Observation 4d4e4a51-aac1-4641-a3e7-efe0001d5e29 · outbound

This paper cites IRBridge: Solving Image Restoration Bridge with Pre-trained Generative Diffusion Models.

Bridging Domain Generalization to Multimodal Domain Generalization via Unified Representations IRBridge: Solving Image Restoration Bridge with Pre-trained Generative Diffusion Models

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-06T20:17:33.709950Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:17:33.709950Z digest=sha256:bd193102fe26fabb7107bdcbfd9c70ddc11ca09c9889675f3b6f8ea0477a0e74

Observation a76050e3-8300-4e75-97d3-7c85a440b111 · outbound

This paper cites Generalizing to unseen domains: A survey on do- main generalization.

Bridging Domain Generalization to Multimodal Domain Generalization via Unified Representations Generalizing to unseen domains: A survey on do- main generalization

Reference 47

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:17:34.024699Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-06T20:17:33.712714Z digest=sha256:22b329494f702ac4b210f90f3115905ba0edb50481edde513b17dc32db069171

Observation 55207443-e347-4e52-9919-a7340aa99d6d · outbound

This paper cites Towards Transformer-Based Aligned Generation with Self-Coherence Guidance.

Bridging Domain Generalization to Multimodal Domain Generalization via Unified Representations Towards Transformer-Based Aligned Generation with Self-Coherence Guidance

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-06T20:17:33.715241Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:17:33.715241Z digest=sha256:3c008832c8d9bd6c8bd71d14aeeda45ee8b998e3b7f06f902c9ad3219d474be4

Observation 5288a065-ccd6-40ea-b0a5-a03739c47ffc · outbound

This paper cites Vlmixer: Unpaired vision-language pre-training via cross-modal cutmix.

Bridging Domain Generalization to Multimodal Domain Generalization via Unified Representations Vlmixer: Unpaired vision-language pre-training via cross-modal cutmix

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-06T20:17:33.717661Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:17:33.717661Z digest=sha256:37a3f384e8a15d0ce9ba2032ab7c39a1042434d6ca7edc1afbd25286279e6a07

Observation 4a04a520-0c9b-4136-b36a-ee9a8c11bacd · outbound

This paper cites Achiev- ing cross modal generalization with multimodal unified rep- resentation.

Bridging Domain Generalization to Multimodal Domain Generalization via Unified Representations Achiev- ing cross modal generalization with multimodal unified rep- resentation

Reference 50

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:17:34.011211Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-06T20:17:33.719984Z digest=sha256:63918148b5fb87b98787c72d7a0aea54c256c87ccff23659c068047d497e3f57

Observation d0b7965f-640f-441b-84b8-2ab63eabcc06 · outbound

This paper cites mixup: Beyond Empirical Risk Minimization.

Bridging Domain Generalization to Multimodal Domain Generalization via Unified Representations mixup: Beyond Empirical Risk Minimization

Reference 51

Resolution
unresolved
no resolver link, observed 2026-08-06T20:17:33.723554Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:17:33.723554Z digest=sha256:fb0d58ebfb356a532e5078e35150f7956665f89d13c414cde8cc04e86cf384a3

Observation c78852ac-f983-4f70-9ce7-22d74d7406c8 · outbound

This paper cites Towards effective multi-modal interchanges in zero-resource sounding object localization.

Bridging Domain Generalization to Multimodal Domain Generalization via Unified Representations Towards effective multi-modal interchanges in zero-resource sounding object localization

Reference 52

Resolution
unresolved
no resolver link, observed 2026-08-06T20:17:33.726778Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:17:33.726778Z digest=sha256:a57a8b5e6912dfcd9df39054299791eb74dc97aab64ce1d81ea8e8ee350bc74e

Observation 274f3c99-efa3-4402-8f0c-0942b630b3b1 · outbound

This paper cites Deep domain-adversarial image generation for do- main generalisation.

Bridging Domain Generalization to Multimodal Domain Generalization via Unified Representations Deep domain-adversarial image generation for do- main generalisation

Reference 53

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:17:33.998204Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-06T20:17:33.729663Z digest=sha256:5d8af7e50bc36ed163e50379b6b576e55bbdd0481ef4052dea8b8c66168a8286

Observation ce04ae21-10af-465f-a1ce-205d586acd9b · outbound

This paper cites The feature dimensions for video, audio, and optical flow are 2304, 512, and 2048, respec- tively.

Bridging Domain Generalization to Multimodal Domain Generalization via Unified Representations The feature dimensions for video, audio, and optical flow are 2304, 512, and 2048, respec- tively

Reference 54

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:17:33.987896Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-06T20:17:33.732280Z digest=sha256:201abcc31729ef698a1278c007b2902a6945a27500fc412670a5d3c5e042a4f1

Observation 198db991-3e51-4860-aba8-d3d4dc5bb88c · outbound

This paper cites In contrast, our proposed approach sub- stantially improves their performance in the MMDG set- ting.

Bridging Domain Generalization to Multimodal Domain Generalization via Unified Representations In contrast, our proposed approach sub- stantially improves their performance in the MMDG set- ting

Reference 55

Resolution
verified exact
raw_fallback, observed 2026-08-06T20:17:33.826037Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-06T20:17:33.735441Z digest=sha256:8f3dde2d986464e0fa6e318fded5c8c27b4a4e105f25065e4456ac3a713e0b65

Observation 39431b76-be70-446e-9b0d-6c8f699dff9e · outbound

This paper cites Notably, our method exhibits minimal fluctuations across all parame- ter settings, indicating a lower sensitivity to hyperparameter selection.

Bridging Domain Generalization to Multimodal Domain Generalization via Unified Representations Notably, our method exhibits minimal fluctuations across all parame- ter settings, indicating a lower sensitivity to hyperparameter selection

Reference 56

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:17:33.977922Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-06T20:17:33.738182Z digest=sha256:8949cc8a1a094f6befd42617ef1698ca48ef38139cd7e0ceb5c1274afea77493

Observation f164fdb5-87d3-4fab-b0a4-8b278bc93993 · outbound

This paper cites We do not ab- late Lcls since it is essential for classification.

Bridging Domain Generalization to Multimodal Domain Generalization via Unified Representations We do not ab- late Lcls since it is essential for classification

Reference 57

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:17:33.968234Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-06T20:17:33.740779Z digest=sha256:8879da4887d99fbf1fd35b793ba17aa158d89fee95a7bbe74565e1d943dbb75a

Observation 19bb939f-7ff0-444f-99ea-caac1b72b22b · outbound

This paper cites It can be observed that the gen- eral and specific information of each modality are well- separated and consistently aligned across domains.

Bridging Domain Generalization to Multimodal Domain Generalization via Unified Representations It can be observed that the gen- eral and specific information of each modality are well- separated and consistently aligned across domains

Reference 58

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:17:33.959004Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-06T20:17:33.743609Z digest=sha256:2881035c0da685167317ed6cd85d6ba3ec284050b61f61304cfe6bcfe14cdea6

Pith citing papers

Observation bdc993b6-96dd-4fdc-b61a-fd7b27d75e34 · inbound

Open-set Cross Modal Generalization via Multimodal Unified Representation cites this paper.

Open-set Cross Modal Generalization via Multimodal Unified Representation Bridging Domain Generalization to Multimodal Domain Generalization via Unified Representations

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-06T15:48:28.447490Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:48:28.447490Z digest=sha256:5be50623e314b734ebd684261a6f2b071f522823a9e6919ad7731d80bb5d001d

Observation 211054c5-9651-4fd4-aa8f-cb0ac8b06d73 · inbound

TAP: Parameter-efficient Task-Aware Prompting for Adverse Weather Removal cites this paper.

TAP: Parameter-efficient Task-Aware Prompting for Adverse Weather Removal Bridging Domain Generalization to Multimodal Domain Generalization via Unified Representations

Reference 20

Resolution
verified exact
local_arxiv, observed 2026-08-05T21:53:34.006173Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-05T21:53:28.082648Z digest=sha256:a0ca4594feb0e4c0c3c913abf8011b220aea5a081308ad66d532d081da998835