Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-15T21:01:38.878311Z
Paper Citation Record · LEDGER
As of 16 August 2026, this Paper Citation Record lists 100 of 111 outbound references and 0 inbound Pith citation observations for arXiv:2505.11216.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-15T21:01:38.878311Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-16T06:30:59.297886+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links
A source-named dated measurement, never combined with another source.
Source: cited_works
100 of 111 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 66cfa46f-a0bd-4a78-a9c2-e6ee31dbede3 · outbound
GeoMM: On Geodesic Perspective for Multi-modal Learning Geometry of oblique projections
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 8659d41e-7dc6-4ffd-93b0-1cb8e5a0b684 · outbound
GeoMM: On Geodesic Perspective for Multi-modal Learning Vqa: Visual question an- swering
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 55c1edf9-c5ee-4ba4-9e2f-9b139486cc82 · outbound
GeoMM: On Geodesic Perspective for Multi-modal Learning Geodesic matting: A framework for fast interactive image and video seg- mentation and matting
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b7dec225-807d-4d7d-82d5-77a26f2fff1c · outbound
GeoMM: On Geodesic Perspective for Multi-modal Learning Vlmo: Uni- fied vision-language pre-training with mixture-of- modality-experts
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 21f0c57b-7496-4784-8aa4-7068f74eefbe · outbound
GeoMM: On Geodesic Perspective for Multi-modal Learning Grit-vlp: Grouped mini-batch sam- pling for efficient vision and language pre-training
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7a3ce4a2-f859-4cbe-af60-3c16ac592db2 · outbound
GeoMM: On Geodesic Perspective for Multi-modal Learning Mafa: Managing false negatives for vision-language pre-training
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a48256e5-9822-428d-9acf-0f65b344a5e5 · outbound
GeoMM: On Geodesic Perspective for Multi-modal Learning End-to-end object detection with trans- formers
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation db3622f6-0593-4b8f-87af-fff6fcf727bb · outbound
GeoMM: On Geodesic Perspective for Multi-modal Learning Un- supervised learning of visual features by contrasting cluster assignments
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fb5ad8ad-31d2-4fa2-a7ab-c9bb930b2de1 · outbound
GeoMM: On Geodesic Perspective for Multi-modal Learning Emerging properties in self-supervised vi- sion transformers
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 99d13011-6d78-4073-b0c4-48eab8b4d2d9 · outbound
GeoMM: On Geodesic Perspective for Multi-modal Learning Conceptual 12m: Pushing web-scale image-text pre-training to recognize long-tail visual concepts
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 274db597-f706-466c-94af-bff1aab1ab82 · outbound
GeoMM: On Geodesic Perspective for Multi-modal Learning STAIR: Learning Sparse Text and Image Representation in Grounded Tokens
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 65cd4779-0669-4769-8fb5-cf7d80842a14 · outbound
GeoMM: On Geodesic Perspective for Multi-modal Learning Vlp: A survey on vision-language pre-training
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2fe0d320-69d6-4761-914f-af8965a929be · outbound
GeoMM: On Geodesic Perspective for Multi-modal Learning A simple framework for con- trastive learning of visual representations
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 004bccc6-a759-477a-b537-6ce78ea73233 · outbound
GeoMM: On Geodesic Perspective for Multi-modal Learning Exploring simple siamese representation learning
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4196cc9e-1fce-4152-ae6c-f046f859e2b4 · outbound
GeoMM: On Geodesic Perspective for Multi-modal Learning Improved Baselines with Momentum Contrastive Learning
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8adae875-cc7c-4b7c-959d-a92eeedff320 · outbound
GeoMM: On Geodesic Perspective for Multi-modal Learning X-volution: On the unification of convolution and self-attention
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation ad3267df-c538-4ee7-b467-0337bcb55ec6 · outbound
GeoMM: On Geodesic Perspective for Multi-modal Learning Uniter: Universal image-text represen- tation learning
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b5b4d7b4-0948-4db1-bff5-bb6998010a34 · outbound
GeoMM: On Geodesic Perspective for Multi-modal Learning Unsupervised Opinion Summarization Using Approximate Geodesics
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 54f20193-659f-4878-8e05-b9d15bf37185 · outbound
GeoMM: On Geodesic Perspective for Multi-modal Learning Geodesics in heat: A new approach to computing distance based on heat flow
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3478c4dc-1800-4b8a-8fa1-150a8fac611b · outbound
GeoMM: On Geodesic Perspective for Multi-modal Learning Imagenet: A large-scale hierar- chical image database
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b6a0e3ef-a335-403a-afe0-dea09844afcb · outbound
GeoMM: On Geodesic Perspective for Multi-modal Learning BERT: Pre-training of Deep Bidirectional Transformers for Language Understanding
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 563e7ed6-9e23-49c7-8249-556a7a4e2d39 · outbound
GeoMM: On Geodesic Perspective for Multi-modal Learning Similarity reasoning and filtration for image-text matching
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8ede5920-4d59-4b00-996e-06fa677e62b4 · outbound
GeoMM: On Geodesic Perspective for Multi-modal Learning A note on two problems in con- nexion with graphs
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2bf51a31-6b6b-42d0-a3bf-253e0f9ca610 · outbound
GeoMM: On Geodesic Perspective for Multi-modal Learning An Image is Worth 16x16 Words: Transformers for Image Recognition at Scale
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 31a91c94-7b3c-49b9-b09a-29f0d3911a19 · outbound
GeoMM: On Geodesic Perspective for Multi-modal Learning Algorithm 97: shortest path
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3ef1901e-d388-47be-9797-5288ce17d716 · outbound
GeoMM: On Geodesic Perspective for Multi-modal Learning Large-scale adversar- ial training for vision-and-language representation learning
Reference 26
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 94e9c6b0-76e3-4cfc-825b-3ce545cc4019 · outbound
GeoMM: On Geodesic Perspective for Multi-modal Learning Imagebind: One embedding space to bind them all
Reference 27
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 53a7fab5-efb6-4ff8-bbec-c2242f7cbfc4 · outbound
GeoMM: On Geodesic Perspective for Multi-modal Learning Momentum contrast for unsupervised visual representation learning
Reference 28
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a3f9074b-fbab-445c-a06a-d73412d91706 · outbound
GeoMM: On Geodesic Perspective for Multi-modal Learning Masked autoen- coders are scalable vision learners
Reference 29
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fc0aedf2-1406-4814-93fc-e27dac959f1f · outbound
GeoMM: On Geodesic Perspective for Multi-modal Learning Geonet: Deep geodesic networks for point cloud analysis
Reference 30
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 276fe121-e048-4455-a1f0-4be44f6352e2 · outbound
GeoMM: On Geodesic Perspective for Multi-modal Learning Pixel-BERT: Aligning Image Pixels with Text by Deep Multi-Modal Transformers
Reference 32
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8ab911b9-f86b-4331-9caf-a74d29e77a59 · outbound
GeoMM: On Geodesic Perspective for Multi-modal Learning Scaling up visual and vision-language representation learning with noisy text supervision
Reference 33
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4da22129-6ec5-4438-9d0f-38cac98f3077 · outbound
GeoMM: On Geodesic Perspective for Multi-modal Learning Vilt: Vision-and-language transformer without convolu- tion or region supervision
Reference 34
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 43dd0a4b-2483-4364-90ad-ed7f8b51290b · outbound
GeoMM: On Geodesic Perspective for Multi-modal Learning Computing geodesic paths on manifolds
Reference 35
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 27c0734b-23d0-43c1-9612-5e6207f46290 · outbound
GeoMM: On Geodesic Perspective for Multi-modal Learning Visual genome: Connecting language and vision using crowdsourced dense image annota- tions
Reference 36
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7e43bedb-4f44-468c-bcf5-b3c469f4d285 · outbound
GeoMM: On Geodesic Perspective for Multi-modal Learning Numba: a llvm-based python JIT compiler
Reference 37
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation aa33debf-79a4-4188-b5a2-4d393a1c1ee8 · outbound
GeoMM: On Geodesic Perspective for Multi-modal Learning Le, Vu Nguyen, Chen-Ping Yu, and Dimitris Samaras
Reference 38
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 46ea7ca2-464f-42cf-bda9-8f1d75c8caee · outbound
GeoMM: On Geodesic Perspective for Multi-modal Learning Stacked cross attention for image- text matching
Reference 39
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 23f519ed-53c6-45ae-a6e6-f7b0681e5b9d · outbound
GeoMM: On Geodesic Perspective for Multi-modal Learning Multimodal Foundation Models: From Specialists to General-Purpose Assistants
Reference 40
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 71eddf66-8905-4e23-8355-6cd0d419bc8e · outbound
GeoMM: On Geodesic Perspective for Multi-modal Learning Align before fuse: Vision and lan- guage representation learning with momentum dis- tillation
Reference 41
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f3f7b385-b0a0-4dbb-88c2-1a9fd92507e0 · outbound
GeoMM: On Geodesic Perspective for Multi-modal Learning Blip: Bootstrapping language-image pre- training for unified vision-language understanding and generation
Reference 42
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 365af228-59ce-46e0-ae1c-7e91497c43c6 · outbound
GeoMM: On Geodesic Perspective for Multi-modal Learning HERO: Hierarchical Encoder for Video+Language Omni-representation Pre-training
Reference 43
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4cb98fbf-d047-4e6d-bfcd-bb8854344691 · outbound
GeoMM: On Geodesic Perspective for Multi-modal Learning VisualBERT: A Simple and Performant Baseline for Vision and Language
Reference 44
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4d8b67ea-7b08-4c7a-89fb-c7d019169cfe · outbound
GeoMM: On Geodesic Perspective for Multi-modal Learning Oscar: Object-semantics aligned pre-training for vision-language tasks
Reference 45
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation b18e0303-3dcd-4727-b2b8-c843df5c473e · outbound
GeoMM: On Geodesic Perspective for Multi-modal Learning Supervision Exists Everywhere: A Data Efficient Contrastive Language-Image Pre-training Paradigm
Reference 46
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c4cc866d-c672-4f5d-9e80-c0fa2e1bc243 · outbound
GeoMM: On Geodesic Perspective for Multi-modal Learning Scaling language- image pre-training via masking
Reference 47
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 9254e8d4-3858-4246-a97a-f3fa3252d5e7 · outbound
GeoMM: On Geodesic Perspective for Multi-modal Learning Geodesic self- attention for 3d point clouds
Reference 48
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 06178822-4c25-4d71-a750-3cdcaa1e655d · outbound
GeoMM: On Geodesic Perspective for Multi-modal Learning Microsoft coco: Common objects in context
Reference 49
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 60a9f332-f09a-4b03-abb3-693a13838ba4 · outbound
GeoMM: On Geodesic Perspective for Multi-modal Learning Unresolved cited work
Reference 50
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 39a9358a-fce3-4c28-a839-3c9deeedb799 · outbound
GeoMM: On Geodesic Perspective for Multi-modal Learning Adap- tive reconstruction network for weakly supervised re- ferring expression grounding
Reference 51
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 2351d9c4-6ef2-45fb-8709-8d500a8fbb70 · outbound
GeoMM: On Geodesic Perspective for Multi-modal Learning Algorithm as 136: A k-means clustering algorithm
Reference 52
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 4d014c8a-5557-4533-8c80-b64489b474eb · outbound
GeoMM: On Geodesic Perspective for Multi-modal Learning Decoupled weight decay regularization
Reference 53
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation fe50d3e3-97e6-4ba1-989f-c0f4b8c4e66e · outbound
GeoMM: On Geodesic Perspective for Multi-modal Learning Vilbert: Pretraining task-agnostic visiolinguistic rep- resentations for vision-and-language tasks.Advances in Neural Information Processing Systems, 32, 2019
Reference 54
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 5246da00-04c3-45f3-8e9a-2f135549850b · outbound
GeoMM: On Geodesic Perspective for Multi-modal Learning Computing geodesics on triangular meshes.Comput- ers & Graphics, 29(5):667–675, 2005
Reference 55
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation b5cd8895-692e-4312-9a36-eb02c979757c · outbound
GeoMM: On Geodesic Perspective for Multi-modal Learning Bron- stein, and Pierre Vandergheynst
Reference 56
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation db0dae6f-1b92-491a-9eec-622b4be87742 · outbound
GeoMM: On Geodesic Perspective for Multi-modal Learning Jensen’s inequality
Reference 57
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 53b6afc6-3b03-4001-99e5-aa01f5ddb3f2 · outbound
GeoMM: On Geodesic Perspective for Multi-modal Learning Towards bridging sample complexity and model capacity
Reference 58
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 037bcfe0-7300-451f-8c0e-90acd50e9d26 · outbound
GeoMM: On Geodesic Perspective for Multi-modal Learning Towards interpreting and utiliz- ing symmetry property in adversarial examples
Reference 59
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation eb72ad7e-e4f0-4cb7-ac55-d7d92f501aee · outbound
GeoMM: On Geodesic Perspective for Multi-modal Learning Exploring and utilizing pattern imbal- ance
Reference 60
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation fc249f86-bb9e-4169-b426-df1f88ff3b4b · outbound
GeoMM: On Geodesic Perspective for Multi-modal Learning MSSIDD: A Benchmark for Multi-Sensor Denoising
Reference 61
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8667fbc0-cf70-4172-a709-7e11fe756667 · outbound
GeoMM: On Geodesic Perspective for Multi-modal Learning Object- oriented anchoring and modal alignment in multi- modal learning
Reference 62
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation ab70750a-7f52-41d8-b35d-9c13d9cf1988 · outbound
GeoMM: On Geodesic Perspective for Multi-modal Learning Unresolved cited work
Reference 63
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation dfa2bc74-bcff-4279-9090-48d22c9a014a · outbound
GeoMM: On Geodesic Perspective for Multi-modal Learning Analytic inequalities
Reference 64
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 779b1cb0-4418-4c9a-b5ee-c5bbc4d44e24 · outbound
GeoMM: On Geodesic Perspective for Multi-modal Learning Slip: Self-supervision meets language- image pre-training
Reference 65
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation fa827551-edcc-4c9e-9912-3714a5411439 · outbound
GeoMM: On Geodesic Perspective for Multi-modal Learning Geodesic-former: A geodesic-guided few-shot 3d point cloud instance segmenter
Reference 66
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 365b3a40-91df-420d-8e36-a9d33cf78743 · outbound
GeoMM: On Geodesic Perspective for Multi-modal Learning Representation Learning with Contrastive Predictive Coding
Reference 67
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 02cfc387-40af-4776-aaca-1ce14fa83ce9 · outbound
GeoMM: On Geodesic Perspective for Multi-modal Learning Im2text: Describing images using 1 million cap- tioned photographs
Reference 68
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation c050648a-b3e4-47df-b98c-50d09c10e2f9 · outbound
GeoMM: On Geodesic Perspective for Multi-modal Learning Automatic differentiation in pytorch
Reference 69
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation cdca1922-b4c5-44fe-b249-9b550c72b70f · outbound
GeoMM: On Geodesic Perspective for Multi-modal Learning BEiT v2: Masked Image Modeling with Vector-Quantized Visual Tokenizers
Reference 70
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 963c7e38-774e-441b-a3ab-26fa8a413f82 · outbound
GeoMM: On Geodesic Perspective for Multi-modal Learning Computational optimal transport: With applications to data science
Reference 71
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation b1c8fd83-b920-4b38-bd53-203dd2f7e424 · outbound
GeoMM: On Geodesic Perspective for Multi-modal Learning Combined scaling for zero-shot transfer learn- ing
Reference 72
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation a45f9e51-ef3b-4c23-9477-b6dc223754f4 · outbound
GeoMM: On Geodesic Perspective for Multi-modal Learning Flickr30k entities: Collecting region-to- phrase correspondences for richer image-to-sentence models
Reference 73
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 3aa8b2b5-f6c8-432a-bfad-9389940629e9 · outbound
GeoMM: On Geodesic Perspective for Multi-modal Learning Straightest geodesics on polyhedral surfaces
Reference 74
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 69d4cb8f-1a1c-4d7f-beaf-ea39ec9da25a · outbound
GeoMM: On Geodesic Perspective for Multi-modal Learning Graphwalks: Efficient shape agnostic geodesic shortest path estimation
Reference 75
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 83ece59c-1713-4f08-9db8-d05488a96e43 · outbound
GeoMM: On Geodesic Perspective for Multi-modal Learning ImageBERT: Cross-modal Pre-training with Large-scale Weak-supervised Image-Text Data
Reference 76
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b91a6b2b-8976-4bc4-9df4-028ed954f459 · outbound
GeoMM: On Geodesic Perspective for Multi-modal Learning Improving language under- standing by generative pre-training
Reference 77
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation cf238571-3c66-4790-a28d-b0db7bb9eb3c · outbound
GeoMM: On Geodesic Perspective for Multi-modal Learning Learning transferable visual models from natural language supervision
Reference 78
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 34b842d2-b60f-445a-a601-22e99fb888e9 · outbound
GeoMM: On Geodesic Perspective for Multi-modal Learning Faster r-cnn: Towards real-time object detection with region proposal networks
Reference 79
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 8da1e150-ae73-490d-a9c7-3157e38dbdf1 · outbound
GeoMM: On Geodesic Perspective for Multi-modal Learning Grad-cam: Visual explanations from deep networks via gradient-based localization
Reference 80
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 71a53bd5-3456-4d4e-aa00-1ceb46daee9c · outbound
GeoMM: On Geodesic Perspective for Multi-modal Learning Conceptual captions: A cleaned, hy- pernymed, image alt-text dataset for automatic image captioning
Reference 81
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation ef6e53a7-c2d1-4362-8650-182f8a060ede · outbound
GeoMM: On Geodesic Perspective for Multi-modal Learning VL-BERT: Pre-training of Generic Visual-Linguistic Representations
Reference 82
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 37bb562b-fa2e-443f-b0de-ce00af6ef2dd · outbound
GeoMM: On Geodesic Perspective for Multi-modal Learning PandaGPT: One Model To Instruction-Follow Them All
Reference 83
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2d304298-9594-4627-bfbf-0b7daf9418fe · outbound
GeoMM: On Geodesic Perspective for Multi-modal Learning A Corpus for Reasoning About Natural Language Grounded in Photographs
Reference 84
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 140a7beb-1be5-4a53-ad80-278df7e9ca50 · outbound
GeoMM: On Geodesic Perspective for Multi-modal Learning Revisiting unreasonable effective- ness of data in deep learning era
Reference 85
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 80e6a150-979d-4e34-96b5-8e289fb4060b · outbound
GeoMM: On Geodesic Perspective for Multi-modal Learning Gortler, and Hugues Hoppe
Reference 86
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 71871314-e8da-4eb1-821a-1a4f8785f83e · outbound
GeoMM: On Geodesic Perspective for Multi-modal Learning LXMERT: learning cross-modality encoder representations from trans- formers
Reference 87
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 06bf8c74-6ea4-4cce-afc8-ea3c2782f5ee · outbound
GeoMM: On Geodesic Perspective for Multi-modal Learning Tenenbaum, Vin de Silva, and John C
Reference 88
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 65913de9-b905-4a15-910a-51ccf44688da · outbound
GeoMM: On Geodesic Perspective for Multi-modal Learning Pigeon hole principle
Reference 89
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 403d64ff-a537-4552-8f49-4041e217dbac · outbound
GeoMM: On Geodesic Perspective for Multi-modal Learning Attention is all you need
Reference 90
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation db607275-2142-4af8-8dd8-14cb766e6385 · outbound
GeoMM: On Geodesic Perspective for Multi-modal Learning Optimal transport: old and new
Reference 91
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 2dcf3c28-e99f-4ec3-ae7c-e8c33c20b304 · outbound
GeoMM: On Geodesic Perspective for Multi-modal Learning Learning to combine: Knowledge aggrega- tion for multi-source domain adaptation
Reference 92
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 7c0b0f61-e2b7-46d9-bb19-459480b1a9e1 · outbound
GeoMM: On Geodesic Perspective for Multi-modal Learning Image as a Foreign Language: BEiT Pretraining for All Vision and Vision-Language Tasks
Reference 93
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 55b012a0-b38b-49f9-b321-51c4f2c794b8 · outbound
GeoMM: On Geodesic Perspective for Multi-modal Learning Mvp: Multimodality-guided visual pre-training
Reference 94
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation deaa2c2d-f7ac-4aec-a500-d1e3e843a593 · outbound
GeoMM: On Geodesic Perspective for Multi-modal Learning Visual Entailment: A Novel Task for Fine-Grained Image Understanding
Reference 95
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 35dd7e54-89fe-4315-8fb4-62c1fc13119e · outbound
GeoMM: On Geodesic Perspective for Multi-modal Learning A fast proximal point method for computing exact wasserstein distance
Reference 96
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 9f1ffb73-0c4f-45eb-8907-460c0cd33c99 · outbound
GeoMM: On Geodesic Perspective for Multi-modal Learning Vision-language pre- training with triple contrastive learning
Reference 97
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 89478e07-0d26-4164-881d-72924a471f13 · outbound
GeoMM: On Geodesic Perspective for Multi-modal Learning Unified contrastive learning in image-text-label space
Reference 98
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 957b594d-6f13-47fe-96d9-93b44a7a7eb9 · outbound
GeoMM: On Geodesic Perspective for Multi-modal Learning FILIP: Fine-grained Interactive Language-Image Pre-Training
Reference 99
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5c558f01-be97-4bf8-aa03-68cd271fc20d · outbound
GeoMM: On Geodesic Perspective for Multi-modal Learning FILIP: fine-grained interactive language-image pre-training
Reference 100
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 132e1aec-03d9-42ba-a78b-d52284e19a45 · outbound
GeoMM: On Geodesic Perspective for Multi-modal Learning CoCa: Contrastive Captioners are Image-Text Foundation Models
Reference 101
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
No inbound Pith citation observations are available.