Pith. sign in

Paper Citation Record · LEDGER

MultiMAE Meets Earth Observation: Pre-training Multi-modal Multi-task Masked Autoencoders for Earth Observation Tasks

As of 10 August 2026, this Paper Citation Record lists 34 of 34 outbound references and 1 inbound Pith citation observation for arXiv:2505.14951.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2505.14951 v1

Coverage vector

measured 34 of 34 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T15:29:54.475780Z

measured 35 of 35 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00

measured 1 of 1 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-06T17:08:13.297661Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-08-06T17:08:16.157991Z

Reference resolution

34 of 34 outbound references displayed

  • verified exact0
  • verified fuzzy22
  • unresolved12
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 8abeaa84-cf6c-43ad-a98e-ef34afc4c09e · outbound

This paper cites Multimae: Multi-modal multi-task masked autoen- coders.

MultiMAE Meets Earth Observation: Pre-training Multi-modal Multi-task Masked Autoencoders for Earth Observation Tasks Multimae: Multi-modal multi-task masked autoen- coders

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:29:58.407649Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T15:29:51.356112Z digest=sha256:20a4ade197a3b43d20adc6389364d98705da95b61cf7800d751688099fa7934d

Observation 103c181d-ca92-4271-8cde-eb9fe46fa57a · outbound

This paper cites Satlaspretrain: A large- scale dataset for remote sensing image understanding.

MultiMAE Meets Earth Observation: Pre-training Multi-modal Multi-task Masked Autoencoders for Earth Observation Tasks Satlaspretrain: A large- scale dataset for remote sensing image understanding

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:29:58.180477Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T15:29:51.441077Z digest=sha256:d08518a26a3d28c95d644ea461706e6c0971b383fc96ab2f5b625e1d87e18d77

Observation c043bd61-342e-4b5a-ba6b-0810ce526277 · outbound

This paper cites HLS Multi Temporal Crop Classification, 2023.

MultiMAE Meets Earth Observation: Pre-training Multi-modal Multi-task Masked Autoencoders for Earth Observation Tasks HLS Multi Temporal Crop Classification, 2023

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:29:57.934653Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T15:29:51.579948Z digest=sha256:2b8a7cc367742aa367a6edf32094fd5ed0f10f4842b8a51a5c60a024b3542935

Observation bd5c0250-b76b-41cf-a2a9-afe910db9356 · outbound

This paper cites Satmae: Pre-training transformers for tem- poral and multi-spectral satellite imagery.NeurIPS, 35:197– 211, 2022.

MultiMAE Meets Earth Observation: Pre-training Multi-modal Multi-task Masked Autoencoders for Earth Observation Tasks Satmae: Pre-training transformers for tem- poral and multi-spectral satellite imagery.NeurIPS, 35:197– 211, 2022

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:29:57.830678Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T15:29:51.695559Z digest=sha256:b1bb00205e4a74e2377a12d92295e02daa984aee69fdd79371ea72fd9cd02c8f

Observation 9b296322-03bd-43ea-9227-6a46225e8db9 · outbound

This paper cites Imagenet: A large-scale hierarchical image database.

MultiMAE Meets Earth Observation: Pre-training Multi-modal Multi-task Masked Autoencoders for Earth Observation Tasks Imagenet: A large-scale hierarchical image database

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-07T15:29:51.806284Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:29:51.806284Z digest=sha256:8e8f6d25df5edf0cc8b704e05ec2cef0d9b1e480b98a1092ba0ad5918bd59efc

Observation 25c6e75f-a217-47e4-9a8c-886635a3df7e · outbound

This paper cites An Image is Worth 16x16 Words: Transformers for Image Recognition at Scale.

MultiMAE Meets Earth Observation: Pre-training Multi-modal Multi-task Masked Autoencoders for Earth Observation Tasks An Image is Worth 16x16 Words: Transformers for Image Recognition at Scale

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-07T15:29:51.936537Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:29:51.936537Z digest=sha256:57d32cda0cbff4a40203e1b37659adf1ee3e2109ec98023f119ef2aefe2609db

Observation c654ea9e-0232-4f7c-9d08-ac6cba409b5a · outbound

This paper cites Croma: Remote sensing representations with contrastive radar- optical masked autoencoders.NeurIPS, 36, 2024.

MultiMAE Meets Earth Observation: Pre-training Multi-modal Multi-task Masked Autoencoders for Earth Observation Tasks Croma: Remote sensing representations with contrastive radar- optical masked autoencoders.NeurIPS, 36, 2024

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:29:57.696017Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T15:29:52.090685Z digest=sha256:6f9ca2814c16d30d0fd1d8cf299ac75a861c817ab909a5ff68db48a2d40fcb88

Observation 30ef9799-0805-4057-a7a9-a770b8ee3c4c · outbound

This paper cites Skysense: A multi-modal remote sens- ing foundation model towards universal interpretation for earth observation imagery.

MultiMAE Meets Earth Observation: Pre-training Multi-modal Multi-task Masked Autoencoders for Earth Observation Tasks Skysense: A multi-modal remote sens- ing foundation model towards universal interpretation for earth observation imagery

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:29:57.586494Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T15:29:52.205724Z digest=sha256:c274247cf2dd5147712949fc1367fffd7d551da4970f6644a5e6584c2e5de603

Observation b3ff44e5-cb70-4706-9018-9270e62cfc1c · outbound

This paper cites Masked autoencoders are scalable vision learners.

MultiMAE Meets Earth Observation: Pre-training Multi-modal Multi-task Masked Autoencoders for Earth Observation Tasks Masked autoencoders are scalable vision learners

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:29:57.495808Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T15:29:52.368114Z digest=sha256:90ac020d6070710760cde634f30ee511dc164f59de5ede58d088d9e93980c273

Observation ece778ca-2009-45ec-b9bd-77e3778c8ff3 · outbound

This paper cites an unresolved cited work.

MultiMAE Meets Earth Observation: Pre-training Multi-modal Multi-task Masked Autoencoders for Earth Observation Tasks Unresolved cited work

Reference 10

Resolution
unresolved
raw_fallback, observed 2026-08-07T15:29:57.344291Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T15:29:52.502692Z digest=sha256:eac022d457540bc10c3790b8c2679fefd909a1a64f622ba9c2bfea42aaa70b61

Observation 304e5bbc-e06a-4e57-a0d0-9a73b31ece0c · outbound

This paper cites Masked Image Modeling: A Survey.

MultiMAE Meets Earth Observation: Pre-training Multi-modal Multi-task Masked Autoencoders for Earth Observation Tasks Masked Image Modeling: A Survey

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-07T15:29:52.563283Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:29:52.563283Z digest=sha256:e1e241297781c4f736ab478bd615ef57de986e80a9508ef6d6d504c13e9913b7

Observation 80fee846-169f-4578-8233-add4a506f278 · outbound

This paper cites Masked autoencoders that listen.

MultiMAE Meets Earth Observation: Pre-training Multi-modal Multi-task Masked Autoencoders for Earth Observation Tasks Masked autoencoders that listen

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:29:57.218245Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T15:29:52.631096Z digest=sha256:b0431b431eb86a43cb09bd7dbf561de0d66d1d06332098482d618626984b6ea5

Observation fc78dbde-2c4c-4696-a112-9268c208b36a · outbound

This paper cites Foundation Models for Generalist Geospatial Artificial Intelligence.

MultiMAE Meets Earth Observation: Pre-training Multi-modal Multi-task Masked Autoencoders for Earth Observation Tasks Foundation Models for Generalist Geospatial Artificial Intelligence

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-07T15:29:52.735664Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:29:52.735664Z digest=sha256:b4115508086eed35d24c25df2618519715c65ca00929b13e8e9d410c781e36af

Observation f32c05a6-7d51-483e-9d03-e173e48032d4 · outbound

This paper cites Geo- bench: Toward foundation models for earth monitoring.

MultiMAE Meets Earth Observation: Pre-training Multi-modal Multi-task Masked Autoencoders for Earth Observation Tasks Geo- bench: Toward foundation models for earth monitoring

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:29:57.077347Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T15:29:52.814349Z digest=sha256:a83d50f8551d4e1ee4651b9ae11e3f6ba8603efebc5a51d4708edfc9eaafee0f

Observation 8630e3c1-09c5-41f1-ba7a-9154ef8524be · outbound

This paper cites Multimodality helps unimodality: Cross- modal few-shot learning with multimodal models.

MultiMAE Meets Earth Observation: Pre-training Multi-modal Multi-task Masked Autoencoders for Earth Observation Tasks Multimodality helps unimodality: Cross- modal few-shot learning with multimodal models

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:29:56.960366Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T15:29:52.905682Z digest=sha256:799333fd79460292b029ac0815ebb9ccaca01b5fb8bfaf77f978dc0fa5d0f0c3

Observation b2ea4498-0e3e-4906-af59-5968c43aef6a · outbound

This paper cites Re- moteclip: A vision language foundation model for remote sensing.IEEE Transactions on Geoscience and Remote Sensing, 2024.

MultiMAE Meets Earth Observation: Pre-training Multi-modal Multi-task Masked Autoencoders for Earth Observation Tasks Re- moteclip: A vision language foundation model for remote sensing.IEEE Transactions on Geoscience and Remote Sensing, 2024

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:29:56.854441Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T15:29:52.967460Z digest=sha256:28448d57f26bc7ddfb793e127f1111774a2fca57c801f7880ab1538d6524ca6c

Observation bb8b4295-8a4b-41c9-aa36-d00daf446721 · outbound

This paper cites A convnet for the 2020s.

MultiMAE Meets Earth Observation: Pre-training Multi-modal Multi-task Masked Autoencoders for Earth Observation Tasks A convnet for the 2020s

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:29:56.716250Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T15:29:53.017671Z digest=sha256:1e25004e56e664a43af52a1031edc168cdaad5b64cefd31e08364379aa4b47d8

Observation c22a0db8-849d-439d-aecf-b0511898dd12 · outbound

This paper cites MMEarth: Exploring Multi-Modal Pretext Tasks For Geospatial Representation Learning.

MultiMAE Meets Earth Observation: Pre-training Multi-modal Multi-task Masked Autoencoders for Earth Observation Tasks MMEarth: Exploring Multi-Modal Pretext Tasks For Geospatial Representation Learning

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-07T15:29:53.100039Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:29:53.100039Z digest=sha256:f8c95bbbbca3ca8b3046f3c724e1b5648117d32409e95094a7274211944990a0

Observation 0b75565f-c42a-4228-a242-2406af6445e1 · outbound

This paper cites Rethinking transformers pre-training for multi- spectral satellite imagery.

MultiMAE Meets Earth Observation: Pre-training Multi-modal Multi-task Masked Autoencoders for Earth Observation Tasks Rethinking transformers pre-training for multi- spectral satellite imagery

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:29:56.488338Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T15:29:53.188818Z digest=sha256:5c789f7962cac55df41e51013b9d347cb03363ab741d1c11d6fcfd97fef8d8ca

Observation e462cdb5-2da5-49d6-8a8c-83fdd31691ce · outbound

This paper cites How Effective is Pre-training of Large Masked Autoencoders for Downstream Earth Observation Tasks?.

MultiMAE Meets Earth Observation: Pre-training Multi-modal Multi-task Masked Autoencoders for Earth Observation Tasks How Effective is Pre-training of Large Masked Autoencoders for Downstream Earth Observation Tasks?

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-07T15:29:53.273043Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:29:53.273043Z digest=sha256:215785b277f191d572f5a3e457aa623729ea5606354d575b0bd9f0a256d1ca5a

Observation 10cc8de0-b885-4996-908d-29d742858b71 · outbound

This paper cites Ssl4eo-l: Datasets and foundation models for landsat imagery.Advances in Neural Information Processing Systems, 36:59787–59807,.

MultiMAE Meets Earth Observation: Pre-training Multi-modal Multi-task Masked Autoencoders for Earth Observation Tasks Ssl4eo-l: Datasets and foundation models for landsat imagery.Advances in Neural Information Processing Systems, 36:59787–59807,

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:29:56.258370Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T15:29:53.395956Z digest=sha256:9eb78734867542a964ffc099c63ef2f7abb71f583aa305f6964f88f86e1295e7

Observation c21eeaab-db24-4163-bf4a-aab4f24a837e · outbound

This paper cites Torch- geo: deep learning with geospatial data.

MultiMAE Meets Earth Observation: Pre-training Multi-modal Multi-task Masked Autoencoders for Earth Observation Tasks Torch- geo: deep learning with geospatial data

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:29:56.024179Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T15:29:53.526054Z digest=sha256:e7830378fca3f8d791d0d9755c7d5f632e914577becb086ff705a8fcf7ca65e3

Observation 9ae240aa-34bf-4b6a-9d9d-a9fb5259c9f6 · outbound

This paper cites Con- vnext v2: Co-designing and scaling convnets with masked autoencoders.

MultiMAE Meets Earth Observation: Pre-training Multi-modal Multi-task Masked Autoencoders for Earth Observation Tasks Con- vnext v2: Co-designing and scaling convnets with masked autoencoders

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:29:55.826704Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T15:29:53.660442Z digest=sha256:302233d16e7ac10d874911988a0d6fea771bb873a99873a770f803f2b208f0f5

Observation 3d7d1e59-244a-4fbc-9c51-2f5aef01a770 · outbound

This paper cites Simmim: A simple framework for masked image modeling.

MultiMAE Meets Earth Observation: Pre-training Multi-modal Multi-task Masked Autoencoders for Earth Observation Tasks Simmim: A simple framework for masked image modeling

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-07T15:29:53.749122Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:29:53.749122Z digest=sha256:ee3ee2851311d240a7196cd08e840557443b262213ebb25f55aa3ed30dfde33c

Observation 0222f482-3379-4dda-a424-a4dd1b8dc624 · outbound

This paper cites EarthNets: Empowering AI in Earth Observation.

MultiMAE Meets Earth Observation: Pre-training Multi-modal Multi-task Masked Autoencoders for Earth Observation Tasks EarthNets: Empowering AI in Earth Observation

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-07T15:29:53.796800Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:29:53.796800Z digest=sha256:3725977d43c7f1b4e51379665fd6fa094bea23cb6f03487fa77b7f0c6e9a0f11

Observation b9a84169-9ba2-4401-bc9c-325d5c3bb5e0 · outbound

This paper cites Neural plasticity-inspired foundation model for observing the earth crossing modalities.arXiv e-prints, pages arXiv–2403, 2024.

MultiMAE Meets Earth Observation: Pre-training Multi-modal Multi-task Masked Autoencoders for Earth Observation Tasks Neural plasticity-inspired foundation model for observing the earth crossing modalities.arXiv e-prints, pages arXiv–2403, 2024

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:29:55.694192Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T15:29:53.836145Z digest=sha256:ad7b51c8b5dbbcf490ed2a5e1184c810ecd7db7f3d2dddfb0a8c92a4eb06f7f7

Observation d41d6eb0-0403-463d-8662-0eba0f109097 · outbound

This paper cites Depth anything: Unleashing the power of large-scale unlabeled data.

MultiMAE Meets Earth Observation: Pre-training Multi-modal Multi-task Masked Autoencoders for Earth Observation Tasks Depth anything: Unleashing the power of large-scale unlabeled data

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:29:55.589747Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T15:29:53.918881Z digest=sha256:bf6aa172bfd91a0f128f5d17d2105e38e97161c8c61fa303c805ac92ab0c791c

Observation 95d7c543-69a8-48b5-8026-4be353fd28b1 · outbound

This paper cites an unresolved cited work.

MultiMAE Meets Earth Observation: Pre-training Multi-modal Multi-task Masked Autoencoders for Earth Observation Tasks Unresolved cited work

Reference 28

Resolution
unresolved
raw_fallback, observed 2026-08-07T15:29:55.493268Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T15:29:53.993947Z digest=sha256:56051040817d13c95a15583b33163785507877b6dbd403b2403e9a6f7b8dbef1

Observation f6b4b861-4512-4648-8e56-2911dac23509 · outbound

This paper cites an unresolved cited work.

MultiMAE Meets Earth Observation: Pre-training Multi-modal Multi-task Masked Autoencoders for Earth Observation Tasks Unresolved cited work

Reference 29

Resolution
unresolved
raw_fallback, observed 2026-08-07T15:29:55.391866Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T15:29:54.050525Z digest=sha256:ae7bb68de1da67900f71f101824351d18c4f3672b9c0b587580c4525351ddeba

Observation ab13c586-0fcc-453f-9e67-00bb951be622 · outbound

This paper cites an unresolved cited work.

MultiMAE Meets Earth Observation: Pre-training Multi-modal Multi-task Masked Autoencoders for Earth Observation Tasks Unresolved cited work

Reference 30

Resolution
unresolved
raw_fallback, observed 2026-08-07T15:29:55.238206Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T15:29:54.147522Z digest=sha256:d7b7bc36f3ecf4302357d6237ef4180427d3dc097842e17cb74b0c05402d256f

Observation 53335596-d842-4a59-8443-1322123d38b7 · outbound

This paper cites Overall, [14] comprises multiple modified versions of standard geospatial datasets for classification and segmen- tation tasks.

MultiMAE Meets Earth Observation: Pre-training Multi-modal Multi-task Masked Autoencoders for Earth Observation Tasks Overall, [14] comprises multiple modified versions of standard geospatial datasets for classification and segmen- tation tasks

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:29:55.031397Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T15:29:54.249838Z digest=sha256:4c5570e41d7140374a5effff33fdf0d524b682bec702fecfb66ef73d34bf96c7

Observation 79d90d38-e04f-4499-af64-20188460bc06 · outbound

This paper cites Pre-training objective We pre-train our approach (depicted in Figure 2) using six input modalities: RGB, IRED, SIRED, EB, DEPTH, and SEG.

MultiMAE Meets Earth Observation: Pre-training Multi-modal Multi-task Masked Autoencoders for Earth Observation Tasks Pre-training objective We pre-train our approach (depicted in Figure 2) using six input modalities: RGB, IRED, SIRED, EB, DEPTH, and SEG

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:29:54.918948Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T15:29:54.305152Z digest=sha256:84375e1a656677a9b113dabbd7bfebccab77b8d958a11a805c9522b096e7d12d

Observation d187ac60-e375-4b51-93e7-a4a2f9d796aa · outbound

This paper cites Fine-tuning setups for segmentation and classification EO tasks.

MultiMAE Meets Earth Observation: Pre-training Multi-modal Multi-task Masked Autoencoders for Earth Observation Tasks Fine-tuning setups for segmentation and classification EO tasks

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:29:54.788137Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T15:29:54.372481Z digest=sha256:97fcceae9dde3235e1d15dbe65d1386b53e256acfb4032b403b66f96f037eee1

Observation 744ff84f-45e1-41f5-adf1-1287d2c7d4c4 · outbound

This paper cites Pre-training visualisations Masked input Prediction Masked input Prediction TargetTarget RGBDEPTHSEGRGBDEPTHSEG Figure 5.

MultiMAE Meets Earth Observation: Pre-training Multi-modal Multi-task Masked Autoencoders for Earth Observation Tasks Pre-training visualisations Masked input Prediction Masked input Prediction TargetTarget RGBDEPTHSEGRGBDEPTHSEG Figure 5

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:29:54.690361Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T15:29:54.475780Z digest=sha256:0eb7a4c005953368498ee407af96eb03829c535427c7308582410700d56b8217

Pith citing papers

Observation d834033a-ee40-40b4-b0da-f4d01a61f1b0 · inbound

Using Multiple Input Modalities Can Improve Data-Efficiency and O.O.D. Generalization for ML with Satellite Imagery cites this paper.

Using Multiple Input Modalities Can Improve Data-Efficiency and O.O.D. Generalization for ML with Satellite Imagery MultiMAE Meets Earth Observation: Pre-training Multi-modal Multi-task Masked Autoencoders for Earth Observation Tasks

Reference 33

Resolution
verified exact
local_arxiv, observed 2026-08-06T17:08:16.257595Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-06T17:08:13.297661Z digest=sha256:3a049012e4fd1b4de61b7fda246216ad4ab65b072b72928dd23c23c92c2f9626