Paper recorded by Signals 4 on 2026-09-04 in cs.CV. Abstract reproduced from arXiv; link to the original below.
Published 2026-09-04 on arXiv · recorded by Signals 4 on 2026-09-07
Category: cs.CV · 计算机视觉 · first seen 2026-09-07
Recent advances in Earth Observation representation learning accommodate heterogeneous sensors and missing observations, often through larger architectures. We present MEOX (Multimodal Earth Observation with eXperts), a multimodal masked autoencoder with a 2.939 million-parameter encoder and 3.115 million parameters in total. Sensor-specific adapters, explicit validity signals, and a shared sparse