arXiv Computer Vision

MarsFM: Shading-Regularized Flow Matching for Martian Relief Estimation

arXiv AI
Sep 15

Multimodal-Multiresolution Foundation Model for Lunar Remote Sensing

The paper introduces a multimodal foundation model for lunar remote sensing, trained from scratch on SomBench—a dataset of nearly two million co‑registered tile bundles across 11 modalities at 1 m and 100 m resolutions. The model extends the TerraMind masked‑token architecture with lunar‑specific features such as explicit acquisition geometry and joint training of two spatial scales, and employs FlexiViT patch embeddings for adaptable patch sizes. Evaluation on crater detection, irregular mare patch segmentation, and polar ice prospectivity regression shows that the pretrained model matches or surpasses ImageNet‑pretrained baselines, with notable label efficiency and effective adaptation via LoRA.

By Paolo Fraccaro, Gabby Nyirjesy, Daniela Szwarcman, Himanshu Patil, Vishal Gaur, Rohit Lal, Rachel A. Slank, Geoffrey Dawson, Hiyam Debary, Michael K. Barker, Andrew Annex, Vishnu Viswanathan, Zachary Morse, Ethan I. Schaefer, Nikolaos Dionelis, Ankur Kumar, Campbell D. Watson, Manil Maskey, Rebekah I. Dawson-Rigas, Juan Bernab\'e-Moreno, Rahul Ramachandran, Sujit Roy
arXiv Computer Vision
Sep 15

Global-Local Contextual Progressive Expansion Network for Martian Landslide Segmentation in Multimodal Remote Sensing Imagery

arXiv:2609.13332v1 Announce Type: new Abstract: Automated landslide segmentation on Mars is one of the important tasks for understanding its surface processes, and all will aid in future space explor...

By Leo Thomas Ramos, Sidike Paheding, Abel A. Reyes-Angulo, Rajaneesh A., Sajinkumar K. S., Angel D. Sappa, Thomas Oommen
arXiv Computer Vision
Sep 3

Adapting a Foundation Model for Lunar Surface Height Estimation

The paper proposes a method to adapt the Depth Anything V2 (DAV2) zero‑shot relative depth model for estimating lunar surface height. By fine‑tuning DAV2 with publicly available stereophotogrammetry‑derived DEM data, the authors achieve a significant performance boost over the unadapted zero‑shot model. This improved estimator can provide more accurate relative height information useful for hazard detection in future ESA lunar landings.

By Patrick Bauer, Marius Schwinning, Melanie Siegel, Andreas Weinmann, Hichem Snoussi
arXiv Machine Learning
Sep 15

SomBench: Benchmark Dataset for Advancing Machine Learning in Lunar Science

arXiv:2609.13277v1 Announce Type: cross Abstract: Lunar orbital missions, such as Lunar Reconnaissance Orbiter, Kaguya/SELENE, Gravity Recovery and Interior Laboratory, and Lunar Prospector, among ot...

By Himanshu Patil, Gabby Nyirjesy, Rachel A. Slank, Vishal Gaur, Daniela Szwarcman, Paolo Fraccaro, Nikolaos Dionelis, Michael K. Barker, Andrew Annex, Vishnu Viswanathan, Zachary Morse, Ethan I. Schaefer, Hiyam Debary, Ankur Kumar, Rohit Lal, Geoffrey Dawson, Campbell Watson, Rebekah I. Dawson-Rigas, Manil Maskey, Juan Bernab\'e-Moreno, Rahul Ramachandran, Sujit Roy
arXiv Computer Vision
Sep 1

MANTLE: A Framework for Adaptive In-Situ Planetary Perception Using a Modular Uplink Principle

MANTLE is a multi‑task adaptive network designed for planetary perception, featuring a shared DINOv2 backbone with separate heads for landform classification and boulder segmentation. Trained on HiRISE and MSL imagery, it achieved 92.56% accuracy on seven Martian terrain classes and a 0.753 IoU for boulder segmentation, with strong cross‑sol generalization. The framework follows the Modular Uplink Principle, allowing lightweight task‑specific heads to be trained on Earth and uplinked to the rover without retraining the core model.

By Pranav Durai, Gary Doran
arXiv Computer Vision
Aug 28

SIMPLER: Efficient Foundation Model Adaptation via Similarity-Guided Layer Pruning for Earth Observation

SIMPLER is a pre‑fine‑tuning method that reduces inference and deployment costs for Earth Observation foundation models by pruning redundant layers. It uses layer‑wise representation similarity on unlabeled task data to identify and remove up to 79% of parameters without requiring gradients, magnitude heuristics, or hyperparameter tuning. Experiments on Prithvi‑EO‑2, TerraMind, and ImageNet‑pretrained ViT‑MAE show that SIMPLER retains 94% of baseline performance while achieving 2.1× faster training and 2.6× faster inference.

By V\'ictor Barreiro, Johannes Jakubik, Francisco Arg\"uello, Dora B. Heras
Hugging Face Trending Papers
Jul 2

Geometric Foundation Model Distillation for Efficient Lunar 3D Reconstruction

Large 3D foundation models such as MASt3R achieve state-of-the-art stereo reconstruction but are computationally demanding for deployment under strict hardware constraints -- a critical limitation in domains such as planetary exploration, where onboard computing is severely restricted. We study how far such models can be compressed through knowledge distillation, using lunar stereo reconstruction as a challenging and practically relevant case study.

arXiv Machine Learning
Sep 4

Distilling deep optical flow stereo methods to retrieve dense three-dimensional wind fields

The paper presents a method to replace traditional window-based tracking in geostationary atmospheric motion vector (AMV) stereo matching with deep optical flow, enabling efficient and accurate retrieval of dense three‑dimensional wind fields. A stereo teacher model is distilled into a single‑satellite student model that emulates the teacher’s uncertainty estimates, allowing global wind generation from full‑disk GEO imagery. Validation against radiosondes, operational AMVs, ERA5 reanalysis, and EarthCARE cloud profiles shows that the stereo winds outperform operational AMVs in water‑vapor bands while performing slightly worse in the long‑wave infrared band.

By Thomas J. Vandal, Dong L. Wu, James L. Carr, Derek J. Posselt, Elise Penn, Tristan Ballard, August Posch, Kate Duffy