arXiv Machine Learning By Dan Zimmerman, Dimitris A. Pados, George Sklivanitis

Taxonomy-aware deep learning for hierarchical marine species classification in underwater imagery

Read the original on arXiv Machine Learning →

arXiv:2606. 25989v1 Announce Type: cross Abstract: Automated classification of marine species from underwater imagery is essential for scalable ocean biodiversity monitoring and conservation policy.

Machine-generated by The Flow from the publisher's headline and feed description — not written or checked by a human. The full article lives at arXiv Machine Learning.

arXiv Machine Learning
Sep 11

Multimodal Taxonomic Conditioning for Generative Plankton Imagery

The paper presents a method for generating synthetic plankton images conditioned on taxonomic labels to address the long‑tailed nature of automated plankton imaging datasets. A CLIP encoder is fine‑tuned on a large plankton corpus using a ranked contrastive objective that accommodates deep, ragged taxonomies, and then frozen to guide a parameter‑efficient diffusion transformer. The quality of the synthetic samples is evaluated both for distributional fidelity and for their usefulness in training downstream classifiers.

By Daniela Ivanova, Ozgu Goksu, Nicolas Pugeault
arXiv Computer Vision
Aug 26

Comparative Assessment of Deep Learning Architectures for Underwater Subsurface Kelp Forest Segmentation with The Kelp-o-Tron

arXiv:2608.24594v1 Announce Type: new Abstract: Submerged kelp forests are vital coastal ecosystems that support marine biodiversity and ecosystem dynamics, yet accurate underwater kelp segmentation...

By Sundarabalan Balasubramanian, C\'esar Borja, Ana C. Murillo, Lexi N. Wilkes, Meredith L. McPherson, Kira A. Krumhansl, Jennifer A. Dijkstra, Jarrett E. K. Byrnes
arXiv Statistics ML
6d ago

SAGE: A sampling-aware global evaluation benchmark for species distribution modeling

The paper introduces SAGE, a Sampling‑Aware Global Evaluation benchmark for species distribution modeling that uses GBIF records for training and sPlotOpen vegetation plots for presence‑absence evaluation across 5,771 plant species. It groups species by sampling effort and relative prevalence to assess how well single‑species and multi‑species deep‑learning SDMs perform under different data conditions. The study finds that Random Forests and DeepSDMs perform best overall, with DeepSDMs excelling for infrequently recorded species only when bias‑correction techniques are applied.

By Emilia Arens, Nina van Tiel, Robin Zbinden, Damien Robert, Lukas Drees, Chiara Vanalli, Benjamin Kellenberger, Niklaus E. Zimmermann, Lo\"ic Pellissier, Devis Tuia, Jan Dirk Wegner