arXiv AI By Alan Gerson Contreras Montanares, Luis Valenzuela, Luis Mart\'i, Nayat Sanchez-Pi

Planktonzilla: Multimodal dataset and models for understanding plankton ecosystems

Read the original on arXiv AI →

arXiv:2606. 00080v1 Announce Type: cross Abstract: Marine plankton underpin aquatic food webs and play a key role in global CO2 sequestration, making reliable species identification critical for understanding ocean health and climate feedbacks.

Machine-generated by The Flow from the publisher's headline and feed description — not written or checked by a human. The full article lives at arXiv AI.

arXiv Machine Learning
Sep 11

Multimodal Taxonomic Conditioning for Generative Plankton Imagery

The paper presents a method for generating synthetic plankton images conditioned on taxonomic labels to address the long‑tailed nature of automated plankton imaging datasets. A CLIP encoder is fine‑tuned on a large plankton corpus using a ranked contrastive objective that accommodates deep, ragged taxonomies, and then frozen to guide a parameter‑efficient diffusion transformer. The quality of the synthetic samples is evaluated both for distributional fidelity and for their usefulness in training downstream classifiers.

By Daniela Ivanova, Ozgu Goksu, Nicolas Pugeault
arXiv Computer Vision
Aug 24

WildFin: An In-the-Wild Dataset for Fish Behavioral Recognition

arXiv:2608.21281v1 Announce Type: new Abstract: Recent advances in field technology have led to a massive influx of in-the-wild video data for ecological science. The primary bottleneck in leveraging...

By Abigail G. Grassick, Jerome Tze-Hou Hsu, Ethan Lin, Ziang Liu, Max Whitton, Madelyn Hair, Liam Gutierrez, Haozheng Yu, Kristin Branson, Vivek Jayaraman, Michael A. Gil, Andrew M. Hein, Jennifer J. Sun
arXiv Machine Learning
Aug 19

Leveraging existing sparse point annotations for benthic imagery dense segmentation

The paper presents a method that leverages sparse expert point annotations from historical benthic surveys to improve dense segmentation of marine imagery. By using these points as visual prompts for the SAM2 foundation model and introducing a mechanism to filter out unreliable points, the authors generate high‑quality pseudo‑ground‑truth masks that train more accurate fine‑grained semantic segmentation models. The approach is validated on public benthic datasets and a new benchmark featuring real‑world sparse annotations, aiming to enable scalable ecological analysis.

By Cesar Borja, Breck A. McCollum, Jarret E. Byrnes, Kenneth Sebens, Ana C. Murillo
arXiv Computer Vision
Sep 14

CoralscapesV2: Panoptic and Fine-Grained Visual Scene Understanding in Coral Reefs

CoralscapesV2 is an expanded dataset for coral reef visual scene understanding, increasing the number of fine‑grained classes from 39 to 95 and adding 65,000 exhaustive fish instance masks. It supports panoptic segmentation by providing high‑quality semantic and instance labels across diverse, unconstrained reef imagery. The dataset serves as a challenging benchmark for modern segmentation models and enables broader applications such as benthic cover mapping and automated fish‑reef interaction analysis.

By Jonathan Sauder, Thomas Ruckli, Gabriel\.e Strodomskyt\.e, Ibrahim Souleiman Abdallah, Rahma Hassan Abdi, Djama Goumaneh Awaleh, Mohamed Houssein Farah, Moustapha Nour, Osama Sharhubil Saad, Mustafa Mohammed Khalafallah Altaib, Maysoon Kteifan, Farah Alsoqi, Eyad Zgool, Jafar Al-Omari, Temesgen Gebremeskel Gebreluel, Zekaria Zekeria Abdulkerim, Meron Ghirmay, Teklehaimanot Beraki, Devis Tuia, Guilhem Banc-Prandi