arXiv AI By Igor Bogdanov, James Green

Infant Care Video Dataset for Classification of Interventions Using Transformers

Read the original on arXiv AI →

The Infant Care Video Dataset (ICVD) contains 4,144 videos covering 12 simulated infant care intervention classes, designed to aid automated documentation in neonatal intensive care units. The dataset was collected using a manikin-based setup that varies camera angles and clinician skin tones while maintaining privacy. Baseline experiments with video transformer models (TimeSformer and MotionFormer) achieved over 93% top‑1 accuracy, whereas a framewise approach scored only 23%, highlighting the importance of temporal modeling for this task.

Machine-generated by The Flow from the publisher's headline and feed description — not written or checked by a human. The full article lives at arXiv AI.

arXiv Computer Vision
Sep 3

Wound3DAssist: A Practical Framework for 3D Wound Assessment

Wound3DAssist is a practical framework that creates 3D wound models from short handheld videos taken with consumer‑grade devices, enabling non‑contact, automatic measurements of wound surfaces. The system integrates 3D reconstruction, wound segmentation, tissue classification, and periwound analysis into a modular workflow. Evaluations on digital models, silicone phantoms, and real patients show millimeter‑scale reconstruction accuracy and multi‑view tissue composition analysis, with full assessments completed in under 20 minutes.

By Remi Chierchia, Rodrigo Santa Cruz, L\'eo Lebrat, Yulia Arzhaeva, Mohammad Ali Armin, Jeremy Oorloff, Chuong Nguyen, Olivier Salvado, Clinton Fookes, David Ahmedt-Aristizabal
arXiv Computer Vision
Sep 3

VideoPulse: Neonatal heart rate and peripheral capillary oxygen saturation (SpO2) estimation from contact free video

VideoPulse is a neonatal dataset and end‑to‑end pipeline that estimates heart rate and peripheral capillary oxygen saturation (SpO2) from facial video without contact. The dataset contains 157 recordings from 52 neonates, and the pipeline uses face alignment, artifact‑aware supervision, and 3D CNN backbones to produce predictions every 2 seconds. On the NBHR dataset the model achieves a heart‑rate MAE of 2.97 bpm and SpO2 MAE of 1.69 %. "whyItMatters":"The results show that short, unaligned neonatal video segments can accurately estimate vital signs, offering a low‑cost, non‑invasive monitoring option for neonatal intensive care."

By Deependra Dewagiri, Kamesh Anuradha, Pabadhi Liyanage, Helitha Kulatunga, Pamuditha Somarathne, Udaya S. K. P. Miriya Thanthrige, Nishani Lucas, Anusha Withana, Joshua P. Kulasingham