arXiv Computer Vision

The Impact of Processing Parameters on High-Accuracy Measurements in UAV Photogrammetry

The paper investigates how processing parameters affect the accuracy of UAV photogrammetry. By testing 768 processing variants on ten datasets over 1.5 years, the study finds that the best configuration yields an RMSE of 16 mm, while the worst reaches 303 mm. Key factors include the number of ground control points, camera calibration corrections, and the use of Post‑Processing Kinematic GNSS for camera center determination, which together reduce systematic errors by more than half.

arXiv Computer Vision
Aug 28

Camera Calibration Using Inaccurate and Asynchronous Discrete GPS Trajectory from Drones

The paper tackles the problem of calibrating a stationary camera’s yaw, pitch, and roll using a drone’s GPS trajectory, which suffers from altitude bias, time offset, and discrete sampling. It formulates a parameter estimation problem that jointly estimates the GPS altitude bias, time offset, and camera orientation biases, and proposes a maximum likelihood estimator based on Iterated Least Squares to handle the asynchronous, discrete GPS data. Simulation results show the estimator achieves accuracy close to the Cramér–Rao Lower Bound, with a recommended drone trajectory yielding calibration errors within 14% of the measurement error standard deviation.

By R. Yang, Y. Bar-Shalom, H. A. J. Huang
arXiv Machine Learning
Aug 10

UAV3DCrop: Benchmarking 3D Reconstruction in Repeated Multi-Angle UAV Crop Surveys

arXiv:2608. 06404v1 Announce Type: cross Abstract: Accurate 3D crop monitoring underpins data-driven precision agriculture by enabling field-scale analysis of plant structure, growth dynamics, and management response.

By Junxiong Zhou, Xuechen Li, Chonghao Qiu, Lang Qiao, Xiaowei Jia, Qi Yang, Chishan Zhang, Leikun Yin, Nanshan You, Vipin Kumar, David Mulla, Ce Yang, Zhenong Jin, Licheng Liu
arXiv Computer Vision
Sep 3

UAV Thermal Imagery for Inert Ordnance Screening: Multi Campaign Dataset Development,Object Detection, and Practical Recommendations

The paper presents a multi‑campaign UAV thermal image dataset for inert ordnance screening, comprising 5,855 labeled image pairs collected in Tennessee across diverse terrains and seasons. The authors trained YOLOV11l and RT‑DETR‑R50 models on 33 m and 15 m altitude data, achieving automated candidate detection, and provided practical guidelines for future humanitarian mine action surveys. The dataset and models aim to aid screening and prioritization for follow‑up technical surveys or EOD assessment, not to replace clearance operations.

By Chad Melton, PhD., Annabelle Kelton
arXiv Computer Vision
Sep 3

A Lightweight Multi-Metric No-Reference Image Quality Assessment Framework for UAV Imaging

The paper presents MM‑IQA, a lightweight no‑reference image quality assessment framework designed for UAV imaging. It fuses interpretable metrics—blur, edge structure, low‑resolution artifacts, exposure imbalance, noise, haze, and frequency content—to output a single quality score between 0 and 100. Evaluated on five benchmark datasets, MM‑IQA achieved SRCC values from 0.647 to 0.830 and runs in about 1.97 s per image with modest memory usage.

By Koffi Titus Sergio Aglin, Anthony K. Muchiri, Celestin Nkundineza
arXiv Computer Vision
Aug 25

3D Point Cloud from Close-Range Photogrammetry for Defect Characterisation of Rubberised Concrete

The paper presents a close‑range photogrammetry workflow using Structure‑from‑Motion and Multi‑View Stereo to generate high‑resolution 3D point clouds of rubberised concrete. By capturing images with a Canon DSLR and an iPhone 16, the authors achieved sub‑millimetre reconstruction accuracy, outperforming traditional LiDAR for fine‑scale defect analysis. An RGB‑guided crack extraction method and deformation analysis further demonstrate the method’s utility for detailed surface monitoring and material performance evaluation.

By Jiacheng Liu, Mohammed Alnahhal, Ailar Hajimohammadi, Sara Gonizzi Barsanti, Jinling Wang, Mohsen Kalantari
Hugging Face Trending Papers
Jun 22

Scene-agnostic ALS boresight self-calibration

ALS boresight calibration has relied for two decades on dedicated flight patterns over structured scenes containing planar surfaces of varied aspect and slope. While reliable, this approach imposes constraints on the scene content and operations, which limits its applicability to boresight recovery within routine mapping missions.

arXiv Computer Vision
Sep 4

Hold-Out Self-Validation Cannot Certify Photogrammetric Accuracy: Saturation and Blindness to Coherent Distortion

The paper argues that internal self-consistency checks cannot guarantee the accuracy of photogrammetric reconstructions, a limitation that is structural rather than a tuning issue. It introduces a track‑leakage‑free hold‑out protocol that withholds a deterministic subset of images and tests each against only 3D points supported by at least two retained images, ensuring no view is evaluated against the structure it helped create. Experiments on diverse datasets show that while the protocol is well‑posed, it saturates at a confidence score of 1.00 and fails to detect coherent distortion, missing large errors that can reach over 100 m. whyItMatters":"The study highlights that hold‑out self‑validation scores, increasingly used as quality evidence for metric deliverables, may be misleading and cannot replace external survey validation."

By Behnam Asadi
arXiv Computer Vision
Sep 24

Beyond Balanced Accuracy: A Resolution and Parity-Controlled Benchmark for Vision-Language and Vision-Only Defect Assessment in UAV Power-Line Inspection

The paper evaluates the claim that vision‑language models (VLMs) outperform task‑specific vision backbones for UAV power‑line defect assessment using the ElecVQA‑Bench benchmark. Across various evaluation settings—partitioning, item sets, label spaces, replication, resolution, and side information—the performance gap between VLMs and traditional backbones is minimal or even reversed when controlling for resolution and token budget. The study concludes that VLM superiority is not universally supported and emphasizes the importance of rigorous benchmark audits.

By Linghao Zhang, Siyu Xiang, Junwei Kuang, Peiyu Yi