arXiv Machine Learning

Synthetic LiDAR Data Generation and Deterministic Downsampling for Point Cloud Classification on the Edge

arXiv:2608. 07106v1 Announce Type: new Abstract: Deploying three-dimensional deep learning frameworks to low-power embedded processors is bottlenecked by the unstructured nature of spatial data and the resource-intensive distance sorting algorithms often used before neural network inference.

arXiv AI
Sep 12

Lightweight LiDAR-Based Cone Detection Framework Using Random Forest for Formula Student Driverless

The paper introduces a lightweight LiDAR-only perception pipeline for Formula Student Driverless vehicles that runs entirely on CPU. It combines ground removal, IMU-based motion compensation, DBSCAN clustering, and a Random Forest classifier, reducing the feature set from 12 to 7 while maintaining high accuracy. On a dataset of 2,371 labeled clusters, the system achieves an F1-score of 98.33% with an end-to-end runtime of 3.13 ms.

By M\'ark Mez\H{o}-Kerekes, P\'eter Praksz, Chang Liu
arXiv Machine Learning
Aug 13

Achieving Near-Zero-Overhead Multi-Model Hierarchical Classification in Real-Time Detection Pipelines

arXiv:2608. 11770v1 Announce Type: cross Abstract: Edge-deployed vision systems in target recognition, surveillance, autonomous vehicles, and drone domains require hierarchical inference pipelines where a detection model identifies objects of interest and downstream classifiers provide fine-grained attribute analysis.

By Vaishnav Raju
arXiv Computer Vision
Sep 25

M3GD: Multi-Modal Multi-View Geometric Diffusion for Camera--LiDAR Novel View Synthesis

M3GD introduces a multimodal representation that fuses pre‑trained 2D image and 3D LiDAR foundation models for robotic novel view synthesis, avoiding the need for a separate cross‑modal translator. By projecting LiDAR onto the image latent grid and injecting the resulting geometry‑aware packets via a lightweight residual adapter, the method enhances both RGB and depth synthesis on the GrandTour dataset compared to an image‑only baseline. Ablation studies confirm that pixel‑aligned LiDAR content drives the performance gains, and real‑world deployment on a ground robot demonstrates a tunable quality–cost trade‑off.

By Yang Zhou, Jiuhong Xiao, Shizhao Ye, Long Quang, Carlos Nieto-Granda, Giuseppe Loianno