arXiv AI By Christoph Legat, Tobias Miller, Marco Riess

Leveraging Deep Learning for Object and Position Recognition of Load Carriers for Autonomous Logistics Vehicles

Read the original on arXiv AI →

arXiv:2606. 16042v1 Announce Type: cross Abstract: This work explores the use of artificial intelligence in mobile robotics to achieve autonomous detection and pose estimation of load carriers for automated pickup.

Machine-generated by The Flow from the publisher's headline and feed description — not written or checked by a human. The full article lives at arXiv AI.

arXiv Computer Vision
Oct 2

GenCOPE: Syn2Real Generalized Category-Level Object Pose Estimation for Robotic Picking

GenCOPE introduces a synthetic-to-real (Syn2Real) approach for category-level object pose estimation (COPE) that eliminates the need for labor-intensive real-world data collection. By learning domain-invariant representations through 2D and 3D semantic consistency constraints and employing an end-to-end pose regression framework with 2D-3D cross consistency, the model achieves robust generalization across synthetic and real domains. The architecture relies solely on global features, resulting in a lightweight and efficient design validated on REAL275, Wild6D, and real-world robotic manipulation scenes.

By Jian Liu, Wei Sun, Zhenqi Dai, Hui Yang, Jian Xiao, Nicu Sebe, Na Zhao
arXiv Computer Vision
Sep 4

An Ensemble-Based Self-Taught Learning Approach for Parking Space Classification Under Limited Data

The paper proposes an ensemble-based self‑taught learning framework for parking space classification that uses unsupervised convolutional autoencoders to learn transferable visual representations from unlabeled data. These learned encoders serve as fixed feature extractors for supervised classification with limited annotated samples, and an ensemble of heterogeneous autoencoders with independent classifier heads is employed to enhance robustness and reduce architectural bias. Experiments on PKLot and CNRPark benchmarks demonstrate that this approach significantly lowers annotation requirements while achieving high accuracies (93–96%) under cross‑dataset evaluation protocols.

By Lucas de Oliveira Cunha, Joelton Deonei Gotz, Paulo Lisboa de Almeida, Andre Gustavo Hochuli