CORTEXA
← Browse
crossrefMachine Learning and Knowledge Extraction2026-02-08Cited by 1

Visual Perception and Robust Autonomous Following for Orchard Transportation Robots Based on DeepDIMP-ReID

Renyuan Shen, Yong Wang, Huaiyang Liu, Haiyang Gu, Changxing Geng, Yun Shi

Dense foliage, severe illumination variations, and interference from multiple individuals with similar appearances in complex orchard environments pose significant challenges for vision-based following robots in maintaining persistent target perception and identity consistency, thereby compromising the stability and safety of fruit transportation operations. To address these challenges, we propose a novel framework, DeepDIMP-ReID, which integrates the Deep Implicit Model Prediction (DIMP) tracker with a person re-identification (ReID) module based on EfficientNet. This visual perception and autonomous following framework is designed for differential-drive orchard transportation robots, aiming to achieve robust target perception and reliable identity maintenance in unstructured orchard settings. The proposed framework adopts a hierarchical perception–verification–control architecture. Visual tracking and three-dimensional localization are jointly achieved using synchronized color and depth data acquired from a RealSense camera, where target regions are obtained via the discriminative model prediction (DIMP) method and refined through an elliptical-mask-based depth matching strategy. Front obstacle detection is performed using DBSCAN-based point cloud clustering techniques. To suppress erroneous following caused by occlusion, target switching, or target reappearance after occlusion, an enhanced HOReID person re-identification module with an EfficientNet backbone is integrated for identity verification at critical decision points. Based on the verified perception results, a state-driven motion control strategy is employed to ensure safe and continuous autonomous following. Extensive long-term experiments conducted in real orchard environments demonstrate that the proposed system achieves a correct tracking rate exceeding 94% under varying human walking speeds, with an average localization error of 0.071 m. In scenarios triggering re-identification, a target discrimination success rate of 93.3% is obtained. These results confirm the effectiveness and robustness of the proposed framework for autonomous fruit transportation in complex orchard environments.

View free PDFSource page

Related papers

crossrefMachine Learning and Knowledge Extraction2023-11-07Cited by 19

Reconstruction-Based Adversarial Attack Detection in Vision-Based Autonomous Driving Systems

Manzoor Hussain, Jang-Eui Hong

The perception system is a safety-critical component that directly impacts the overall safety of autonomous driving systems (ADSs). It is imperative to ensure the robustness of the deep-learning model used in the perception system. However, studies have shown that these models ar…

View free PDFSource page
crossrefMachine Learning and Knowledge Extraction2026-07-07

Sensor Fusion and Perception for Autonomous Driving: A Critical Review of Modalities, AI Models, Algorithms, and Industry Configurations

Esraa Khatab, Fares Fathy, Abdallah AlKholy, Omar Shalash

Autonomous driving systems rely on a sophisticated pipeline of artificial intelligence models to perceive, predict, and plan in dynamic environments. This review presents a systematic analysis of the machine learning and deep learning models underpinning vehicle autonomy, spannin…

View free PDFSource page
crossrefMachine Learning and Knowledge Extraction2026-05-11

Knowledge Graphs in Autonomous Driving: Construction, Integration, and Real-Time Reasoning

Patrik Viktor, Gábor Kiss

Autonomous driving systems require the integration of heterogeneous sensor data, distributed V2X communication, and safety-critical decision-making into coherent and interpretable world models. This review provides a systematic analysis of knowledge graph (KG)-based approaches in…

View free PDFSource page
crossrefMachine Learning and Knowledge Extraction2025-08-13Cited by 2

Unsupervised Knowledge Extraction of Distinctive Landmarks from Earth Imagery Using Deep Feature Outliers for Robust UAV Geo-Localization

Zakhar Ostrovskyi, Oleksander Barmak, Pavlo Radiuk, Iurii Krak

Vision-based navigation is a common solution for the critical challenge of GPS-denied Unmanned Aerial Vehicle (UAV) operation, but a research gap remains in the autonomous discovery of robust landmarks from aerial survey imagery needed for such systems. In this work, we propose a…

View free PDFSource page
crossrefMachine Learning and Knowledge Extraction2021-07-15Cited by 69

Recent Advances in Deep Reinforcement Learning Applications for Solving Partially Observable Markov Decision Processes (POMDP) Problems: Part 1—Fundamentals and Applications in Games, Robotics and Natural Language Processing

Xuanchen Xiang, Simon Foo

The first part of a two-part series of papers provides a survey on recent advances in Deep Reinforcement Learning (DRL) applications for solving partially observable Markov decision processes (POMDP) problems. Reinforcement Learning (RL) is an approach to simulate the human’s nat…

View free PDFSource page
crossrefMachine Learning and Knowledge Extraction2025-11-13Cited by 2

Learning to Navigate in Mixed Human–Robot Crowds via an Attention-Driven Deep Reinforcement Learning Framework

Ibrahim K. Kabir, Muhammad F. Mysorewala, Yahya I. Osais, Ali Nasir

The rapid growth of technology has introduced robots into daily life, necessitating navigation frameworks that enable safe, human-friendly movement while accounting for social aspects. Such methods must also scale to situations with multiple humans and robots moving simultaneousl…

View free PDFSource page