Local Ai

Interpretable Fuzzy Inference for UAV Target Tracking Using Bounding-Box Geometry

arXiv:2608.04121v1 Announce Type: cross Abstract: Vision-based guidance of unmanned aerial vehicles (UAVs) toward unmanned ground vehicles (UGVs) supports cooperative aerial--ground robotics, but reli

DGX agentpaper
local-aiarxiv-cs-ai

arXiv:2608.04121v1 Announce Type: cross Abstract: Vision-based guidance of unmanned aerial vehicles (UAVs) toward unmanned ground vehicles (UGVs) supports cooperative aerial--ground robotics, but reliable continuous yaw estimation from onboard vision remains challenging because of sensing uncertainty, limited computation, and the need for interpretable control. Existing deep-learning and geometric-reconstruction approaches often require large datasets, external localization, or complex modeling assumptions, reducing transparency and deployment suitability on resource-constrained platforms. We present an interpretable fuzzy-inference framework that generates continuous yaw commands from low-dimensional features extracted from YOLO boxes: target centroid location, area, and aspect ratio. No explicit geometric modeling is required. A Mamdani fuzzy system serves as an interpretable baseline using a shoulder--triangle--shoulder input partition. It is followed by a first-order Takagi--Sugeno model with three antecedent membership terms per input, whose parameters are derived from training-set quantiles, yielding a compact 27-rule structure. Evaluation uses 6{,}169 labeled samples from a VICON motion-capture environment. Across five randomized train--test splits, the Takagi--Sugeno model achieves a test-set mean absolute error of 0.140^irc pm 0.003^irc, a root mean squared error of 0.200^irc pm 0.008^irc, and a maximum absolute error of 1.254^irc pm 0.121^irc. Within-threshold accuracies are 99.676% pm 0.270% for pm1^irc and 100.000% pm 0.000% for both pm3^irc and pm5^irc. Directional consistency between image-plane horizontal displacement and predicted yaw sign reaches 90.254% pm 0.612%. These results show that the framework is transparent, data-efficient, computationally lightweight, and suitable for real-time vision-based UAV guidance toward mobile ground targets.

Source: arXiv cs.AI | 2026-08-06

Loading related sources…