Paper Reading AI Learner

Genetic Algorithms For Parameter Optimization for Disparity Map Generation of Radiata Pine Branch Images

2025-12-05 04:00:18
Yida Lin, Bing Xue, Mengjie Zhang, Sam Schofield, Richard Green

Abstract

Traditional stereo matching algorithms like Semi-Global Block Matching (SGBM) with Weighted Least Squares (WLS) filtering offer speed advantages over neural networks for UAV applications, generating disparity maps in approximately 0.5 seconds per frame. However, these algorithms require meticulous parameter tuning. We propose a Genetic Algorithm (GA) based parameter optimization framework that systematically searches for optimal parameter configurations for SGBM and WLS, enabling UAVs to measure distances to tree branches with enhanced precision while maintaining processing efficiency. Our contributions include: (1) a novel GA-based parameter optimization framework that eliminates manual tuning; (2) a comprehensive evaluation methodology using multiple image quality metrics; and (3) a practical solution for resource-constrained UAV systems. Experimental results demonstrate that our GA-optimized approach reduces Mean Squared Error by 42.86% while increasing Peak Signal-to-Noise Ratio and Structural Similarity by 8.47% and 28.52%, respectively, compared with baseline configurations. Furthermore, our approach demonstrates superior generalization performance across varied imaging conditions, which is critcal for real-world forestry applications.

Abstract (translated)

传统的立体匹配算法,如半全局块匹配(Semi-Global Block Matching, SGBM)结合加权最小二乘滤波(Weighted Least Squares, WLS),在无人飞行器(UAV)应用中比神经网络具有速度优势,能够在大约0.5秒内生成每帧的视差图。然而,这些算法需要仔细调整参数。我们提出了一种基于遗传算法(Genetic Algorithm, GA)的参数优化框架,该框架系统地搜索SGBM和WLS的最佳参数配置,使UAV能够以更高的精度测量到树干的距离,同时保持处理效率。我们的贡献包括:(1) 一种新颖的GA基参数优化框架,消除了手动调优的需求;(2) 使用多个图像质量度量的全面评估方法;以及 (3) 针对资源受限UAV系统的实用解决方案。 实验结果表明,我们基于GA优化的方法与基准配置相比,在减少均方误差(Mean Squared Error)方面提高了42.86%,同时将峰值信噪比(Peak Signal-to-Noise Ratio)和结构相似性(Structural Similarity)分别提升了8.47% 和 28.52%。此外,我们的方法在不同的成像条件下表现出色的泛化性能,这对于现实世界的林业应用至关重要。

URL

https://arxiv.org/abs/2512.05410

PDF

https://arxiv.org/pdf/2512.05410.pdf


Tags
3D Action Action_Localization Action_Recognition Activity Adversarial Agent Attention Autonomous Bert Boundary_Detection Caption Chat Classification CNN Compressive_Sensing Contour Contrastive_Learning Deep_Learning Denoising Detection Dialog Diffusion Drone Dynamic_Memory_Network Edge_Detection Embedding Embodied Emotion Enhancement Face Face_Detection Face_Recognition Facial_Landmark Few-Shot Gait_Recognition GAN Gaze_Estimation Gesture Gradient_Descent Handwriting Human_Parsing Image_Caption Image_Classification Image_Compression Image_Enhancement Image_Generation Image_Matting Image_Retrieval Inference Inpainting Intelligent_Chip Knowledge Knowledge_Graph Language_Model LLM Matching Medical Memory_Networks Multi_Modal Multi_Task NAS NMT Object_Detection Object_Tracking OCR Ontology Optical_Character Optical_Flow Optimization Person_Re-identification Point_Cloud Portrait_Generation Pose Pose_Estimation Prediction QA Quantitative Quantitative_Finance Quantization Re-identification Recognition Recommendation Reconstruction Regularization Reinforcement_Learning Relation Relation_Extraction Represenation Represenation_Learning Restoration Review RNN Robot Salient Scene_Classification Scene_Generation Scene_Parsing Scene_Text Segmentation Self-Supervised Semantic_Instance_Segmentation Semantic_Segmentation Semi_Global Semi_Supervised Sence_graph Sentiment Sentiment_Classification Sketch SLAM Sparse Speech Speech_Recognition Style_Transfer Summarization Super_Resolution Surveillance Survey Text_Classification Text_Generation Time_Series Tracking Transfer_Learning Transformer Unsupervised Video_Caption Video_Classification Video_Indexing Video_Prediction Video_Retrieval Visual_Relation VQA Weakly_Supervised Zero-Shot