paper-with-me

홈 › Papers

Global optimization tailored for graphics processing units: Complete and rigorous search for large-scale nonlinear minimization

2025-07-02 · Guanglu Zhang, Qihang Shan, Jonathan Cagan arxiv

This paper introduces a numerical method to enclose the global minimum of a nonlinear function subject to simple bounds on the variables. Using interval analysis, coupled with the computational power and architecture of graphics processing units (GPUs), the method iteratively rules out the regions in the search domain where the global minimum cannot exist and leaves a finite set of regions where the global minimum must exist. For effectiveness, because of the rigor of interval analysis, the method is guaranteed to enclose the global minimum even in the presence of rounding errors. For efficiency, the method employs a novel GPU-based single program, single data parallel programming style to circumvent major GPU performance bottlenecks, and a variable cycling technique is also integrated into the method to reduce computational cost when minimizing large-scale nonlinear functions. The method is validated by minimizing 11 benchmark test functions with scalable dimensions, including the well-known Ackley function, Griewank function, Levy function, Rastrigin function, and Rosenbrock function. These benchmark test functions represent grand challenges of global optimization, and enclosing the guaranteed global minimum of these benchmark test functions with more than 80 dimensions has not been reported in the literature. Our method completely searches the feasible domain and successfully encloses the guaranteed global minimum of these 11 benchmark test functions with up to 10,000 dimensions using only one GPU in a reasonable computation time, far exceeding the reported results in the literature due to the unique method design and implementation based on GPU architecture.

📄 PDF Abstract BibTeX arXiv:2507.01770

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Accelerating Discrete Wavelet Transforms on GPUs

2017-05-18 · David Barina, Michal Kula, Michal Matysek, Pavel Zemcik

The two-dimensional discrete wavelet transform has a huge number of applications in image-processing techniques. Until now, several papers compared the performance of such transform on graphics processing units (GPUs). H…

On the Suitability of Neural Networks as Building Blocks for The Design of Efficient Learned Indexes

2022-02-21 · Domenico Amato, Giosue' Lo Bosco, Raffaele Giancarlo

With the aim of obtaining time/space improvements in classic Data Structures, an emerging trend is to combine Machine Learning techniques with the ones proper of Data Structures. This new area goes under the name of Lear…

LANCE: Efficient Low-Precision Quantized Winograd Convolution for Neural Networks Based on Graphics Processing Units

2020-03-19 · Guangli Li, Lei Liu, Xueying Wang, Xiu Ma 외

Accelerating deep convolutional neural networks has become an active topic and sparked an interest in academia and industry. In this paper, we propose an efficient low-precision quantized Winograd convolution algorithm, …

image-classificationImage ClassificationQuantization

Implementation of Real-Time Automotive SAR Imaging

2023-06-16 · Marcel Hoffmann, Theresa Noegel, Christian Schüßler, Lars Schwenger 외

This paper presents measures to reduce the computation time of automotive synthetic aperture radar (SAR) imaging to achieve real-time capability. For this, the image formation, which is based on the Back-Projection algor…

GPU

Computer Vision Accelerators for Mobile Systems based on OpenCL GPGPU Co-Processing

2014-03-17 · Guohui Wang, Yingen Xiong, Jay Yun, Joseph R. Cavallaro

In this paper, we present an OpenCL-based heterogeneous implementation of a computer vision algorithm -- image inpainting-based object removal algorithm -- on mobile devices. To take advantage of the computation power of…

CPUGPUImage Inpainting