Papers CPU
“CPU” 태그가 달린 논문 2,231편 · 필터 해제
Hear Your Code Fail, Voice-Assisted Debugging for Python
This research introduces an innovative voice-assisted debugging plugin for Python that transforms silent runtime errors into actionable audible diagnostics. By implementing a global exception hook architecture with pytts…
CPUMedical Diagnosistext-to-speechText to Speech3C-FBI: A Combinatorial method using Convolutions for Circle Fitting in Blurry Images
This paper addresses the fundamental computer vision challenge of robust circle detection and fitting in degraded imaging conditions. We present Combinatorial Convolution-based Circle Fitting for Blurry Images (3C-FBI), …
CPUDensity EstimationInter2Former: Dynamic Hybrid Attention for Efficient High-Precision Interactive
Interactive segmentation (IS) improves annotation efficiency by segmenting target regions from user prompts, with widespread applications in real-world scenarios. Current approaches face a critical trade-off: dense-token…
CPUInteractive SegmentationMixture-of-ExpertsMathOptAI.jl: Embed trained machine learning predictors into JuMP models
We present \texttt{MathOptAI.jl}, an open-source Julia library for embedding trained machine learning predictors into a JuMP model. \texttt{MathOptAI.jl} can embed a wide variety of neural networks, decision trees, and G…
CPUGaussian ProcessesGPULoRA Fine-Tuning Without GPUs: A CPU-Efficient Meta-Generation Framework for LLMs
Low-Rank Adapters (LoRAs) have transformed the fine-tuning of Large Language Models (LLMs) by enabling parameter-efficient updates. However, their widespread adoption remains limited by the reliance on GPU-based training…
CPUGPUAUTOMATIC ROOM LIGHT CONTROLLER MANAGEMENT SYSTEM.
The AT89S51 is a low-power, high- performance CMOS 8-bit microcontroller with 4K bytes of In-System Programmable Flash memory. The device is manufactured using Atmel’s high-density non-volatile memory technology and is c…
4kCPUManagementMNN-AECS: Energy Optimization for LLM Decoding on Mobile Devices via Adaptive Core Selection
As the demand for on-device Large Language Model (LLM) inference grows, energy efficiency has become a major concern, especially for battery-limited mobile devices. Our analysis shows that the memory-bound LLM decode pha…
CPULarge Language ModelCausal-Aware Intelligent QoE Optimization for VR Interaction with Adaptive Keyframe Extraction
The optimization of quality of experience (QoE) in multi-user virtual reality (VR) interactions demands a delicate balance between ultra-low latency, high-fidelity motion synchronization, and equitable resource allocatio…
Causal InferenceCPUFairnessReinforcement Learning (RL)LIGHTHOUSE: Fast and precise distance to shoreline calculations from anywhere on earth
We introduce a new dataset and algorithm for fast and efficient coastal distance calculations from Anywhere on Earth (AoE). Existing global coastal datasets are only available at coarse resolution (e.g. 1-4 km) which lim…
CPUVariational Bayesian Channel Estimation and Data Detection for Cell-Free Massive MIMO with Low-Resolution Quantized Fronthaul Links
We study the joint channel estimation and data detection (JED) problem in a cell-free massive multiple-input multiple-output (CF-mMIMO) network, where access points (APs) communicate with a central processing unit (CPU) …
CPUQuantizationConsumerBench: Benchmarking Generative AI Applications on End-User Devices
The recent shift in Generative AI (GenAI) applications from cloud-only environments to end-user devices introduces new challenges in resource management, system efficiency, and user experience. This paper presents Consum…
BenchmarkingCPUGPUSchedulingSpeeding up Local Optimization in Vehicle Routing with Tensor-based GPU Acceleration
Local search plays a central role in many effective heuristic algorithms for the vehicle routing problem (VRP) and its variants. However, neighborhood exploration is known to be computationally expensive and time consumi…
AttributeComputational EfficiencyCPUGPUWavelet-based Global Orientation and Surface Reconstruction for Point Clouds
Unoriented surface reconstruction is an important task in computer graphics and has extensive applications. Based on the compact support of wavelet and orthogonality properties, classic wavelet surface reconstruction ach…
CPUSurface ReconstructionDistributed Activity Detection for Cell-Free Hybrid Near-Far Field Communications
A great amount of endeavor has recently been devoted to activity detection for massive machine-type communications in cell-free massive MIMO. However, in practice, as the number of antennas at the access points (APs) inc…
Action DetectionActivity DetectionCPUParallel Branch Model Predictive Control on GPUs
We present a parallel GPU-accelerated solver for branch Model Predictive Control problems. Based on iterative LQR methods, our solver exploits the tree-sparse structure and implements temporal parallelism using the paral…
CPUGPUmodelModel Predictive ControlVersatile and Fast Location-Based Private Information Retrieval with Fully Homomorphic Encryption over the Torus
Location-based services often require users to share sensitive locational data, raising privacy concerns due to potential misuse or exploitation by untrusted servers. In response, we present VeLoPIR, a versatile location…
CPUGPUInformation RetrievalSecONNds: Secure Outsourced Neural Network Inference on ImageNet
The widespread adoption of outsourced neural network inference presents significant privacy challenges, as sensitive user data is processed on untrusted remote servers. Secure inference offers a privacy-preserving soluti…
CPUGPUPrivacy PreservingMNN-LLM: A Generic Inference Engine for Fast Large Language Model Deployment on Mobile Devices
Large language models (LLMs) have demonstrated exceptional performance across a variety of tasks. However, their substantial scale leads to significant computational resource consumption during inference, resulting in hi…
CPUGPULanguage ModelingLanguage Modelling+2RT-VC: Real-Time Zero-Shot Voice Conversion with Speech Articulatory Coding
Voice conversion has emerged as a pivotal technology in numerous applications ranging from assistive communication to entertainment. In this paper, we present RT-VC, a zero-shot real-time voice conversion system that del…
CPUVoice ConversionHPCTransCompile: An AI Compiler Generated Dataset for High-Performance CUDA Transpilation and LLM Preliminary Exploration
The rapid growth of deep learning has driven exponential increases in model parameters and computational demands. NVIDIA GPUs and their CUDA-based software ecosystem provide robust support for parallel computing, signifi…
CPUData Augmentation