paper-with-me

홈 › Papers

Query Processing on Tensor Computation Runtimes

2022-03-03 · Dong He, Supun Nakandala, Dalitso Banda, Rathijit Sen, Karla Saur, Kwanghyun Park, Carlo Curino, Jesús Camacho-Rodríguez, Konstantinos Karanasos, Matteo Interlandi

The huge demand for computation in artificial intelligence (AI) is driving unparalleled investments in hardware and software systems for AI. This leads to an explosion in the number of specialized hardware devices, which are now offered by major cloud vendors. By hiding the low-level complexity through a tensor-based interface, tensor computation runtimes (TCRs) such as PyTorch allow data scientists to efficiently exploit the exciting capabilities offered by the new hardware. In this paper, we explore how database management systems can ride the wave of innovation happening in the AI space. We design, build, and evaluate Tensor Query Processor (TQP): TQP transforms SQL queries into tensor programs and executes them on TCRs. TQP is able to run the full TPC-H benchmark by implementing novel algorithms for relational operators on the tensor routines. At the same time, TQP can support various hardware while only requiring a fraction of the usual development effort. Experiments show that TQP can improve query execution time by up to 10$\times$ over specialized CPU- and GPU-only systems. Finally, TQP can accelerate queries mixing ML predictions and SQL end-to-end, and deliver up to 9$\times$ speedup over CPU baselines.

📄 PDF Abstract BibTeX arXiv:2203.01877

Code (0)

등록된 구현이 없습니다.

Tasks

CPUGPUManagement

Similar Papers 제목 키워드 기반

Minimizing the Number of Matching Queries for Object Retrieval

2014-12-18 · Johannes Niedermayer, Peer Kröger

To increase the computational efficiency of interest-point based object retrieval, researchers have put remarkable research efforts into improving the efficiency of kNN-based feature matching, pursuing to match thousands…

Computational EfficiencyObjectRetrieval

Iskra: A System for Inverse Geometry Processing

2026-02-12 · Ana Dodik, Ahmed H. Mahmoud, Justin Solomon arxiv

We propose a system for differentiating through solutions to geometry processing problems. Our system differentiates a broad class of geometric algorithms, exploiting existing fast problem-specific schemes common to geom…

Fast Linear Interpolation for Piecewise-Linear Functions, GAMs, and Deep Lattice Networks

2019-09-25 · Nathan Zhang, Kevin Canini, Sean Silva, and Maya R. Gupta

We present fast implementations of linear interpolation operators for both piecewise linear functions and multi-dimensional look-up tables. We use a compiler-based solution (using MLIR) for accelerating this family of wo…

CPU

Rollplex: Cross-Phase GPU Spatial Sharing for Vision Language Model Post-Training

2026-08-14 · Hanfeng Lu, Tianyu Feng, Suyi Li, Yuheng Zhao 외 arxiv

Vision-language models (VLMs) enable embodied agents to reason and act from visual observations and language instructions. Reinforcement learning (RL) post-training enhances these capabilities using task feedback, but cu…

Reinforcement Learning

A Tensor Compiler for Unified Machine Learning Prediction Serving

2020-10-09 · Supun Nakandala, Karla Saur, Gyeong-In Yu, Konstantinos Karanasos 외

Machine Learning (ML) adoption in the enterprise requires simpler and more efficient software infrastructure---the bespoke solutions typical in large web companies are simply untenable. Model scoring, the process of obta…

BIG-bench Machine LearningCPUGPU