paper-with-me

홈 › Papers

A Multi-Dimensional Quality Scoring Framework for Decentralized LLM Inference with Proof of Quality

2026-03-04 · Arther Tian, Alex Ding, Frank Chen, Simon Wu, Aaron Chan arxiv

Decentralized large language model (LLM) inference networks can pool heterogeneous compute to scale serving, but they require lightweight and incentive-compatible mechanisms to assess output quality. Prior work introduced cost-aware Proof of Quality (PoQ) and adaptive robust PoQ to allocate rewards under evaluator heterogeneity and adversarial behavior. In this paper, we focus on the quality signal itself and propose a multi-dimensional quality scoring framework that decomposes output quality into modular dimensions, including model and cost priors, structure quality, semantic quality, query-output alignment, and agreement/uncertainty. Using logged outputs from QA and summarization tasks, we systematically audit dimension reliability and show that seemingly reasonable dimensions can be task-dependent and even negatively correlated with reference quality without calibration. While the default composite underperforms a strong single semantic evaluator, ablations reveal that removing unreliable dimensions and re-normalizing weights yields a calibrated composite that matches or exceeds the best single- evaluator and consensus baselines. Finally, we integrate the composite score as a drop-in quality signal in PoQ and demonstrate complementary benefits with robust aggregation and adaptive trust weighting under adversarial evaluator attacks.

📄 PDF Abstract BibTeX arXiv:2603.04028

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

A Decentralized Retrieval Augmented Generation System with Source Reliabilities Secured on Blockchain

2025-11-10 · Yining Lu, Wenyi Tang, Max Johnson, Taeho Jung 외 arxiv

Existing retrieval-augmented generation (RAG) systems typically use a centralized architecture, causing a high cost of data collection, integration, and management, as well as privacy concerns. There is a great need for …

PoQ-Judge: A Multi-Architecture Evaluation Framework for Cost-Aware Proof-of-Quality in Decentralized LLM Inference

2026-04-20 · Arther Tian, Alex Ding, Frank Chen, Simon Wu 외 arxiv

Decentralized LLM inference networks need lightweight, reference-free quality evaluation for Proof of Quality (PoQ). We present PoQ-Judge, a framework that trains dedicated judge models to score query-output pairs withou…

Auction-Consensus Algorithm with Learned Bidding Scheme for Multi-Robot Systems

2026-05-21 · Jose Rodriguez, Constantine Tarawneh, Sven Koenig, Wenjie Dong 외 arxiv

Multi-Robot Task Allocation (MRTA) is a central challenge in decentralized multi-agent systems, where teams of robots must cooperatively assign and execute tasks under limited communication while optimizing global perfor…

Reinforcement Learning

Multidimensional Service Quality Scoring System

2022-12-09 · Shiyang Lai

This supplementary paper aims to introduce the Multidimensional Service Quality Scoring System (MSQs), a review-based method for quantifying host service quality mentioned and employed in the paper Exit and transition: E…

The Multi-Range Theory of Translation Quality Measurement: MQM scoring models and Statistical Quality Control

2024-05-27 · Arle Lommel, Serge Gladkoff, Alan Melby, Sue Ellen Wright 외

The year 2024 marks the 10th anniversary of the Multidimensional Quality Metrics (MQM) framework for analytic translation quality evaluation. The MQM error typology has been widely used by practitioners in the translatio…

Machine TranslationSentenceTranslation