paper-with-me

홈 › Papers

Making AI-Assisted Grant Evaluation Auditable without Exposing the Model

2026-04-28 · Kemal Bicakci arxiv

Public agencies are beginning to consider large language models (LLMs) as decision-support tools for grant evaluation. This creates a practical governance problem: the model and scoring rubric should not be exposed in a way that allows applicants to optimize against them, yet the evaluation process must remain auditable, contestable, and accountable. We propose a TEE-based architecture that helps reconcile these requirements through remote attestation. The architecture allows an external verifier to check which model, rubric, prompt template, and input representation were used, without exposing model weights, proprietary scoring logic, or intermediate reasoning to applicants or infrastructure operators. The main artifact is an attested evaluation bundle: a signed, timestamped record linking the original submission hash, the canonical input hash, the model-and-rubric measurement, and the evaluation output. The paper also considers a scenario-specific prompt injection risk: applicant-controlled documents may contain hidden or indirect instructions intended to influence the LLM evaluator. We therefore include a canonicalization and sanitization layer that normalizes document representations and records suspicious transformations before inference. We position the design relative to confidential AI inference, attestable AI audits, zero-knowledge machine learning, algorithmic accountability, and AI-assisted peer review. The resulting claim is deliberately narrow: remote attestation does not prove that an evaluation is fair or scientifically correct, but it can make part of the evaluation process externally verifiable.

📄 PDF Abstract BibTeX arXiv:2604.25200

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Telling stories, making Hanzi: AI-assisted co-creation with elderly migrants in urban China

2025-07-02 · Yunfei Chen, Wen Zhan, Peiyue Lin, Ziqun Hua 외 arxiv

This paper explores how older migrants in urban China can record stories that everyday language and design often miss. We ran two co-creation workshops with 10 elders. Activities combined oral storytelling, facilitator-m…

Making Clinical Language Models Auditable: Concept-Guided Fine-Tuning for Robust Prediction

2026-08-27 · Jin Mu, Guanhua Chen arxiv

Clinical language models can achieve strong in-hospital accuracy yet fail under deployment shifts because they exploit note-specific artifacts (e.g., templates, separators, boilerplate) that do not reflect patient state.…

Mortality PredictionText Classification

RW-Post: Auditable Evidence-Grounded Multimodal Fact-Checking in the Wild

2025-12-28 · Danni Xu, Shaojing Fan, Harry Cheng, Mohan Kankanhalli arxiv

Multimodal misinformation increasingly leverages visual persuasion, where repurposed or manipulated images strengthen misleading text. We introduce RW-Post, a post-aligned text--image benchmark for real-world multimodal …

Visual Grounding

Evaluating LLM-Based Grant Proposal Review via Structured Perturbations

2026-03-09 · William Thorne, Joseph James, Yang Wang, Chenghua Lin 외 arxiv

As AI-assisted grant proposals outpace manual review capacity in a kind of ``Malthusian trap'' for the research ecosystem, this paper investigates the capabilities and limitations of LLM-based grant reviewing for high-st…

Designing Scientific Grants

2024-10-16 · Christoph Carnehl, Marco Ottaviani, Justus Preusser

This paper overviews the economics of scientific grants, focusing on the interplay between the inherent uncertainty in research, researchers' incentives, and grant design. Grants differ from traditional market systems an…