paper-with-me

홈 › Papers

Delineating Knowledge Boundaries for Honest Large Vision-Language Models

2026-04-29 · Junru Song, Yimeng Hu, Yijing Chen, Huining Li, Qian Li, Lizhen Cui, Yuntao Du arxiv

Large Vision-Language Models (VLMs) have achieved remarkable multimodal performance yet remain prone to factual hallucinations, particularly in long-tail or specialized domains. Moreover, current models exhibit a weak capacity to refuse queries that exceed their parametric knowledge. In this paper, we propose a systematic framework to enhance the refusal capability of VLMs when facing such unknown questions. We first curate a model-specific "Visual-Idk" (Visual-I don't know) dataset, leveraging multi-sample consistency probing to distinguish between known and unknown facts. We then align the model using supervised fine-tuning followed by preference-aware optimization (e.g., DPO, ORPO) to effectively delineate its knowledge boundaries. Results on the Visual-Idk dataset show our method improves the Truthful Rate from 57.9\% to 67.3\%. Additionally, internal probing also demonstrates that the model genuinely recognizes its boundaries instead of just memorizing refusal patterns. Our framework further generalizes to out-of-distribution medical and perceptual domains, providing a robust path toward more trustworthy and prudent visual assistants.

📄 PDF Abstract BibTeX arXiv:2604.26419

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Annotation-Efficient Universal Honesty Alignment

2025-10-20 · Shiyu Ni, Keping Bi, Jiafeng Guo, Minghao Tang 외 arxiv

Honesty alignment-the ability of large language models (LLMs) to recognize their knowledge boundaries and express calibrated confidence-is essential for trustworthy deployment. Existing methods either rely on training-fr…

Alignment for Honesty

2023-12-12 · Yuqing Yang, Ethan Chern, Xipeng Qiu, Graham Neubig 외

Recent research has made significant strides in aligning large language models (LLMs) with helpfulness and harmlessness. In this paper, we argue for the importance of alignment for \emph{honesty}, ensuring that LLMs proa…

BeHonest: Benchmarking Honesty in Large Language Models

2024-06-19 · Steffi Chern, Zhulin Hu, Yuqing Yang, Ethan Chern 외

Previous works on Large Language Models (LLMs) have mainly focused on evaluating their helpfulness or harmlessness. However, honesty, another crucial alignment criterion, has received relatively less attention. Dishonest…

BenchmarkingMisinformation

Unlocking large-scale crop field delineation in smallholder farming systems with transfer learning and weak supervision

2022-01-13 · Sherrie Wang, Francois Waldner, David B. Lobell

Crop field boundaries aid in mapping crop types, predicting yields, and delivering field-scale analytics to farmers. Recent years have seen the successful application of deep learning to delineating field boundaries in i…

Transfer Learning

COB-GS: Clear Object Boundaries in 3DGS Segmentation Based on Boundary-Adaptive Gaussian Splitting

2025-03-25 · CVPR 2025 1 · Jiaxin Zhang, Junjun Jiang, Youyu Chen, Kui Jiang 외

Accurate object segmentation is crucial for high-quality scene understanding in the 3D vision domain. However, 3D segmentation based on 3D Gaussian Splatting (3DGS) struggles with accurately delineating object boundaries…

3DGSObjectScene UnderstandingSegmentation+1