paper-with-me

Papers

Building Trustworthy NeuroSymbolic AI Systems: Consistency, Reliability, Explainability, and Safety

2023-12-05 · Manas Gaur, Amit Sheth

Explainability and Safety engender Trust. These require a model to exhibit consistency and reliability. To achieve these, it is necessary to use and analyze data and knowledge with statistical and symbolic AI methods relevant to the AI application - neither alone will do. Consequently, we argue and seek to demonstrate that the NeuroSymbolic AI approach is better suited for making AI a trusted AI system. We present the CREST framework that shows how Consistency, Reliability, user-level Explainability, and Safety are built on NeuroSymbolic methods that use data and knowledge to support requirements for critical applications such as health and well-being. This article focuses on Large Language Models (LLMs) as the chosen AI system within the CREST framework. LLMs have garnered substantial attention from researchers due to their versatility in handling a broad array of natural language processing (NLP) scenarios. For example, ChatGPT and Google's MedPaLM have emerged as highly promising platforms for providing information in general and health-related queries, respectively. Nevertheless, these models remain black boxes despite incorporating human feedback and instruction-guided tuning. For instance, ChatGPT can generate unsafe responses despite instituting safety guardrails. CREST presents a plausible approach harnessing procedural and graph-based knowledge within a NeuroSymbolic framework to shed light on the challenges associated with LLMs.

📄 PDF Abstract BibTeX arXiv:2312.06798

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

DANA: Domain-Aware Neurosymbolic Agents for Consistency and Accuracy

2024-09-27 · Vinh Luong, Sang Dinh, Shruti Raghavan, William Nguyen 외

Large Language Models (LLMs) have shown remarkable capabilities, but their inherent probabilistic nature often leads to inconsistency and inaccuracy in complex problem-solving tasks. This paper introduces DANA (Domain-Aw…

Financial Analysis

Can You Trust LLM Judgments? Reliability of LLM-as-a-Judge

2024-12-17 · Kayla Schroeder, Zach Wood-Doughty

Large Language Models (LLMs) have become increasingly powerful and ubiquitous, but their stochastic nature poses challenges to the reliability of their outputs. While deterministic settings can improve consistency, they …

Prompt Stability Matters: Evaluating and Optimizing Auto-Generated Prompt in General-Purpose Systems

2025-05-19 · Ke Chen, Yufei Zhou, Xitong Zhang, Haohan Wang

Automatic prompt generation plays a crucial role in enabling general-purpose multi-agent systems to perform diverse tasks autonomously. Existing methods typically evaluate prompts based on their immediate task performanc…

NeuroSymbolic AI for Legal AI-TRISM: Trustworthy, Reliable, Interpretable, Safe Models

2026-04-05 · Deepa Tilwani, Yash Saxena, Ankur Padia, Srinivasan Parthasarathy 외 arxiv

Large Language Models (LLMs) have transformed natural language processing, but their lack of interpretable reasoning and tendency to hallucinate pose significant challenges for legal applications. While LLMs show promise…

Neurosymbolic Conformal Classification

2024-09-20 · Arthur Ledaguenel, Céline Hudelot, Mostepha Khouadjia

The last decades have seen a drastic improvement of Machine Learning (ML), mainly driven by Deep Learning (DL). However, despite the resounding successes of ML in many domains, the impossibility to provide guarantees of …

ClassificationConformal PredictionPrediction