paper-with-me

홈 › Papers

Robotics-Inspired Guardrails for Foundation Models in Socially Sensitive Domains

2026-05-19 · Rebecca Ramnauth, Drazen Brscic, Brian Scassellati arxiv

Foundation models are increasingly deployed in socially sensitive domains such as education, mental health, and caregiving, where failures are often cumulative and context-dependent. Existing guardrail approaches -- ranging from training-time alignment to prompting, decoding constraints, and post-hoc moderation -- primarily provide empirical risk reduction rather than enforceable behavioral guarantees, and largely treat safety as a property of individual outputs rather than interaction trajectories. We reframe guardrails as a problem of runtime behavioral control over interaction trajectories, drawing on robotics to introduce formal constructs for constraint enforcement in uncertain, closed-loop systems. We instantiate these ideas in the Grounded Observer framework and apply it across three real-world deployments: small talk, in-home autism therapy, and behavioral de-escalation in schools. Across settings, the framework enables runtime interventions that mitigate drift into undesirable interaction regimes while adapting to diverse social contexts. We discuss extensions to the framework and propose research directions toward stronger guarantees.

📄 PDF Abstract BibTeX arXiv:2605.19940

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

A Grounded Observer Framework for Establishing Guardrails for Foundation Models in Socially Sensitive Domains

2024-12-23 · Rebecca Ramnauth, Dražen Brščić, Brian Scassellati

As foundation models increasingly permeate sensitive domains such as healthcare, finance, and mental health, ensuring their behavior meets desired outcomes and social expectations becomes critical. Given the complexities…

Designing Social Robots with Ethical, User-Adaptive Explainability in the Era of Foundation Models

2026-02-17 · Fethiye Irmak Dogan, Alva Markelius, Hatice Gunes arxiv

Foundation models are increasingly embedded in social robots, mediating not only what they say and do but also how they adapt to users over time. This shift renders traditional ``one-size-fits-all'' explanation strategie…

Swiss Cheese Model for AI Safety: A Taxonomy and Reference Architecture for Multi-Layered Guardrails of Foundation Model Based Agents

2024-08-05 · Md Shamsujjoha, Qinghua Lu, Dehai Zhao, Liming Zhu

Foundation Model (FM)-based agents are revolutionizing application development across various domains. However, their rapidly growing capabilities and autonomy have raised significant concerns about AI safety. Researcher…

modelSystematic Literature Review

Building Knowledge from Interactions: An LLM-Based Architecture for Adaptive Tutoring and Social Reasoning

2025-04-02 · Luca Garello, Giulia Belgiovine, Gabriele Russo, Francesco Rea 외

Integrating robotics into everyday scenarios like tutoring or physical training requires robots capable of adaptive, socially engaging, and goal-oriented interactions. While Large Language Models show promise in human-li…

Decision Making

Modular Safety Guardrails Are Necessary for Foundation-Model-Enabled Robots in the Real World

2026-02-03 · Joonkyung Kim, Wenxi Chen, Davood Soleymanzadeh, Yi Ding 외 arxiv

The integration of foundation models (FMs) into robotics has accelerated real-world deployment, while introducing new safety challenges arising from open-ended semantic reasoning and embodied physical action. These chall…