paper-with-me

홈 › Papers

Foundation models on the bridge: Semantic hazard detection and safety maneuvers for maritime autonomy with vision-language models

2025-12-30 · Kim Alexander Christensen, Andreas Gudahl Tufte, Alexey Gusev, Rohan Sinha, Milan Ganai, Ole Andreas Alsos, Marco Pavone, Martin Steinert arxiv

The draft IMO MASS Code requires autonomous and remotely supervised maritime vessels to detect departures from their operational design domain, enter a predefined fallback that notifies the operator, permit immediate human override, and avoid changing the voyage plan without approval. Meeting these obligations in the alert-to-takeover gap calls for a short-horizon, human-overridable fallback maneuver. Classical maritime autonomy stacks struggle when the correct action depends on meaning (e.g., diver-down flag means people in the water, fire close by means hazard). We argue (i) that vision-language models (VLMs) provide semantic awareness for such out-of-distribution situations, and (ii) that a fast-slow anomaly pipeline with a short-horizon, human-overridable fallback maneuver makes this practical in the handover window. We introduce Semantic Lookout, a camera-only, candidate-constrained VLM fallback maneuver selector that selects one cautious action (or station-keeping) from water-valid, world-anchored trajectories under continuous human authority. On 40 harbor scenes we measure per-call scene understanding and latency, alignment with human consensus (model majority-of-three voting), short-horizon risk-relief on fire hazard scenes, and an on-water alert->fallback maneuver->operator handover. Sub-10 s models retain most of the awareness of slower state-of-the-art models. The fallback maneuver selector outperforms geometry-only baselines and increases standoff distance on fire scenes. A field run verifies end-to-end operation. These results support VLMs as semantic fallback maneuver selectors compatible with the draft IMO MASS Code, within practical latency budgets, and motivate future work on domain-adapted, hybrid autonomy that pairs foundation-model semantics with multi-sensor bird's-eye-view perception and short-horizon replanning. Website: kimachristensen.github.io/bridge_policy

📄 PDF Abstract BibTeX arXiv:2512.24470

Code (0)

등록된 구현이 없습니다.

Tasks

Scene Understanding

Similar Papers 제목 키워드 기반

Towards a Multi-Agent Vision-Language System for Zero-Shot Novel Hazardous Object Detection for Autonomous Driving Safety

2025-04-18 · Shashank Shriram, Srinivasa Perisetla, Aryan Keskar, Harsha Krishnaswamy 외

Detecting anomalous hazards in visual data, particularly in video streams, is a critical challenge in autonomous driving. Existing models often struggle with unpredictable, out-of-label hazards due to their reliance on p…

Anomaly DetectionAutonomous DrivingDenoisingLanguage Modeling+7

Learning to Navigate Under Imperfect Perception: Conformalised Segmentation for Safe Reinforcement Learning

2025-10-21 · Daniel Bethell, Simos Gerasimou, Radu Calinescu, Calum Imrie arxiv

Reliable navigation in safety-critical environments requires both accurate hazard perception and principled uncertainty handling to strengthen downstream safety handling. Despite the effectiveness of existing approaches,…

Reinforcement LearningSemantic Segmentation

From 3D Perception to Safety Reasoning: A Graph-Based Framework for Real-Time Underground Mine Monitoring

2026-06-02 · Pasindu Ranasinghe, Simit Raval, Dibyayan Patra, Bikram Banerjee 외 arxiv

Underground coal mining requires personnel and heavy equipment to operate within shared, confined, and poorly illuminated spaces where hazards such as equipment proximity violations, structural instabilities, and occlude…

Scene UnderstandingAnomaly DetectionPoint Clouds

A Hazard-Informed Data Pipeline for Robotics Physical Safety

2026-03-06 · Alexei Odinokov, Rostislav Yavorskiy arxiv

This report presents a structured Robotics Physical Safety Framework based on explicit asset declaration, systematic vulnerability enumeration, and hazard-driven synthetic data generation. The approach bridges classical …

Synthetic Data Generation

Ensuring Safe Physical AI in Urban Mobility via Hazard-Informed Synthesized Envelopes

2026-08-14 · Alexei Odinokov, Rostislav Yavorskiy arxiv

As heterogeneous robotic systems deploy across diverse urban zones, maintaining safety amid complex human-robot interactions remains a critical challenge. We present a unified framework that bridges systematic hazard ana…