Negative Human Rights as a Basis for Long-term AI Safety and Regulation
If autonomous AI systems are to be reliably safe in novel situations, they will need to incorporate general principles guiding them to recognize and avoid harmful behaviours. Such principles may need to be supported by a binding system of regulation, which would need the underlying principles to be widely accepted. They should also be specific enough for technical implementation. Drawing inspiration from law, this article explains how negative human rights could fulfil the role of such principles and serve as a foundation both for an international regulatory system and for building technical safety constraints for future AI systems.
Code (0)
등록된 구현이 없습니다.
Similar Papers 제목 키워드 기반
Cultural Rights and the Rights to Development in the Age of AI: Implications for Global Human Rights Governance
Cultural rights and the right to development are essential norms within the wider framework of international human rights law. However, recent technological advances in artificial intelligence (AI) and adjacent digital f…
Artificial intelligence, human rights, democracy, and the rule of law: a primer
In September 2019, the Council of Europe's Committee of Ministers adopted the terms of reference for the Ad Hoc Committee on Artificial Intelligence (CAHAI). The CAHAI is charged with examining the feasibility and potent…
GET-AID: Visual Recognition of Human Rights Abuses via Global Emotional Traits
In the era of social media and big data, the use of visual evidence to document conflict and human rights abuse has become an important element for human rights organizations and advocates. In this paper, we address the …
A mathematical definition of property rights in a Debreu economy
We present mathematical definitions for rights structures, government and non-attenuation in a generalized N-person game (Debreu abstract economy), thus providing a formal basis for property rights theory.
Do LLMs exhibit demographic parity in responses to queries about Human Rights?
This research describes a novel approach to evaluating hedging behaviour in large language models (LLMs), specifically in the context of human rights as defined in the Universal Declaration of Human Rights (UDHR). Hedgin…
Attribute