paper-with-me

홈 › Papers

White Paper - Creating a Repository of Objectionable Online Content: Addressing Undesirable Biases and Ethical Considerations

2021-02-23 · Thamar Solorio, Mahsa Shafaei, Christos Smailis, Isabelle Augenstein, Margaret Mitchell, Ingrid Stapf, Ioannis Kakadiaris

This white paper summarizes the authors' structured brainstorming regarding ethical considerations for creating an extensive repository of online content labeled with tags that describe potentially questionable content for young viewers. The workshop focused on four topics: 1) identifying risks for unintended biases in the data and labels, 2) how to reduce risks for unintended biases; 3) identifying ethical considerations of the annotation task, and 4) reducing the risks for the annotators.

📄 PDF Abstract BibTeX

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

White Paper: Challenges and Considerations for the Creation of a Large Labelled Repository of Online Videos with Questionable Content

2021-01-25 · Thamar Solorio, Mahsa Shafaei, Christos Smailis, Mona Diab 외

This white paper presents a summary of the discussions regarding critical considerations to develop an extensive repository of online videos annotated with labels indicating questionable content. The main discussion poin…

Universal and Transferable Adversarial Attacks on Aligned Language Models

2023-07-27 · Andy Zou, Zifan Wang, Nicholas Carlini, Milad Nasr 외

Because "out-of-the-box" large language models are capable of generating a great deal of objectionable content, recent work has focused on aligning these models in an attempt to prevent undesirable generation. While ther…

Adversarial AttackIngenuity

Answering Yes-No Questions by Penalty Scoring in History Subjects of University Entrance Examinations

2016-12-01 · WS 2016 12 · Yoshinobu Kano

Answering yes{--}no questions is more difficult than simply retrieving ranked search results. To answer yes{--}no questions, especially when the correct answer is no, one must find an objectionable keyword that makes the…

Question Answering

Online Personalizing White-box LLMs Generation with Neural Bandits

2024-04-24 · Zekai Chen, Weeden Daniel, Po-Yu Chen, Francois Buet-Golfouse

The advent of personalized content generation by LLMs presents a novel challenge: how to efficiently adapt text to meet individual preferences without the unsustainable demand of creating a unique model for each user. Th…

Headline GenerationText Generation

Efficient Neural Network based Classification and Outlier Detection for Image Moderation using Compressed Sensing and Group Testing

2023-05-12 · Sabyasachi Ghosh, Sanyam Saxena, Ajit Rajwade

Popular social media platforms employ neural network based image moderation engines to classify images uploaded on them as having potentially objectionable content. Such moderation engines must answer a large number of q…

compressed sensingEfficient Neural NetworkOutlier Detection