paper-with-me

홈 › Papers

Killing Two Flies with One Stone: An Attempt to Break LLMs Using English->Icelandic Idioms and Proper Names

2024-10-04 · Bjarki Ármannsson, Hinrik Hafsteinsson, Atli Jasonarson, Steinþór Steingrímsson

This paper presents the submission of the \'Arni Magn\'usson Institute's team to the WMT24 test suite subtask, focusing on idiomatic expressions and proper names for the English->Icelandic translation direction. Intuitively and empirically, idioms and proper names are known to be a significant challenge for modern translation models. We create two different test suites. The first evaluates the competency of MT systems in translating common English idiomatic expressions, as well as testing whether systems can distinguish between those expressions and the same phrases when used in a literal context. The second test suite consists of place names that should be translated into their Icelandic exonyms (and correctly inflected) and pairs of Icelandic names that share a surface form between the male and female variants, so that incorrect translations impact meaning as well as readability. The scores reported are relatively low, especially for idiomatic expressions and place names, and indicate considerable room for improvement.

📄 PDF Abstract BibTeX arXiv:2410.03394

Code (1)

stofnun-arna-magnussonar/idioms_names_test_suite 공식 구현

Tasks

Translation

Similar Papers 제목 키워드 기반

LLM-Virus: Evolutionary Jailbreak Attack on Large Language Models

2024-12-28 · Miao Yu, Junfeng Fang, Yingjie Zhou, Xing Fan 외

While safety-aligned large language models (LLMs) are increasingly used as the cornerstone for powerful systems such as multi-agent frameworks to solve complex real-world problems, they still suffer from potential advers…

Heuristic SearchTransfer Learning

Extracting books from production language models

2026-01-06 · Ahmed Ahmed, A. Feder Cooper, Sanmi Koyejo, Percy Liang arxiv

Many unresolved legal questions over LLMs and copyright center on memorization: whether specific training data have been encoded in the model's weights during training, and whether those memorized data can be extracted i…

Gradient Cuff: Detecting Jailbreak Attacks on Large Language Models by Exploring Refusal Loss Landscapes

2024-03-01 · Xiaomeng Hu, Pin-Yu Chen, Tsung-Yi Ho

Large Language Models (LLMs) are becoming a prominent generative AI tool, where the user enters a query and the LLM generates an answer. To reduce harm and misuse, efforts have been made to align these LLMs to human valu…

RICoTA: Red-teaming of In-the-wild Conversation with Test Attempts

2025-01-29 · Eujeong Choi, Younghun Jeong, SooMin Kim, Won Ik Cho

User interactions with conversational agents (CAs) evolve in the era of heavily guardrailed large language models (LLMs). As users push beyond programmed boundaries to explore and build relationships with these systems, …

ChatbotRed Teaming

Flies as Ship Captains? Digital Evolution Unravels Selective Pressures to Avoid Collision in Drosophila

2016-03-02 · Ali Tehrani-Saleh, Christoph Adami

Flies that walk in a covered planar arena on straight paths avoid colliding with each other, but which of the two flies stops is not random. High-throughput video observations, coupled with dedicated experiments with con…

Collision Avoidance