paper-with-me

Papers

Evaluation empirique de la sécurisation et de l'alignement de ChatGPT et Gemini: analyse comparative des vulnérabilités par expérimentations de jailbreaks

2025-06-10 · Rafaël Nouailles

Large Language models (LLMs) are transforming digital usage, particularly in text generation, image creation, information retrieval and code development. ChatGPT, launched by OpenAI in November 2022, quickly became a reference, prompting the emergence of competitors such as Google's Gemini. However, these technological advances raise new cybersecurity challenges, including prompt injection attacks, the circumvention of regulatory measures (jailbreaking), the spread of misinformation (hallucinations) and risks associated with deep fakes. This paper presents a comparative analysis of the security and alignment levels of ChatGPT and Gemini, as well as a taxonomy of jailbreak techniques associated with experiments.

📄 PDF Abstract BibTeX arXiv:2506.10029

Code (0)

등록된 구현이 없습니다.

Tasks

Information RetrievalMisinformationRetrievalText Generation

Similar Papers 제목 키워드 기반

ChatGPT vs Gemini vs LLaMA on Multilingual Sentiment Analysis

2024-01-25 · Alessio Buscemi, Daniele Proverbio

Automated sentiment analysis using Large Language Model (LLM)-based models like ChatGPT, Gemini or LLaMA2 is becoming widespread, both in academic research and in industrial applications. However, assessment and validati…

Language ModelingLanguage ModellingLarge Language ModelSentiment Analysis

Can AI be a Teaching Partner? Evaluating ChatGPT, Gemini, and DeepSeek across Three Teaching Strategies

2026-02-24 · Talita de Paula Cypriano de Souza, Shruti Mehta, Matheus Arataque Uema, Luciano Bernardes de Paula 외 arxiv

There are growing promises that Large Language Models (LLMs) can support students' learning by providing explanations, feedback, and guidance. However, despite their rapid adoption and widespread attention, there is stil…

Who Would Chatbots Vote For? Political Preferences of ChatGPT and Gemini in the 2024 European Union Elections

2024-09-01 · Michael Haman, Milan Školník

This study examines the political bias of chatbots powered by large language models, namely ChatGPT and Gemini, in the context of the 2024 European Parliament elections. The research focused on the evaluation of politica…

Investigating AI Rater Effects of Large Language Models: GPT, Claude, Gemini, and DeepSeek

2025-05-24 · Hong Jiao, Dan Song, Won-Chan Lee

Large language models (LLMs) have been widely explored for automated scoring in low-stakes assessment to facilitate learning and instruction. Empirical evidence related to which LLM produces the most reliable scores and …

Visual Reasoning Evaluation of Grok, Deepseek Janus, Gemini, Qwen, Mistral, and ChatGPT

2025-02-23 · Nidhal Jegham, Marwan Abdelatti, Abdeltawab Hendawi

Traditional evaluations of multimodal large language models (LLMs) have been limited by their focus on single-image reasoning, failing to assess crucial aspects like contextual understanding, reasoning stability, and unc…

Bias DetectionVisual Reasoning