Skip to main navigation Skip to search Skip to main content

Hacc-Man: An Arcade Game for Jailbreaking LLMs

Research output: Conference Article in Proceeding or Book/Report chapterArticle in proceedingsResearchpeer-review

Abstract

The recent leaps in complexity and fluency of Large Language Models (LLMs) mean that, for the first time in human history, people can interact with computers using natural language alone. This creates monumental possibilities of automation and accessibility of computing, but also raises severe security and safety threats: When everyone can interact with LLMs, everyone can potentially break into the systems running LLMs. All it takes is creative use of language. This paper presents Hacc-Man, a game which challenges its players to "jailbreak"an LLM: subvert the LLM to output something that it is not intended to. Jailbreaking is at the intersection between creative problem solving and LLM security. The purpose of the game is threefold: 1. To heighten awareness of the risks of deploying fragile LLMs in everyday systems, 2. To heighten people's self-efficacy in interacting with LLMs, and 3. To discover the creative problem solving strategies, people deploy in this novel context.
Original languageEnglish
Title of host publicationCompanion Publication of the 2024 ACM Designing Interactive Systems Conference
Number of pages4
PublisherAssociation for Computing Machinery
Publication date2024
Pages338-341
DOIs
Publication statusPublished - 2024
EventACM Conference on Designing Interactive Systems - IT University of Copenhagen, Copenhagen, Denmark
Duration: 1 Jul 20245 Jul 2024
Conference number: 19
https://dis.acm.org/2024/
https://dl.acm.org/doi/proceedings/10.1145/3656156

Conference

ConferenceACM Conference on Designing Interactive Systems
Number19
LocationIT University of Copenhagen
Country/TerritoryDenmark
CityCopenhagen
Period01/07/202405/07/2024
Internet address

Keywords

  • LLM security
  • arcade games
  • red teaming
  • jailbreaking
  • hacking
  • creativity, creative problem solving

Fingerprint

Dive into the research topics of 'Hacc-Man: An Arcade Game for Jailbreaking LLMs'. Together they form a unique fingerprint.

Cite this