Skip to main navigation Skip to search Skip to main content

Backdoor threats in large language models—a survey

  • Xi'an Jiaotong University

Research output: Contribution to journalReview articlepeer-review

Abstract

Large language models (LLMs), with their advanced language comprehension and text generation capabilities, have demonstrated remarkable performance across diverse application scenarios involving code processing, search engines, and translation, among others. However, these models have become increasingly vulnerable to security threats, particularly to backdoor attacks. Therefore, a timely and comprehensive review of the existing backdoor threats is urgently required. In this paper, we present a systematic and timely review of the research on backdoor attacks on LLMs, categorising existing attack and defence methods according to the LLM. Additionally, we draw comparisons with backdoor attacks in traditional deep learning to provide a more intuitive understanding of backdoor threats in LLMs. Through this effective analysis and an evaluation of the reviewed studies, we identify the current research challenges and propose potential future research directions to address these issues.

Original languageEnglish
Article number191101
JournalScience China Information Sciences
Volume68
Issue number9
DOIs
StatePublished - Sep 2025

Keywords

  • LLM life cycle
  • artificial intelligence security
  • attack and defence category
  • backdoor threats
  • large language models

Fingerprint

Dive into the research topics of 'Backdoor threats in large language models—a survey'. Together they form a unique fingerprint.

Cite this