The Discord Leak is Bigger Than Expected - Threat Wire
Quick Overview
The apparent size of the Discord data leak, which compromised a third-party customer service provider, was significantly larger than initially reported, ultimately exposing government IDs and impacting around 70,000 users, while concurrently, a new study revealed that large language model (LLM) data poisoning attacks are surprisingly easy to execute, requiring only a small, fixed percentage of malicious training data to backdoor models.
Key Points: The Discord security incident, involving a third-party customer service provider, compromised data affecting approximately 70,000 users, including government IDs. Discord published an update on October 3rd, 2025, confirming the breach originated from the compromised third-party vendor. The attacker demanded a financial ransom, but Discord revoked the vendor's access and engaged law enforcement and a computer forensics firm. A new paper by researchers at UK AI Security Institute and the Alan Turing Institute demonstrated that LLM poisoning is easier than expected. The study found that poisoning attacks require a near-constant number of malicious documents regardless of the LLM's parameter size (tested up to 138 billion parameters). Only 250 malicious documents (roughly 420k tokens, representing 0.00016% of total training tokens) were sufficient to successfully backdoor models. BreachForums, a data leak extortion site used by the ShinyHunters group, was seized by the FBI and French law enforcement on October 7th, 2025.
Context: This episode of Threat Wire covers two significant cybersecurity events: the escalation of a data breach at Discord involving a third-party vendor, and alarming new research detailing the ease of data poisoning attacks against large language models (LLMs). The Discord situation highlights risks associated with third-party access, while the LLM research underscores a growing vulnerability in AI model training data.
Detailed Analysis
The video first addresses the Discord data leak, revealing that the initial reports understated the impact. The breach occurred via a third-party customer service provider, leading to the exposure of approximately 70,000 users' data, including government IDs. Discord took immediate steps upon discovery on September 20th, 2025, by revoking access, launching an internal investigation, hiring a forensics firm, and involving law enforcement. The attackers had demanded a ransom, but their access was revoked. Subsequently, the host discusses a new paper detailing how easy it is to poison Large Language Models (LLMs). The research, conducted by UK AI Security Institute and the Alan Turing Institute, found that poisoning attacks require a nearly constant number of malicious documents, irrespective of the model size, even for models up to 138 billion parameters. Just 250 malicious documents, equating to only 0.00016% of total training tokens, were enough to successfully backdoor the models, causing them to generate incoherent responses when prompted with a specific trigger word. Finally, the video reports that BreachForums, a notorious hacking forum used by the ShinyHunters group for data leak extortion following the Salesforce theft, was seized on October 7th, 2025, through a joint operation between the FBI and French law enforcement, although the site briefly came back online before being fully controlled by authorities.