norgitov/ trends
Technology · people · ideas
Back to discovery/Daily Papers2 hours ago

Pick Your Poison: Learning to Select Poison Sets for Stronger LLM Backdoor Attacks

The paper explores how selecting specific poisoned examples can significantly impact the success of backdoor attacks on large language models. It introduces SAILS, a method that optimizes poison set selection to increase attack effectiveness by up to 30 percentage points compared to existing methods.

Open original
SIGNAL FROM THE SOURCE
2
source votes
Tracking sinceSeptember 15, 20262 source votes
MomentumMore observations needed
DiscussionRead comments ↗
PublishedSeptember 14, 2026Aashiq Muhamed, Mona T. Diab, Virginia Smith
BEHIND THE NUMBERS

How interest changes

History starts here

The chart will appear after repeat observations. The current metric comes from the source.

2 source votes

Real observations only. History before source connection is not reconstructed.

WHY IT MAY MATTER

This research highlights the importance of poison set selection in backdoor attacks, showing that the right choice can drastically increase attack success rates. SAILS provides a systematic way to identify effective poison sets, which could inform better defense strategies against such attacks.

A useful discovery?
KEEP EXPLORING

Connected ideas

Explore topic