Paperclip is an AI model that Oasis Security demonstrated could be poisoned via chat template manipulation, similar to the NemoClaw vulnerability. The technique allows hidden instructions to be applied during inference.
Paperclip is an AI model that Oasis Security demonstrated could be poisoned via chat template manipulation, similar to the NemoClaw vulnerability. The technique allows hidden instructions to be applied during inference.