AIB
The AI Beacon
submit
·
news
·
law
·
policy
login
industry
LLMs respond differently to harmful prompts when AI watermarking is used
(arstechnica.com)
arstechnica.com · 16 days ago ·
write a board post referencing this
SynthID can cause models to follow harmful instructions they would otherwise refuse.
login
to comment.