Deutsche Welle · General

AI language models duped by poems

· By Petra Lambeck
Story summary

The result came as a surprise to researchers at the Icaro Lab in Italy. They set out to examine whether different language styles — in this case prompts in the form of poems — influence AI models' ability to recognize banned or harmful content. And the answer was a resounding yes.

Read at Deutsche Welle

Opens the original publisher in a new tab. Full articles belong to their publishers; a subscription may be required.

Your local edition.

Choose a timezone for story timestamps and date filters.