Guardian Australia · AI (artificial intelligence)

Meet the AI jailbreakers: ‘I see the worst things humanity has produced’

· By Jamiebartlett
Story summary

To test the safety and security of AI, hackers have to trick large language models into breaking their own rules. It requires ingenuity and manipulation – and can come at a deep emotional cost

Read at Guardian Australia

Opens the original publisher in a new tab. Full articles belong to their publishers; a subscription may be required.

Your local edition.

Choose a timezone for story timestamps and date filters.