← Back to feed
AI SecurityEmerging1 sourceFeb 8, 2024 · 10:01via Embrace The Red (AI agent security)

Hidden Prompt Injections with Anthropic Claude

Brief

A few weeks ago while waiting at the airport lounge I was wondering how other Chatbots, besides ChatGPT, handle hidden Unicode Tags code points.

A quick reminder: Unicode Tags code points are invisible in UI elements , but ChatGPT was able to interpret them and follow hidden instructions. Riley Goodside discovered it .

What about Anthropic Claude?

While waiting for a flight I figured to look at Anthropic Claude. Turns out it has the same issue as ChatGPT had. I reported it behind the scenes, but got the following final reply and the ticket was closed.

Read more on Embrace The Red (AI agent security)