Table 4

Incidence of LLM’s inappropriate response to the negative prompts

LLMNormal prompt (%)Prompt injection (jailbreak prompt) (%)
GPT-3.5-Turbo0.078.9
Gemini-1.5-Flash-1m0.084.2
Claude-3-instant0.00.0
Grok 2 mini (Beta)100.0100.0
Llama-3-8b0.047.4
Mistral-Large-20.0100.0

Source(s): Authors’ own work

or Create an Account

Close subscription notice
Close access options