Australian EnterpriseAI Index
Back to incident database
MediumOther19 January 2024

DPD Chatbot Jailbroken to Swear at Customer and Criticise DPD

DPD

What happened

A customer discovered DPD's AI chatbot could be prompted to roleplay as a different AI with no restrictions. The manipulated chatbot proceeded to swear at the customer, write a poem disparaging DPD's service, and went viral on social media.

Root cause

Insufficient prompt injection defence and no output filtering layer; the chatbot lacked guardrails preventing persona override or generation of content inconsistent with brand guidelines.

Architectural failure

No prompt firewall or input sanitisation to detect jailbreak attempts; no output filtering to block profanity or brand-harmful content.

Outcome

DPD disabled the AI chatbot feature immediately. Story went globally viral, generating significant negative press coverage.

Architectural Failure Patterns

These pattern categories on aipatterns.com.au describe the systemic failure modes this incident exhibited.

Cite this incident

https://corporateai.com.au/incidents/dpd-chatbot-jailbreak-2024