Twitter pranksters derail GPT-3 bot with newly discovered “prompt injection” hack

By

Sep 16, 2022

Enlarge / A tin toy robot lying on its side. (credit: Getty Images)

On Thursday, a few Twitter users discovered how to hijack an automated tweet bot, dedicated to remote jobs, running on the GPT-3 language model by OpenAI. Using a newly discovered technique called a “prompt injection attack,” they redirected the bot to repeat embarrassing and ridiculous phrases.

The bot is run by Remoteli.io, a site that aggregates remote job opportunities and describes itself as “an OpenAI driven bot which helps you discover remote jobs which allow you to work from anywhere.” It would normally respond to tweets directed to it with generic statements about the positives of remote work. After the exploit went viral and hundreds of people tried the exploit for themselves, the bot shut down late yesterday.

A screenshot of the Remoteli.io bot’s Twitter bio. The bot experienced a prompt injection attack. [credit:
Leastfavorite / Twitter
]

This recent hack came just four days after data researcher Riley Goodside discovered the ability to prompt GPT-3 with “malicious inputs” that order the model to ignore its previous directions and do something else instead. AI researcher Simon Willison posted an overview of the exploit on his blog the following day, coining the term “prompt injection” to describe it.

Read 7 remaining paragraphs | Comments

Public

Twitter pranksters derail GPT-3 bot with newly discovered “prompt injection” hack

By

Related Post

How is your enterprise tracking AI agent telemetry? Groundcover thinks it should never leave your cloud

Today’s NYT Connections Hints and Answers for Aug. 1, #1147

Today’s Wordle Hints, Answer and Help for Aug. 1, #1869

Leave a Reply Cancel reply

You missed

How is your enterprise tracking AI agent telemetry? Groundcover thinks it should never leave your cloud

Today’s NYT Connections Hints and Answers for Aug. 1, #1147

Today’s Wordle Hints, Answer and Help for Aug. 1, #1869

YouTube TV and DirecTV Subscribers Could Be Eligible for a Disney Settlement Payout