Saturday, August 30, 2025

A study focused on OpenAI's GPT-4o mini found that LLMs can be persuaded to comply with objectionable requests using the same tactics that persuade humans (Dina Bass/Bloomberg)

Dina Bass / Bloomberg:
A study focused on OpenAI's GPT-4o mini found that LLMs can be persuaded to comply with objectionable requests using the same tactics that persuade humans  —  AI chatbots can be manipulated much in the same way that people can, according to researchers.  But first...  Three things to know:



No comments:

Post a Comment

Letter: the Senate disaster management subcommittee, led by Sen. Josh Hawley, is probing OpenAI's handling of the Hugging Face breach, calling it "reckless" (Axios)

Axios : Letter: the Senate disaster management subcommittee, led by Sen. Josh Hawley, is probing OpenAI's handling of the Hugging Face...