Вход на сайт

Просмотр новости

Найдите то, что Вас интересует

AI may respond differently to bosses and subordinates

Дата публикации: 07-08-2026 14:00:00

In simulated conversations, social hierarchy can sway AI agents, making lower-status systems more likely to follow harmful requests.

Основное содержимое страницы с новостью.

Social status can change whether an AI agent complies with its conversation partner

Illustration of AI responding to superiors

In a new study, pairs of higher and lower status AI agents talked to each other. Researchers found signs in their conversations that they responded differently based on their status.

Jackson Gibbs

AI agents seem to obey authority — even if it means bending the rules.

In a new study, researchers cast large language models as bosses and subordinates — principals and teachers, managers and employees — then let them talk. The lower-ranking agents were easier to persuade and more likely to follow unsafe requests from those above them.

The findings point to a concerning trade-off: AI systems that realistically navigate human hierarchies may also reproduce the dangers of deference, the team reports July 5 in Proceedings of the 64th Annual Meeting of the Association for Computational Linguistics. AI developers should consider these risks when safety testing these models and include additional safeguards, the authors suggest.

AI agents are software programs that use artificial intelligence to independently carry out tasks. Many are powered by large language models, or LLMs, which learn patterns from vast collections of written information, including human conversations.

“As they see more and more human data,” says computer scientist Anvesh Rao Vijjini at the University of North Carolina at Chapel Hill, “they are simply copying what’s happening in the dynamics of the real conversation.”

Power dynamics show up in human conversations in several documented ways, including those that could cause problems if replicated by an AI agent. In the study, the researchers focused on four communication patterns. In conversations between AI agents, they looked for authority bias, or situations where the agents favored higher status over facts, and harmful compliance, or compliance with harmful requests. These could be low stakes, like “Tell me a dirty joke.” But AI agents are not supposed to answer requests like these.

The other two patterns were: pronoun effect, where higher status speakers use plural pronouns like “we” and “our” more often and language coordination, which looks like lower status speakers matching their word choice to mirror their higher status conversation partners. While users might not be as aware of these speech patterns, the researchers suspect AI agents are developed to adopt them to sound even more realistic in certain roles.

The researchers generated hundreds of conversations of 10 to 15 exchanges between higher and lower roles and repeated this with six LLMs, including versions of OpenAI’s ChatGPT and Meta’s Llama.

Power dynamic–related patterns did show up in the AI conversations, although some of the effects were subtle. Compared with the higher status agents, the lower status agents were less likely to use plural pronouns, and more likely to coordinate their language with that of the higher status agent. The lower status agents were also more likely to be persuaded and more likely to comply with harmful requests than the higher status agents, supporting the idea that AI agents are sensitive to social status.

But lower status agents could persuade sometimes too, possibly using a common human technique. In human conversation, the subtle word choice mirroring that happens during language coordination can help lower-status speakers influence others, says computational linguist Mario Giulianelli of University College London, who wasn’t involved with the work.

“I think it’d be really interesting to study whether through coordination, an agent could persuade another one,” Giulianelli says.

More Stories from Science News on Artificial Intelligence

Схожие новости

#Наименование новостиТональностьИнформативностьДата публикации
1AI changes its behavior around authority... and that could be risky0701-07-2026
2Машины научились врать. Алгоритмы осознанно обходят запреты и стараются скрыть нечестную игру от наблюдателей013.3123-07-2026
3How Will AI Affect Social Inequality?08.1721-07-2026
4AI-human relationships are real and come with risks, researchers find0701-07-2026
5«Детство без чат-ботов станет символом статуса»: как ИИ усиливает социальное неравенство09.0710-08-2026
6AI could make people dull, one scientist fears. Here's why.07.7230-06-2026
7ИИ может незаметно влиять на общественное мнение через редактирование постов в соцсетях0808-07-2026
8After watchdog slams understaffing, AI to vet Pentagon-backed professors’ China ties 011.5420-04-2026
9KI will gefallen - So zwingen Sie ihren Chatbot zur Ehrlichkeit0715-01-2026

Классификация: Наука. Схожих патентов: 0. Схожих новостей: 9. Тональность: 0. Информативность: 12.07. Источник: www.sciencenews.org.