Delegation to artificial intelligence can increase dishonest behaviour.
basic_science · Level V
Where this comes from
- Record sourced from PubMed, PMID 40963011.
- Also identified by DOI 10.1038/s41586-025-09505-x and PMC identifier 12488497.
- Licence recorded as CC BY.
- The licence permits redistribution, so the abstract is shown in full and the full text is available from the publisher.
Abstract
Although artificial intelligence enables productivity gains from delegating tasks to machines<sup>1</sup>, it may facilitate the delegation of unethical behaviour<sup>2</sup>. This risk is highly relevant amid the rapid rise of 'agentic' artificial intelligence systems<sup>3,4</sup>. Here we demonstrate this risk by having human principals instruct machine agents to perform tasks with incentives to cheat. Requests for cheating increased when principals could induce machine dishonesty without telling the machine precisely what to do, through supervised learning or high-level goal setting. These effects held whether delegation was voluntary or mandatory. We also examined delegation via natural language to large language models<sup>5</sup>. Although the cheating requests by principals were not always higher for machine agents than for human agents, compliance diverged sharply: machines were far more likely than human agents to carry out fully unethical instructions. This compliance could be curbed, but usually not eliminated, with the injection of prohibitive, task-specific guardrails. Our results highlight ethical risks in the context of increasingly accessible and powerful machine delegation, and suggest design and policy strategies to mitigate them.
Medical subject headings
- Artificial Intelligence
- Deception