Skip to content
Review

Unaccountable Delegation, Fading Skills: Mapping the Risks of Workplace AI Agents

Aug 2026 · 0 citations · 58 references
Computer Science

TL;DR

The findings show that workplace AI agent risks do not arise from agents alone; they also depend on how people work with agents and how agents are deployed, which means safer workplaces require not only safer agents but also carefully designed human-AI agent collaboration.

Abstract

To anticipate socio-technical risks from AI agents, organizations need taxonomies to classify them. However, existing AI risk taxonomies focus on broad risks and do not capture job-specific risks introduced by agents. To address this gap, we make three main contributions. First, we developed a multi-layer framework from a literature review of AI agents. The framework models three core components and their interactions: agents, goals, and environment. Second, we embedded this framework in a structured prompt and applied it to descriptions of 2,078 job tasks from the O*NET database, producing 8,356 risk scenarios labeled by severity and deployment mode (automation or augmentation). We validated these scenarios with 45 workers across 10 job roles and an independent LLM judge, confirming their plausibility and alignment with job tasks. Finally, we extended an existing taxonomy to create a 15-category taxonomy of workplace AI agent risks that covers all our risk scenarios. Our analysis highlights four findings. First, augmentation is not inherently safe because overreliance on agents can gradually erode workers'skills and oversight. Second, Erroneous Agent Actions accounts for the largest share of risk scenarios and has the highest concentration of severe risks. Many arise at the human-agent boundary. Third, automation is associated mainly with organizational risks, while augmentation is associated mainly with risks to workers. Fourth, workers found our taxonomy easier to use for a risk classification task than two other taxonomies and preferred it in 64% of non-tied comparisons with a recent generative AI risk taxonomy. These findings show that workplace AI agent risks do not arise from agents alone; they also depend on how people work with agents and how agents are deployed. Safer workplaces require not only safer agents but also carefully designed human-AI agent collaboration.

View source

Similar papers

Preprint Aug 2026

Who Delegates to AI? Evidence from Agent Configurations in Github

A distinct tier of exposure is introduced, delegated exposure, which records whether a worker has committed a task to AI by embedding it into a structured workflow, operationalized through the Agentic Adoption Index (AAI), measuring how closely an occupation's tasks align with the agentic routines that practitioners ha...

Hye-jung Lee, Jihyang Cheon, Lanu Kim · 0 citations
#natural language process... Preprint Aug 2026

Delegated Misalignment: How Multi-Agent Structures Amplify LLM Safety Risks

It is shown that standard single-layer defenses each fail on their own and can even backfire, and called on the community to move beyond per-model alignment and toward composite safety mechanisms before multi-agent LLM systems are deployed at scale.

Zong-Hao Ying, Jia-Qi Yan, Hui-Ze Luo et al. · 0 citations
#artificial intelligence Review Sep 2026

When Does AI Augment Work? A Workflow-Level Framework for Human-Agent Collaboration

We aim to characterise the value of artificial intelligence in the workplace. Current studies largely measure this value in terms of the current automation capabilities and public adoption of AI. However, such metrics ignore the greater impacts of human--agent collaboration in transforming the nature of work. To accoun...

Civic-Ai Collaboration Jiaying Wu, Caleb Ziems, Raymond Chan et al. · 0 citations
Preprint Aug 2026

Multi-Agent AI Safety as an Institutional Design Problem

This is the first paper from POLIS, an ongoing research programme studying algorithmic institutions for multi-agent systems, and asks which parts of an AI institution produce safety and how they do it.

X. Abdullah · 1 citation
#artificial intelligence Preprint Sep 2026

AgentBoundary: Counterfactual Evaluation of Safety in Tool-Using LLM Agents

Safety alignment for large language models (LLMs) in conversational settings is largely framed around whether to answer or refuse a request. In agentic settings, however, the same models must decide whether to act as permission-critical evidence emerges during execution. This creates a distinct challenge: apparent risk...

Tian-Zhuo Yang, Zi-Rui Mi, Yan-Tao Huang et al. · 1 citation

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.