Skip to content

Autonomy, Social Norms, and Alignment: Towards a Developmental Framework for Autonomous Artificial Agents

Sep 2026 · 0 citations · 15 references
Computer Science

TL;DR

The regulatory sandboxes can be viewed as pedagogical environments for AI: dynamic spaces where alignment develops as a formative process, progressively shaping autonomous behaviors through interaction and cooperation in scenarios of increasing complexity.

Abstract

In recent years, artificial intelligence has made extraordinary progress thanks to large-scale models capable of generalization and the generation of complex outputs. However, transferring this potential into embodied agents reveals a significant limitation: the most advanced systems rely on pre-existing datasets and human feedback strategies that are powerful but insufficient in dynamic or unknown contexts. To adapt, an agent must acquire knowledge through direct interaction with its environment. One strategy to address this challenge involves introducing higher-level mechanisms, such as intrinsic motivations, which leverage curiosity and competence, to guide exploration and learning in complex environments. While this flexibility expands autonomy, it complicates the task of ensuring agents remain aligned with human goals. Alignment, already a challenge for artificial systems in general, becomes even more complex in unstructured and dynamic contexts where predefined rules prove insufficient. To be effective and adaptable, norms must be rooted in experience through an epistemological process that starting from simple, situated principles allows for the gradual construction of more complex rules through experience, autonomous learning, and cooperation with other moral agents. Similarly to children learning social norms by exploring their environment and participating in collective practices, artificial agents must also be educated toward alignment. Following Dennett, the status of a moral agent is not innate but is attributed gradually based on the ability to responsibly manage increasing degrees of freedom. From this perspective, the regulatory sandboxes can be viewed as pedagogical environments for AI: dynamic spaces where alignment develops as a formative process, progressively shaping autonomous behaviors through interaction and cooperation in scenarios of increasing complexity.

View source

Similar papers

Review Open access Aug 2026

Towards safe and trustworthy agentic AI: foundations, taxonomy, technologies, applications, and future directions

This work synthesizes perspectives from philosophy, cognitive science, and AI to define agency, outline its key properties, and situate it in relation to existing paradigms such as reinforcement learning, symbolic reasoning, Belief–Desire–Intention (BDI) architectures, and embodied cognition.

Vijayrajsinh Gohil, Siddhant Bikram Shah, Kritesh Rauniyar et al. · 0 citations
Review Open access Aug 2026

From Language Models to Agentic AI: A Survey of Autonomous, Action-Enabled, and Collaborative LLM Agents

A unified, taxonomy-driven, and deployment-oriented survey of agentic AI systems, synthesizing recent advances through a modular reference architecture and a four-dimensional taxonomy that characterizes agents along the axes of autonomy, tool use, collaboration, and safety–governance is presented.

Sparsh Bajoria, Shreyanshu Ranjan, Adhitya M et al. · 0 citations
Review Open access Sep 2026

Toward Collaborative AI: A Framework for the Transition from Autonomous Agents to Adaptive Human-AI Partners

This work defines collaborative AI as a class of systems that combine generative exploration with autonomous action and calibrate between them based on context, uncertainty, and task demands, and identifies four required capabilities: metacognition, contextual mode-switching, uncertainty-aware action, and adaptive huma...

Nalan Karunanayake, Savindu Nanayakkara, Kasun Gayashan Hettihewa et al. · 0 citations
#natural language process... Preprint Aug 2026

Agents in the Large: Perception-Centered Architecture for Persistent Agents

Pera describes a persistent agent organized around perception and control components that continually perceive service-relevant signals from episodic task executions, internal context, and changes in the surrounding environment, and use these signals to construct lifecycle tasks.

Shi-Han Dou, Haoxiang Jia, Shichun Liu et al. · 1 citation
#artificial intelligence Review Sep 2026

From Language Models to World-Acting Systems: Progress and Limits of Agentic AI across Digital, Social, Virtual, and Physical Environments

Large language models become consequential agents when surrounding systems let outputs change external state. Models now call tools, operate interfaces, delegate work, retain state, inhabit generated worlds, and control robots or laboratory equipment. Such advances are often narrated as one march toward autonomy, confl...

Lin-Sen Zhu, Meng-Qing Cai · 1 citation
#artificial intelligence Preprint Sep 2026

Robots That Take Initiative: A Framework for Building and Evaluating Proactive Robots

This work introduces a unified formalism for proactive robot assistance, organize it into three levels, and provides a framework to address the highest level of unprompted proactive assistance, and presents a method, GAP, that instantiates the framework, learning from passive observation to anticipate user goals and ac...

Maithili Patel, Sonia Chernova · 0 citations

Related blog posts

MIT News · Artificial Intelligence Sep 29, 2026

Who we become when we talk to machines

Professor Sherry Turkle’s new book, “Artificial Intimacy,” offers a withering critique of chatbots and the antisocial dynamics she believes they encourage.

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.