Skip to content

AI Value Alignment for Evolving Social Norms

Jul 2026 · arXiv.org · Vol abs/2607.18506 · 1 citation · 77 references
Computer Science

TL;DR

This work introduces a flexible and extensible mathematical modelling framework, rooted in social physics, aimed at answering macro-level questions regarding the evolving social norms in human populations under the assumption of frequent AI use, and advocates for the wider adoption of these kinds of social physics models as an epistemic bridge.

Abstract

AI alignment is essential for the safe deployment of advanced AI systems. Given that values and preferences change over time, culture, social roles, and context, we need to develop a better understanding of the possible long-term consequences of AI alignment, in particular considering the likely ubiquitous future use of personalized AI assistants. We introduce a flexible and extensible mathematical modelling framework, rooted in social physics, aimed at answering macro-level questions regarding the evolving social norms in human populations under the assumption of frequent AI use. Our analysis is part-analytical, and part-simulation, enabling us to characterize the long-term dynamical consequences under a diverse set of starting assumptions. We highlight the risk of value lock-in, and normative mode collapse, prominently featured in non-adaptive alignment formulations. Beyond alignment, we advocate for the wider adoption of these kinds of social physics models as an epistemic bridge: enabling rapid, rigorous, and quantitatively-grounded hypothesis testing for sociotechnical foresight in general AI futures, and acting as a tractable precursor to more computationally expensive large-scale agentic evaluations.

View source

Similar papers

Preprint Aug 2026

Toward a Theory of Value in AI Alignment

Can AI systems be aligned to human values? The popularization of large language models (LLMs) and multi-modal foundation models has seen a rise in harms spanning from toxic speech and hallucinations to AI agents executing unauthorized actions. Within the field of AI safety, these harmful instances are often framed as t...

Andrew Smart, Shazeda Ahmed, Jackie Kay et al. · 0 citations
Review Open access Aug 2026

A Review of Human-AI Complementarities Across Multiple Dimensions of Organisational Complexity

A five-dimensional diagnostic framework that maps the challenges of human-AI collaboration across Integration, Representation, Scale, Temporality, and Adequacy gaps and shows that augmentation remains the dominant and most viable mode of use in complex environments.

Ganesh Sankaran, Marco A. Palomino, G. Siestrup · 0 citations
Open access 2026

AI and the Transformation of Epistemic Regimes: the Problem of Hybrid Dependency

The article examines how the pursuit of technological sovereignty in the field of artificial intelligence, instead of bringing true autonomy, generates a wide range of hybrid dependencies. These dependencies arise at the levels of computing infrastructure, data governance, AI system evaluation methodologies, and deploy...

Alexey Frolov · 0 citations
#artificial intelligence Review Aug 2026

AI Alignment through a Game-theoretic Lens: A Survey

This perspective clarifies where current alignment methods genuinely benefit from game-theoretic analysis, where the framework is looser, and what challenges remain in building robust, adaptive, and verifiable AI systems.

Yaxin Cai, Zhong-Rui Zhao, Zhigang Lu et al. · 0 citations
#artificial intelligence Preprint Sep 2026

Autonomy, Social Norms, and Alignment: Towards a Developmental Framework for Autonomous Artificial Agents

The regulatory sandboxes can be viewed as pedagogical environments for AI: dynamic spaces where alignment develops as a formative process, progressively shaping autonomous behaviors through interaction and cooperation in scenarios of increasing complexity.

Marica Notte, Ludovica Marinucci, V. Santucci · 0 citations

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.