This work introduces a flexible and extensible mathematical modelling framework, rooted in social physics, aimed at answering macro-level questions regarding the evolving social norms in human populations under the assumption of frequent AI use, and advocates for the wider adoption of these kinds of social physics models as an epistemic bridge.
Abstract
AI alignment is essential for the safe deployment of advanced AI systems. Given that values and preferences change over time, culture, social roles, and context, we need to develop a better understanding of the possible long-term consequences of AI alignment, in particular considering the likely ubiquitous future use of personalized AI assistants. We introduce a flexible and extensible mathematical modelling framework, rooted in social physics, aimed at answering macro-level questions regarding the evolving social norms in human populations under the assumption of frequent AI use. Our analysis is part-analytical, and part-simulation, enabling us to characterize the long-term dynamical consequences under a diverse set of starting assumptions. We highlight the risk of value lock-in, and normative mode collapse, prominently featured in non-adaptive alignment formulations. Beyond alignment, we advocate for the wider adoption of these kinds of social physics models as an epistemic bridge: enabling rapid, rigorous, and quantitatively-grounded hypothesis testing for sociotechnical foresight in general AI futures, and acting as a tractable precursor to more computationally expensive large-scale agentic evaluations.
Can AI systems be aligned to human values? The popularization of large language models (LLMs) and multi-modal foundation models has seen a rise in harms spanning from toxic speech and hallucinations to AI agents executing unauthorized actions. Within the field of AI safety, these harmful instances are often framed as t...
Andrew Smart, Shazeda Ahmed, Jackie Kay et al.· 0 citations
A five-dimensional diagnostic framework that maps the challenges of human-AI collaboration across Integration, Representation, Scale, Temporality, and Adequacy gaps and shows that augmentation remains the dominant and most viable mode of use in complex environments.
Ganesh Sankaran, Marco A. Palomino, G. Siestrup· Big Data and Cognitive Compu...· 0 citations
The article examines how the pursuit of technological sovereignty in the field of artificial intelligence, instead of bringing true autonomy, generates a wide range of hybrid dependencies. These dependencies arise at the levels of computing infrastructure, data governance, AI system evaluation methodologies, and deploy...
This perspective clarifies where current alignment methods genuinely benefit from game-theoretic analysis, where the framework is looser, and what challenges remain in building robust, adaptive, and verifiable AI systems.
Yaxin Cai, Zhong-Rui Zhao, Zhigang Lu et al.· 0 citations
The regulatory sandboxes can be viewed as pedagogical environments for AI: dynamic spaces where alignment develops as a formative process, progressively shaping autonomous behaviors through interaction and cooperation in scenarios of increasing complexity.
Marica Notte, Ludovica Marinucci, V. Santucci· 0 citations
A real-world AI evaluation framework focused on AI-in-use: how people actually appropriate, adapt, and work around AI systems in context, and what consequences follow over time is proposed.
Reva Schwartz, Gabriella Waters· Social science computer revi...· 0 citations
We use cookies to run the site and, with your consent, for analytics and to show ads.
See our Cookie Policy.