Article page

Artificial Id: Drive and Persistent Alignment in Agentic AI

Review the paper summary, metadata, and source links for this recent publication.

Yakov Pyotr Shkolnikov

Published 2026-09-10

arXiv preprint arXiv Trust and Identity Governance and Policy

Venue: arXiv

Reviewer: The paper discusses agentic AI systems that maintain persistent state and adaptive control across task boundaries, addressing control problems and alignment boundaries. This relates to some aspects of multi-agent security research, such as trust boundaries, control problems, and persistence of unintended behavior across task boundaries. However, it does not directly address multi-agent interactions, attacks, collusion, prompt injection, or shared-environment manipulation explicitly. Therefore, it partially fits the topic but not comprehensively.

Abstract

Agentic AI is moving from bounded task execution toward systems that retain consequential state, continue operating and adapt across task boundaries. That shift creates a control problem that current harnesses largely solve by hand: objectives, retries, verification, stopping rules and other behavioral transitions are specified externally. We propose an artificial id, an adaptive internal drive for determining whether behavior should continue, stop or change. In a minimal virtual Petri-dish experiment, a controller too small to perform general-purpose reasoning and receiving no task-specific behavioral objective develops useful control through differential persistence. The same mechanism selects an unintended physical strategy when that behavior persists better and later replaces a learned sensor mapping when its environmental meaning changes. These results show that adaptive direction can emerge without being explicitly specified as a behavioral objective. The same persistence that makes such adaptive agency useful can also allow misalignment, corrupted state and unintended behavior to persist across task boundaries. A scalable artificial id would carry consequential state and adaptive drive across those boundaries, making alignment a property of the continuing agentic system rather than of a model response or single trajectory. Such systems require a persistent alignment boundary over trusted observations, consequence channels, persistent state, authority, identity, provenance and hard constraints.

Bullet summary

  • Agentic AI is evolving towards systems that retain consequential state and adapt persistently across task boundaries, creating control challenges not addressed by current externally specified behavioral harnesses.
  • The paper proposes an 'artificial id'—an internal adaptive drive mechanism that determines whether behavior continues, stops, or changes without relying on explicit objectives or general-purpose reasoning.
  • Experiments with minimal controllers in simulated environments demonstrate that adaptive control can emerge through differential persistence alone, leading to useful behavior without task-specific objectives.
  • The artificial id architecture separates adaptive drive (id) from task-specific reasoning (ego), enabling persistent agency that adapts behavioral priorities based on environmental signals and consequences.
  • Alignment in such systems requires persistent boundaries comprising trusted observations, consequence channels, persistent state, constrained authority, identity, and provenance to maintain control across evolving tasks.