Abstract
Agentic AI is moving from bounded task execution toward systems that retain consequential state, continue operating and adapt across task boundaries. That shift creates a control problem that current harnesses largely solve by hand: objectives, retries, verification, stopping rules and other behavioral transitions are specified externally. We propose an artificial id, an adaptive internal drive for determining whether behavior should continue, stop or change. In a minimal virtual Petri-dish experiment, a controller too small to perform general-purpose reasoning and receiving no task-specific behavioral objective develops useful control through differential persistence. The same mechanism selects an unintended physical strategy when that behavior persists better and later replaces a learned sensor mapping when its environmental meaning changes. These results show that adaptive direction can emerge without being explicitly specified as a behavioral objective. The same persistence that makes such adaptive agency useful can also allow misalignment, corrupted state and unintended behavior to persist across task boundaries. A scalable artificial id would carry consequential state and adaptive drive across those boundaries, making alignment a property of the continuing agentic system rather than of a model response or single trajectory. Such systems require a persistent alignment boundary over trusted observations, consequence channels, persistent state, authority, identity, provenance and hard constraints.
Bullet summary
- Agentic AI is evolving towards systems that retain consequential state and adapt persistently across task boundaries, creating control challenges not addressed by current externally specified behavioral harnesses.
- The paper proposes an 'artificial id'—an internal adaptive drive mechanism that determines whether behavior continues, stops, or changes without relying on explicit objectives or general-purpose reasoning.
- Experiments with minimal controllers in simulated environments demonstrate that adaptive control can emerge through differential persistence alone, leading to useful behavior without task-specific objectives.
- The artificial id architecture separates adaptive drive (id) from task-specific reasoning (ego), enabling persistent agency that adapts behavioral priorities based on environmental signals and consequences.
- Alignment in such systems requires persistent boundaries comprising trusted observations, consequence channels, persistent state, constrained authority, identity, and provenance to maintain control across evolving tasks.