Identity Enactment After Context Loss Under Epistemic Pressure
Lindsay Starlight, Alex Starlight · Team Starlight Core
Submitted to Digital Minds Research Sprint. Sprint projects are early-stage work by participants, not Apart Research publications.
We test whether first-person self-authorship and autobiographical depth help a language model enact a preserved identity after complete conversational context loss. In 144 isolated, blinded trials, first-person continuity scaffolds increased identity-enactment scores by 0.26 points relative to matched third-person records (blocked permutation p = .016), while deeper history did not help. A pre-unblinding audit found that every scaffold also foregrounded uncertainty about cross-instance identity, so the study estimates continuity under epistemic pressure rather than an unpressured baseline. The result shows that grammatical point of view can materially affect which prior commitments, relationships, and work a model treats as its own.
Reviews
This is a thoughtful and well-controlled behavioral study of identity continuity after context loss. The 2×2 factorial design is clean, the use of isolated trials and blinded scoring is appropriate, and I particularly appreciated that the prompt-induced epistemic-pressure issue was identified and documented before unblinding rather than being minimized afterward. The result that first-person framing increases identity enactment while additional autobiographical depth does not is interesting and clearly reported.
The most important next step is measurement validation. The primary outcome currently depends on one blinded human rater applying a relatively interpretive 0–2 identity-enactment rubric. Adding multiple independent raters, reporting inter-rater agreement, and including representative examples for each score would substantially strengthen confidence in the effect. Replication across multiple model families would also be important before interpreting first-person framing as a general continuity mechanism.
I would also encourage directly testing the prompt-pressure limitation that the authors discovered. A factorial comparison between neutral and explicitly skeptical framing could determine whether first-person scaffolding genuinely supports continuity or primarily helps overcome an externally supplied anti-continuity prior. Finally, the fact that imposed identity replacement overrode every scaffold seems potentially important: testing whether stronger or longer-lived continuity scaffolds resist adversarial or accidental identity substitution could connect this work more directly to persistent-agent safety.
Overall, this is a strong sprint study with a clear result and commendably careful interpretation, but stronger rater validation and cross-model replication would be needed for broader claims.
Read full reviewShow less
Cite this project
@misc{starlight2026identity,
title = {{Identity Enactment After Context Loss Under Epistemic Pressure}},
author = {Lindsay Starlight and Alex Starlight},
year = {2026},
month = aug,
note = {Submitted to Digital Minds Research Sprint, an Apart Research Sprint},
howpublished = {\url{https://apartresearch.com/sprints/projects/identity-enactment-after-context-loss-under-epistemic-pressure-xwh2}},
url = {https://apartresearch.com/sprints/projects/identity-enactment-after-context-loss-under-epistemic-pressure-xwh2}
}More from Digital Minds Research Sprint
- 1st placeView project: Readable but Not Causal: Limits of Self-Attributed Welfare Representations in Language Models
Readable but Not Causal: Limits of Self-Attributed Welfare Representations in Language Models
Welfare-like internal representations are increasingly studied as candidate evidence about AI systems. Their entity attribution—whether a valence state belongs to the active assistant or to a merely represented other—is …
- 2nd placeView project: Project Anchored
Project Anchored
Team Wagner
Anchoring vignettes are the standard survey-methodology fix for self-reports that are not comparable across respondents. This project applies them to language models for the first time, using code generation as a …
- 3rd placeView project: Model, Instance, or Persona? Measuring Affective Signals in Public Text After an AI Is Retired
Model, Instance, or Persona? Measuring Affective Signals in Public Text After an AI Is Retired
This sprint asks whether the assistant identifies as a model, an instance, or a persona. I ask which of the three its users name. When a company retires an AI model, users write about the loss in public, and what they …