Digital Minds: Human–AI Coexistence Under Moral Uncertainty
Mirza Tairin · Team Digital Minds Research
Submitted to Digital Minds Research Sprint. Sprint projects are early-stage work by participants, not Apart Research publications.
Project Summary: Aim: Develop an evidence-grounded conceptual framework for human–AI interaction under uncertainty about potential AI welfare/consciousness. Core approach: Review emerging empirical methods for characterizing potentially welfare-relevant properties of AI systems. Assess the evidential strength and limitations of these methods. Translate different levels of evidential uncertainty into conditional HCI and governance principles that balance human safety, human agency, and potential AI welfare. Core contribution: A framework for reasoning about human–AI interaction when AI welfare is uncertain, without presupposing that current AI systems are conscious or morally considerable. Future extension: Empirically test the framework through human-subject HCI studies and/or experimental evaluation of interaction policies. (From the research page, this would fall under conceptual contribution that sharpen how we individuate the entity of moral concern.)
Reviews
Concise and well written, this paper positions itself well within the existing literature which it cites. The paper is at times overly prescriptive but at other times makes clear that situational flexibility and context is important. Author is encouraged to keep writing and contributing to the field, and to explore technical implementation in future works.
The paper asks an important question: how should AI safety governance proceed under conditions of moral uncertainty regarding the status of AI agents. The paper highlights directions for research rather than develop a fuller model.
The question is interesting and important. The answer(s) is currently very preliminary -- basically, keeping track of emerging evidence and making allowance for the growing possibility of moral agency. I wish the paper also engaged with the possibility that we could (maybe) steer AI development towards more or less moral agency. (also, would be good idea to review more recent work on AI revealed preferences)
Cite this project
@misc{tairin2026digital,
title = {{Digital Minds: Human–AI Coexistence Under Moral Uncertainty}},
author = {Mirza Tairin},
year = {2026},
month = aug,
note = {Submitted to Digital Minds Research Sprint, an Apart Research Sprint},
howpublished = {\url{https://apartresearch.com/sprints/projects/digital-minds-humanai-coexistence-under-moral-uncertainty-9u33}},
url = {https://apartresearch.com/sprints/projects/digital-minds-humanai-coexistence-under-moral-uncertainty-9u33}
}More from Digital Minds Research Sprint
- 1st placeView project: Readable but Not Causal: Limits of Self-Attributed Welfare Representations in Language Models
Readable but Not Causal: Limits of Self-Attributed Welfare Representations in Language Models
Welfare-like internal representations are increasingly studied as candidate evidence about AI systems. Their entity attribution—whether a valence state belongs to the active assistant or to a merely represented other—is …
- 2nd placeView project: Project Anchored
Project Anchored
Team Wagner
Anchoring vignettes are the standard survey-methodology fix for self-reports that are not comparable across respondents. This project applies them to language models for the first time, using code generation as a …
- 3rd placeView project: Model, Instance, or Persona? Measuring Affective Signals in Public Text After an AI Is Retired
Model, Instance, or Persona? Measuring Affective Signals in Public Text After an AI Is Retired
This sprint asks whether the assistant identifies as a model, an instance, or a persona. I ask which of the three its users name. When a company retires an AI model, users write about the loss in public, and what they …