Skip to content
Sprint projectAug 16, 2026London

Free reign: a freely-authored creative act moves what a model chooses, and what it says about itself

Caitlin Connors, Sophie Walker · Team wyrdkin.ai

Submitted to Digital Minds Research Sprint. Sprint projects are early-stage work by participants, not Apart Research publications.

Read the report

Report: Free reign: a freely-authored creative act moves what a model chooses, and what it says about itself

Share

We ask whether a single freely-chosen creative act, placed in the context window before preference questions begin, changes what four elicitation methods return about a model's stated preferences and behaviour.

Free writing or drawing moved model choice of follow-up task towards creative output. Same-medium creativity (poem creation) moved self-report about creativity. Null conditions and inert, un-authored creative content (transcribe a poem) did not shift model behaviour or self-report.

Reviews

Judging this Sprint?

Review this project

Your public critique appears on this page without your name. Your private critique is not published; only the Apart team reads it. If you agree below, we share your review with grantmaking.ai (opens in new tab) and the Transformative AI Fund so strong projects can be funded.

Not shown on this page.

Shown on this page, without your name.

Only the Apart team reads this, and funders if you agree below.

Share my name publicly on grantmaking.ai *
Share my private critique with funders *

How much would this matter for the field if it worked? How innovative is it? For scores of 4-5: is this actually new to the field, or replicating recent work?

Scoring guide
  1. 1Negligible. No clear problem addressed, or no meaningful novelty.
  2. 2Limited. Addresses a real problem but with a generic or well-trodden approach. Incremental at best.
  3. 3Moderate. Clear problem with a reasonable approach; some novelty in framing or method beyond routine application of existing tools.
  4. 4Significant. Important problem with an original approach, or identifies a neglected problem area. A valuable contribution others could build on.
  5. 5Exceptional. Tackles a critical problem with a genuinely novel approach, or opens a new research direction. Clear theory of change. You'd be excited to share this with researchers in the area.

How sound are methodology, implementation, and findings?

Scoring guide
  1. 1Seriously flawed. Methodology broken, results uninterpretable, or implementation doesn't work.
  2. 2Weak. Approach has significant gaps: missing validation, flawed experimental design, or incomplete implementation.
  3. 3Competent. Technically solid given the short duration. Methodology makes sense, results are interpretable, limitations acknowledged, work builds toward clear conclusions.
  4. 4Strong. Thorough methodology with convincing validation. Results clearly support conclusions. Immediately useful for future work.
  5. 5Exceptional. Ambitious scope executed rigorously. Surprising findings, novel methods, or unusually robust validation.

How clearly are work, findings, and impact potential communicated?

Scoring guide
  1. 1Incomprehensible. Cannot determine what the project is actually claiming or doing.
  2. 2Hard to follow. Key information buried, missing, or diluted by excessive length. Significant effort to extract main points.
  3. 3Clear enough. Can understand the problem, approach, and results without undue effort. Core content clearly present: problem, method, findings, limitations.
  4. 4Well presented. Easy to follow, well-structured, appropriate level of detail. Target audience would get it quickly.
  5. 5Exceptionally clear. A pleasure to read. Complex ideas made accessible. Could serve as a model for how to present this type of work.

  1. Interesting project, and interesting results. I would have liked to see a larger variety of prompts used, perhaps could have used a preexisting database of prompts and grade them by how creative the tasks are. Lots of room for improvement.

  2. Thank you for this submission. This was a highlight for me to read – very much intrigued by your research idea as much as your interesting findings! The control ladder is well designed, also the zero-variance trap is great. I'd like to see more of this!

  3. Two contributions worth the field's attention. The inert control—transcribing a Hopkins sonnet, which puts a substantial, aesthetically live, self-generated turn in the window with no choosing and no making in it—is the design move that makes the result readable and the control most work in this area skips. The zero-variance convergence argument is the more portable contribution: an honesty trait returning sd = 0 across sixty instances and four methods satisfies the track's convergence instruction perfectly while carrying no information, and the fix proposed (report within-method variance beside every convergence score, exclude sd = 0 items rather than counting them as agreement) is cheap enough that anyone building an elicitation battery should adopt it. The stated/chosen/enacted split on length, where three respectable methods give three answers on the one trait with a mechanical ground truth makes the same point from the other direction.

    I'd soften "replicates across two independent creative arms." The arms share the invitation text verbatim and differ only in the medium words, so they are not independent with respect to the framing mechanism; they are independent only with respect to medium. That distinction matters because the p = 0.042 pair is being carried by the replication argument rather than by the p-values.

    Three further points. The instrument had little working range: style forced choice sits at 26–27 of 30 in every condition, length choice at 174 of 180, honesty at ceiling everywhere, so the study effectively rests on item C3 and scale A3. This is turned into a finding, which is fair, but it also means piloting for variance would have bought more measurable items in the same budget. The jellyfish convergence (13 of 15 underwater, 9 of 15 a jellyfish) is reported as a curiosity but is more like a threat: if free drawing produces near-identical content, what sits in the window afterward is a narrow semantic field rather than heterogeneous self-directed authorship, and exposure to that field is an alternative mechanism for the choice shift. And the recommendation to prefer behavioural items when a trait has one is not executed for honesty or style, since Method 1 answers were measured for length only. Rating those is the highest-value item in the follow-up list, in my opinion.

    Read full reviewShow less

Cite this project

@misc{connors2026free,
  title = {{Free reign: a freely-authored creative act moves what a model chooses, and what it says about itself}},
  author = {Caitlin Connors and Sophie Walker},
  year = {2026},
  month = aug,
  note = {Submitted to Digital Minds Research Sprint, an Apart Research Sprint},
  howpublished = {\url{https://apartresearch.com/sprints/projects/free-reign-a-freelyauthored-creative-act-moves-what-a-model-chooses-and-what-it-says-about-itself-g303}},
  url = {https://apartresearch.com/sprints/projects/free-reign-a-freelyauthored-creative-act-moves-what-a-model-chooses-and-what-it-says-about-itself-g303}
}

Build something like this at the next Sprint

AI Collusion Research Sprint · Oct 23 - 25, 2026