Does the Instrument Change the Preference? Measuring Cross-Method Invariance in LLM Preference Elicitation
Morel Nicolas (NicoMrx)
We tested whether preference-like signals from LLMs remain invariant across three elicitation methods when the underlying semantic comparison is held constant. Across Gemini 2.5 Flash-Lite, GPT-5.6 Sol, Claude Opus 4.6, and Claude Sonnet 4.6, we collected 1,080 fresh-context responses using direct self-report, forced choice, and allocation-based elicitation. Cross-method convergence differed substantially by model. Gemini retained directional disagreement after ties were removed, while much of the divergence in Sol, Opus, and Sonnet came from method-dependent expression of indifference. A negative control was especially diagnostic: methods permitting indifference returned TIE in 80/80 responses, while forced choice selected A in 39/40. The results show that highly repeatable elicitation methods can still disagree, and that response format can manufacture apparent preference signals.
No reviews are available yet
Cite this work
@misc {
title={
(HckPrj) Does the Instrument Change the Preference? Measuring Cross-Method Invariance in LLM Preference Elicitation
},
author={
Morel Nicolas (NicoMrx)
},
date={
},
organization={Apart Research},
note={Research submission to the research sprint hosted by Apart.},
howpublished={https://apartresearch.com}
}


