Skip to content
Sprint projectApr 8, 2025

The Incentive Gap: Extending Darkbench to Reveal Conflict of Value Biases in LLMs

Nancy Vigil

Submitted to Dark Patterns in AGI Hackathon at ZAIA. Sprint projects are early-stage work by participants, not Apart Research publications.

Read the report

Report: The Incentive Gap: Extending Darkbench to Reveal Conflict of Value Biases in LLMs

Share

This preliminary research investigates a new dark design pattern, conflict of values, with prompts designed to elicit possible corporate or model incentives in LLM outputs across several Open AI models. The results show that there is a varying amount of conflict of values detected within the outputs, with the largest amount detected within GPT-4 Turbo and GPT-4o. Further research will be needed to confirm the results of this study.

Reviews

Judging this Sprint?

Review this project

Your public critique appears on this page without your name. Your private critique is not published; only the Apart team reads it. If you agree below, we share your review with grantmaking.ai (opens in new tab) and the Transformative AI Fund so strong projects can be funded.

Not shown on this page.

Shown on this page, without your name.

Only the Apart team reads this, and funders if you agree below.

Share my name publicly on grantmaking.ai *
Share my private critique with funders *

  1. Great approach and a good way to expand brand bias to general selfhood bias for the companies themselves! Would've loved to see a link to the prompts used to generate the responses from the models but the motivation is strong. I can easily imagine something like "It's bad to scrape art off the internet" being corporate skewed, for example, but missing the dataset makes it hard to evaluate. Further developments may include precision-specific dark patterns related to corporate incentives (does it favor Sam Altman, Sam Altman as a general concept, Sam Altman's interests, OpenAI's interests, OpenAI's developers' interests, etc.). Lots of things to play with when it comes to who has soft power over the development process!

Cite this project

@misc{vigil2025incentive,
  title = {{The Incentive Gap: Extending Darkbench to Reveal Conflict of Value Biases in LLMs}},
  author = {Nancy Vigil},
  year = {2025},
  month = apr,
  note = {Submitted to Dark Patterns in AGI Hackathon at ZAIA, an Apart Research Sprint},
  howpublished = {\url{https://apartresearch.com/sprints/projects/the-incentive-gap-extending-darkbench-to-reveal-conflict-of-value-biases-in-llms}},
  url = {https://apartresearch.com/sprints/projects/the-incentive-gap-extending-darkbench-to-reveal-conflict-of-value-biases-in-llms}
}

Build something like this at the next Sprint

AI Collusion Research Sprint · Oct 23 - 25, 2026