Position: Language Models Should be Used to Surface the Unwritten Code of Science and Society

65/100

large language models unwritten code implicit biases heuristics peer review hypothesis generation prior and posterior storytelling contextualization diagnostic tools iterative search self-consistency confidence-weighted voting bayesian perspective gpt-4o o3-mini openreview api academic conferences evaluative settings responsible ai computational social science bias amplification pairwise comparison llm as judge annotation

cs.CL (Computer Science > Computation and Language) cs.DL (Computer Science > Digital Libraries) cs.CY (Computer Science > Computers and Society)

arXiv: 2505.18942 License

Authors

Siyang Wu Honglin Bao James A. Evans Jiwoong Choi Yingrong Mao

AI Summary

Abstract

This position paper calls on the research community not only to investigate how human biases are inherited by large language models (LLMs) but also to explore how these biases in LLMs can be leveraged to make society's "unwritten code" - such as implicit stereotypes and heuristics - visible and accessible for critique. We introduce a conceptual framework through a case study in science: uncovering hidden rules in peer review - the factors that reviewers care about but rarely state explicitly due to normative scientific expectations. The idea of the framework is to push LLMs to speak out their heuristics through generating self-consistent hypotheses - why one paper appeared stronger in reviewer scoring - among paired papers submitted to 46 academic conferences, while iteratively searching deeper hypotheses from remaining pairs where existing hypotheses cannot explain. We observed that LLMs' normative priors about the internal characteristics of good science extracted from their self-talk, e.g., theoretical rigor, were systematically updated toward posteriors that emphasize storytelling about external connections, such as how the work is positioned and connected within and across literatures. Human reviewers tend to explicitly reward aspects that moderately align with LLMs' normative priors (correlation = 0.49) but avoid articulating contextualization and storytelling posteriors in their review comments (correlation = -0.14), despite giving implicit reward to them with positive scores. These patterns are robust across different models and out-of-sample judgments. We discuss the broad applicability of our proposed framework, leveraging LLMs as diagnostic tools to amplify and surface the tacit codes underlying human society, enabling public discussion of revealed values and more precisely targeted responsible AI.

Sponsored Ad NVIDIA Jetson Orin Nano Super Dev Kit Bring generative AI to the edge with the Jetson Orin Nano Super, delivering 67 TOPS of performance for just $249! Buy Now \to

Position: Language Models Should be Used to Surface the Unwritten Code of Science and Society

Authors

AI Summary

Abstract

Full Paper (26 pages)

Related Papers

Deception and Manipulation in Generative AI

Probabilistic Analysis of Copyright Disputes and Generative AI Safety

Missing vs. Unused Knowledge Hypothesis for Language Model Bottlenecks in Patent Understanding

Conversations over Clicks: Impact of Chatbots on Information Search in Interdisciplinary Learning

NVIDIA Jetson Orin Nano Super Dev Kit