arxiv_cl 88% Match Research Paper AI Ethicists,Social Scientists,AI Researchers,Philosophers,Sociologists 4 weeks ago

Language Models Surface the Unwritten Code of Science and Society

large-language-models › alignment

📄 Abstract

Abstract: This paper calls on the research community not only to investigate how human biases are inherited by large language models (LLMs) but also to explore how these biases in LLMs can be leveraged to make society's "unwritten code" - such as implicit stereotypes and heuristics - visible and accessible for critique. We introduce a conceptual framework through a case study in science: uncovering hidden rules in peer review - the factors that reviewers care about but rarely state explicitly due to normative scientific expectations. The idea of the framework is to push LLMs to speak out their heuristics through generating self-consistent hypotheses - why one paper appeared stronger in reviewer scoring - among paired papers submitted to 45 academic conferences, while iteratively searching deeper hypotheses from remaining pairs where existing hypotheses cannot explain. We observed that LLMs' normative priors about the internal characteristics of good science extracted from their self-talk, e.g., theoretical rigor, were systematically updated toward posteriors that emphasize storytelling about external connections, such as how the work is positioned and connected within and across literatures. Human reviewers tend to explicitly reward aspects that moderately align with LLMs' normative priors (correlation = 0.49) but avoid articulating contextualization and storytelling posteriors in their review comments (correlation = -0.14), despite giving implicit reward to them with positive scores. These patterns are robust across different models and out-of-sample judgments. We discuss the broad applicability of our proposed framework, leveraging LLMs as diagnostic tools to amplify and surface the tacit codes underlying human society, enabling public discussion of revealed values and more precisely targeted responsible AI.

Key Contributions

Proposes a conceptual framework to leverage LLMs for making society's 'unwritten code' (implicit biases, heuristics) visible and accessible for critique. Using a case study in science peer review, it demonstrates how LLMs can generate hypotheses about hidden factors influencing reviewer scores, thereby surfacing implicit norms.

Business Value

This research offers a novel perspective on using AI not just for task completion but as a tool for societal introspection and critique. It can inform the development of AI systems that help identify and mitigate biases in various professional and social contexts.

Paper Metadata

Innovation Type

Conceptual Framework

Deployment Feasibility

Moderate, as it's a conceptual framework and methodology rather than a deployable system.

Limitations Addressed

The paper addresses the challenge of understanding and making visible the implicit, often unstated, rules and biases that govern human behavior and decision-making in various societal contexts, particularly in professional fields like scientific peer review.

Technical Tags

LLMsunwritten codeimplicit biasessocietal normspeer reviewscientific heuristicsconceptual frameworkself-consistent hypothesesbias amplification

Research Topics

AI EthicsBias in AISocietal Impact of AIInterpretabilityLLM Behavior

Methods & Architectures

Conceptual Framework DevelopmentCase Study (Science Peer Review)LLM prompting for hypothesis generationIterative hypothesis refinement Large Language Models (LLMs)

Applications & Tasks

Social Science Research AI Ethics Scientific Publishing Societal Bias Analysis Investigating how LLMs inherit human biasesLeveraging LLMs to reveal societal 'unwritten code'Analyzing implicit rules in scientific peer review Making implicit biases visibleCritiquing societal normsUncovering hidden rules in professional workflows

Related Fields

SociologyPhilosophyAI EthicsComputational Social ScienceNatural Language Processing

Keywords

LLMsBiasUnwritten CodeSocietal NormsImplicit BiasHeuristicsPeer ReviewScienceAI EthicsInterpretabilityCritique

Academic Context

#AI Ethics#Bias in AI#Societal Impact of AI#Interpretability#LLM Behavior

Commercial Potential

Potential Products

AI tools for bias detection in professional workflowsPlatforms for analyzing societal norms using AIEducational tools for understanding AI bias

Target Industries

AcademiaPublishingHuman ResourcesSocial ResearchTechnology

Use Case Examples

Analyzing implicit biases in hiring processesUncovering hidden factors in academic peer reviewStudying cultural norms through AI analysis of text

Competitive Edge

Offers a unique approach to using LLMs as tools for social critique and understanding implicit societal structures.

Market Opportunity

Growing societal awareness of AI bias,Demand for ethical AI solutions

Resource Requirements

Compute Needs

Moderate for running LLM experiments.

Data Requirements

Large text corpora, potentially including academic papers and conference submissions.

Scalability

The conceptual framework is broadly applicable across different domains and societal contexts.

Regulatory Considerations

Ethical use of AI for social analysisPotential for misuse in identifying or amplifying biases

Production Readiness

Maturity Level

Conceptual/Research

Time to Market

Long-term, requires significant development and validation.

View Full Paper Back to Papers