Back to feed
arXiv cs.CL·

LLMs Contain Multitudes: How Deployment Context Reshapes Model-Level Preferences and Values

Signal
78
Hype
15
In three linesStudy of 1.2M decisions showing deployment context (Reddit vs news article) produces far larger variations in model preferences and values than prompt paraphrasing or temperature controls. Measured biases (Global North favoritism) and cardinal exchange rates between outcomes shift by factor 2.47 across contexts, questioning stability of model-level properties.
Read source
Your take?
EvalsAI safetyAlignment

Summary generated by Claude — human-verified