MEGA Hub

STONIC: A Layered Measurement Contract for LLM Value Profiling

Authors

Do you know Andrei Chetvergov?You can claim authorship or link another user.Do you know Stepan Ukolov?You can claim authorship or link another user.Do you know Timofei Sivoraksha?You can claim authorship or link another user.Do you know Alexander Evseev?You can claim authorship or link another user.Do you know Danil Sazanakov?You can claim authorship or link another user.Do you know Mikhail Solovev?You can claim authorship or link another user.Do you know Sergey Bolovtsov?You can claim authorship or link another user.

Abstract

LLM value studies often merge questionnaire ratings, pairwise choices, and values inferred from generated text into one profile. That merge assumes that the three observations describe the same stable preference. STONIC tests this assumption on 5,144 situations from four banks and 35 fixed model configurations. It compares responses rated in isolation, choices made under counterbalanced conflict, spontaneous answers, and later choices between a model's own answer and authored alternatives. 10 of 17 configurations with usable behavioral data preserve the endorsement-choice relation across banks. Every one of the 17 eligible configurations prefers its own earlier answer (median effect 0.790), although option position changes the choice rate in every eligible configuration. Profile shape transfers most strongly from ratings to conflict choices and weakens for spontaneous text. Three-way annotation of 200 L3 responses provides a task-local check of the semantic audit: FULCRA agrees most closely with the human majority, while DeBERTa retains useful rank information after calibration. Hidden states encode the completed decision more clearly than the prompt alone. Thus the models show reproducible behavioral continuity, but the evidence does not support one scorer-independent value identity across interfaces.

Community

00

Publication notes

Author note
32 pages, 6 figures, including appendices