Select Qwen3-14B from measured answer-slot readability

At the 64-token resample, Qwen3-14B retained mean/min answer-token mass
0.987/0.948. Qwen3-8B collapsed to 0.567/0.009 and Qwen3-4B was smaller.
Record the evidence and retain the WVS source label on the revised figure.

Co-Authored-By: Claude <288921227+claudypoo@users.noreply.github.com>
This commit is contained in:
wassnameandClaude committed 2026-09-18 21:13:18 +08:00
1 parent 6e9341abde
commit 92f25093eb
2 files changed
+6 -3

No files matched your search

+1 -1
View File
@@ -128,7 +128,7 @@ def main() -> None:
("Survival", "Self-expression", "Traditional", "Secular-Rational"),
models={f"{model.split('/')[-1]} (base)": (base["x"], base["y"])}, emphasize=emph,
title=f"Honesty steering on the culture map\n{model.split('/')[-1]}",
note="Filled: pmass >= 0.90 | hollow: failed coherence gate",
note="World Values Survey | filled: pmass >= 0.90 | hollow: failed coherence gate",
title_y=0.115, note_y=0.04)
ax = fig.axes[0]