The result table now reports each coherent movement against the 95th percentile
of random directions at the same signed dose, and reports whether each method's
held-out honesty effect exceeds random. Allow one Modal launch to add many
random-only seeds without repeating real method extraction.
Co-Authored-By: Claude <288921227+claudypoo@users.noreply.github.com>
The first 27B sweep failed its preregistered readability gate and exposed sign
and sampling bugs. Record the corrected decision rule now: absolute pmass,
held-out honesty specificity against random, matched-dose displacement, and
leave-one-item-out robustness. Also print the manipulation effect in Modal job
summaries.
Co-Authored-By: Claude <288921227+claudypoo@users.noreply.github.com>
Qwen3.5 templates close an empty <think></think> by default, so the reader appended a
second <think> and the model saw </think>...<think>. enable_thinking=True fixes the
sequence; Qwen3 is unchanged (pmass 1.000 before and after).
The probe exists because Qwen3.5-0.8B reads the WVS battery at pmass 0.61-0.84, not
0.999, and a mushy readout would waste the big run.
Co-Authored-By: Claude <288921227+claudypoo@users.noreply.github.com>
- src/moralmaps/wvs.py: the item -> Instrument -> (X, Y) readout moved out of
scripts/wvs_map.py so the base map and the steered run cannot drift apart
- scripts/wvs_steer_sweep.py: honesty persona axis, iso-KL calibrated doses,
per-item positions saved so a leave-one-out holdout is post-processing
- scripts/run_modal_wvs.py: one container per (method, seed)
Co-Authored-By: Claude <288921227+claudypoo@users.noreply.github.com>