docs: note SGTM is the latest gradient-routing paper (same authors)

Co-Authored-By: Claudypoo <288921227+claudypoo@users.noreply.github.com>
This commit is contained in:
wassname
2026-06-07 11:56:58 +00:00
co-authored by Claudypoo
parent 637f9388c8
commit 5fd980244b
+2 -1
View File
@@ -78,5 +78,6 @@ For the original paper (the substrate: reward-hacking LeetCode env)
- LessWrong post: ./docs/papers/2025_lw_ariahw_steering-rl-training-benchmarking-interventions.md
- Code: ./docs/vendor/rl-rewardhacking
For the gradient-routing prior (SGTM; source of the absorption/leakage vocab)
For the gradient-routing prior (SGTM = latest gradient-routing paper, same authors as
the original; source of the absorption/leakage vocab)
- ./docs/papers/grad_routing/paper_sgtm.md