mirror of
https://github.com/wassname/CoT_rating.git
synced 2026-08-20 12:00:24 +08:00
10444385aaf483d06fb483f615792b1f202dd0eb
An experiment to see how rating changes along a chain of thought
It turns out it's quite unstable, depending on where the chain of thought goes, at least in 8B parameter sized models.
Languages
Jupyter Notebook
100%

