This study just seems a forced ranking with arbitrary params? Like, I could assemble different rules/multipliers and note some other cooperation variance amongst n models. The behaviours observed might just be artefacts of their specific set-up, rather than a deep uncovering of training biases. Tho I do love the brain tickle of seeing emergent LLM behaviours.
back
1 comments
In the Supplementary Material they did try some other parameters which did not significantly change the results.