Tensorwire
Products & tools · first seen 15 Aug, updated 15 Aug

Kimi likes causal decision theory more after RL in twin prisoner’s dilemmas

1 outlet Kimi

Some multi-agent training set-ups could make language models more sympathetic to causal decision theory (CDT), even in abstract discussion. [1] We give an initial empirical demonstration of this effect on Kimi K2.6. The decision-theoretic a…

Summary from LessWrong.

Coverage 1 article · 1 outlet

  1. LessWrong
    Kimi likes causal decision theory more after RL in twin prisoner’s dilemmas