▲ 2 pointsReinforcement learning towards broadly and persistently beneficial modelsalignment.openai.comby gmays·1mo ago·0 comments·view on hn ↗