Reddit
Sam Altman, 차세대 모델 RL 학습 중단 배경 설명
모델 성능 향상 속도가 안전성 및 정렬 기술을 앞지르는 현상에 대해 선제적 대응 조치임을 시사.
원문 제목 Explanation from @sama on RL training pause: "Model progress is now extremely rapid, and we always said we would take action if we felt that model capabilities were outstripping the pace of safety and alignment."
원문 보기 ↗