News
OpenAI는 24개 환경에서 13가지 평가를 포함하는 사고 연쇄 모니터링 기능을 위한 새로운 프레임워크와 평가 스위트를 도입했습니다. 우리의 연구 결과는 모델의 내부 리얼을 모니터링하는 것이...
출처 제공 본문
OpenAI introduces a new framework and evaluation suite for chain-of-thought monitorability, covering 13 evaluations across 24 environments. Our findings show that monitoring a model’s internal reasoning is far more effective than monitoring outputs alone, offering a promising path toward scalable control as AI systems grow more capable.
댓글 0
댓글을 불러오는 중입니다.