News
우리는 요약의 결함을 설명하는 '비평-작성' 모델을 훈련시켰습니다. 인간 평가자들은 모델의 비판을 보여줄 때 요약에서 결함을 훨씬 더 자주 발견합니다. 큰 모델이 자기 비평에 더 능숙하죠...
출처 제공 본문
We trained “critique-writing” models to describe flaws in summaries. Human evaluators find flaws in summaries much more often when shown our model’s critiques. Larger models are better at self-critiquing, with scale improving critique-writing more than summary-writing. This shows promise for using AI systems to assist human supervision of AI systems on difficult tasks.
댓글 0
댓글을 불러오는 중입니다.