News
RL-Teacher는 수작업으로 만든 보상 함수 대신 가끔씩 인간의 피드백을 통해 AI를 훈련시키기 위한 오픈소스 인터페이스 구현체입니다. 기저 기술은 단계적으로 발전된 것이었다...
출처 제공 본문
RL-Teacher is an open-source implementation of our interface to train AIs via occasional human feedback rather than hand-crafted reward functions. The underlying technique was developed as a step towards safe AI systems, but also applies to reinforcement learning problems with rewards that are hard to specify.
댓글 0
댓글을 불러오는 중입니다.