News
우리는 학습 에이전트의 손실 함수를 진화시켜 새로운 과제에 대한 빠른 훈련을 가능하게 하는 실험적 메타러닝 접근법인 Evolved Policy Gradients를 출시합니다. 나이...
출처 제공 본문
We’re releasing an experimental metalearning approach called Evolved Policy Gradients, a method that evolves the loss function of learning agents, which can enable fast training on novel tasks. Agents trained with EPG can succeed at basic tasks at test time that were outside their training regime, like learning to navigate to an object on a different side of the room from where it was placed during training.
댓글 0
댓글을 불러오는 중입니다.