News
OpenAI는 강화 학습으로 훈련된 자동 레드티칭을 사용해 ChatGPT Atlas를 prompt injection 공격에 대비해 강화하고 있습니다. 이 능동적 발견 및 패치 루프는 새로운 전자를 식별하는 데 도움을 줍니다...
출처 제공 본문
OpenAI is strengthening ChatGPT Atlas against prompt injection attacks using automated red teaming trained with reinforcement learning. This proactive discover-and-patch loop helps identify novel exploits early and harden the browser agent’s defenses as AI becomes more agentic.
댓글 0
댓글을 불러오는 중입니다.