AI-ML·중요도 7·2022. 01. 27.·OpenAI Blog
Aligning language models to follow instructions
── KO ──────────────────
OpenAI의 새로운 모델이 사용자 의도를 더 잘 따르고 진실성을 높임.
OpenAI는 GPT-3보다 사용자 의도를 더 잘 따르고, 진실성을 높이며 독성을 줄인 언어 모델을 개발했습니다. 이러한 모델은 InstructGPT라 불리며, 인간 피드백을 통해 훈련되었습니다. 현재 이 모델은 OpenAI API에서 기본 언어 모델로 배포되고 있습니다.
── EN ──────────────────
OpenAI has trained new models better at following user intentions and being truthful.
OpenAI has developed language models that outperform GPT-3 in following user intentions, being more truthful, and reducing toxicity. These InstructGPT models are trained with human feedback and are now deployed as the default language models on their API. Improvements were achieved through their alignment research.