OpenAI has disclosed that a workforce of hundreds of contracted workers reviews and verifies a portion of user interaction data for its generative AI service, "ChatGPT." This initiative is carried out as part of a continuous process to maintain AI safety and improve quality.
Reviews of ChatGPT conversation logs are conducted to improve the accuracy of AI-generated responses, suppress the generation of inappropriate content, and curb the spread of misinformation. Based on the collected data, contracted workers verify the appropriateness of datasets required for training and fine-tuning the AI models.
Human feedback (RLHF: Reinforcement Learning from Human Feedback) is indispensable for advancing large language models (LLMs). To align model behavior in a more helpful and safe direction for humans, a monitoring process involving hundreds of personnel has been established to ensure AI quality and detect defects.
OpenAI plans to continue its research into AI safety and pursue initiatives that balance user privacy protection with model performance improvements. As the societal implementation of AI accelerates, such transparent monitoring processes will continue to be operated as a vital component of the technological foundation.