Human Reviewers Read Private ChatGPT Conversations

OpenAI contractors are manually reading user chats to refine AI responses, raising urgent questions about privacy and data handling.
A recent report by 404 Media reveals that OpenAI employs hundreds of human contractors to manually review ChatGPT transcripts. This initiative, internally known as Project Lily, involves these reviewers assessing the quality of AI responses to ensure they are helpful and free from overly robotic language. The process is not just a technical audit; it is a direct human intervention in the model’s behavior.
For users, this means that their conversations may be read by people outside of OpenAI’s core engineering team. While the company claims these chats are anonymized, the report highlights that personal details and sensitive information often remain visible to these reviewers. This creates a significant trade-off: better conversational quality for the AI, at the cost of user privacy and the potential exposure of private data.
Reviewers Assess Tone And Style
The primary task for these contractors is to judge how well the AI answers questions and whether its tone is appropriate. Reviewers look for specific flaws, such as excessive use of emojis, patronizing language, or sycophantic behavior. They also check to prevent the AI from anthropomorphizing itself, ensuring it does not claim to have personal experiences or feelings. This work is described as repetitive and routine, yet it commands a high hourly wage, suggesting a significant financial investment in human oversight.
Personal Data May Slip Through
Despite efforts to anonymize data, OpenAI admits that the filtering process can miss personal information, particularly in shorter conversations. Reviewers do not see user names, but they do access a summary of the user's interests, context, and even location. This is particularly concerning for users who treat the chatbot as a confidant or therapist, sharing deep secrets with the assumption of privacy. The existence of these reviews undermines the notion that AI interactions are a private, closed loop.
Default Settings And Deletion Limits
Most consumer chatbots have a setting that allows companies to use user data for improvement, and this is often turned on by default. Even if a user deletes a chat, it may have already been processed and stored in a dataset. This lack of retroactive control means that past conversations remain vulnerable to review and potential use in future model training. The reliance on human reviewers also signals that technological advancement alone is not sufficient to ensure high-quality AI interactions.






