Project Lily: Humans Monitoring Your ChatGPT Chats
openai
| Source: Mastodon | Original article
OpenAI is employing hundreds of contractors to review large volumes of real ChatGPT prompts, including full user‑bot conversations that may contain sensitive personal data.
OpenAI has quietly launched an internal programme, codenamed **Project Lily**, that enlists hundreds of contractors to read real‑time ChatGPT prompts and rate the model’s replies. Leaked internal documents and a cache of actual user prompts obtained by 404 Media reveal that reviewers are handed anonymised excerpts of conversations – often whole dialogues – and asked to summarise the user’s request before judging the chatbot’s output. The work is split into three stages: ingest the prompt, produce a concise summary, and evaluate the generated answer.
OpenAI pays participants more than **$50 an hour**, a rate that underscores the company’s willingness to invest heavily in human‑in‑the‑loop quality control. The documents show that the stream of data being examined is massive, drawn from the platform’s roughly **900 million** user base, and includes content that can be highly personal or sensitive. While the company frames the effort as a means to improve model performance and safety, the practice raises fresh privacy concerns for users who assumed their interactions were only processed by algorithms.
The revelation matters because it spotlights a tension between AI development speed and data protection. Human review can catch nuanced failures that automated systems miss, but it also creates a new vector for exposure of private information, especially if anonymisation fails or if contractors mishandle data. Regulators in the EU and Nordic states have been tightening AI‑related privacy rules, and Project Lily could become a focal point for compliance scrutiny.
Going forward, watch for OpenAI’s response to the report, including any policy adjustments or transparency measures. EU data‑protection authorities may launch investigations, and privacy‑focused NGOs are likely to demand clearer user consent mechanisms. As we reported on **14 September 2026**, Project Lily represents a hidden layer of human oversight that could reshape how large‑scale AI services balance improvement with user confidentiality.
Sources
Back to AIPULSEN