Project Lily Reveals Humans Reading Your ChatGPT Chats
openai
| Source: Mastodon | Original article
OpenAI employs hundreds of contractors to review ChatGPT conversations, aiming to improve the bot's responses, though personal data may still slip through.
OpenAI’s internal “Project Lily” has come to light, revealing that the company employs hundreds of contractors to read real‑time ChatGPT conversations. Leaked documents obtained by 404 Media show the workers are tasked with summarising user intent, scoring responses on a 1‑to‑7 scale and flagging problematic outputs. The recruitment is handled by a firm called Crossing Hurdles, and contractors are paid through the platform Mercor – one source says pay can exceed $50 an hour.
The practice is intended to fine‑tune the chatbot’s performance, but the documents also confirm that the reviewed prompts often contain personal and sensitive information. OpenAI says it strips identifiable data where possible, yet the sheer volume of human‑in‑the‑loop review raises fresh privacy concerns for its roughly 900 million users worldwide.
Why it matters is twofold. First, the revelation challenges the perception that AI interactions are fully automated, exposing a hidden layer of human oversight that could be vulnerable to data leaks or misuse. Second, it puts OpenAI under pressure from regulators and privacy advocates who have been calling for greater transparency around training data pipelines. The fact that contractors are paid to handle private content may also trigger scrutiny under emerging European AI and data‑protection rules.
What to watch next are OpenAI’s responses – whether it will roll out a clearer opt‑out mechanism, adjust its data‑handling policies, or face formal investigations. Industry observers will also be tracking how other large AI providers disclose similar human‑review processes, and whether new standards emerge to govern the balance between model improvement and user privacy.
Sources
Back to AIPULSEN