
Inside ‘Project Lily’: The Humans Reading Your ChatGPT Chats
OpenAI is employing hundreds of contractors to read and evaluate real user prompts from ChatGPT to enhance the quality and safety of the chatbot's responses. These human reviewers assess conversations that may contain sensitive personal data, despite OpenAI's efforts to anonymize the inputs using a privacy filter. The practice highlights a significant privacy concern for users who may not realize their intimate details are being scrutinized by third-party workers.
- ▪OpenAI hires contractors to review real user prompts and rate the chatbot's generated replies to improve model performance.
- ▪Internal documents indicate that reviewers are tasked with reducing anthropomorphism and sycophancy in ChatGPT's responses.
- ▪OpenAI uses a Privacy Filter to remove personal information, but acknowledges that sensitive details can still reach the human reviewers.
- ▪Anthropic has also confirmed that it utilizes human review processes to improve its own AI models.
- ▪Contractors work in three stages: reading the prompt, summarizing the user's intent, and critiquing the AI's response.
404 Media files mainly under tech. We currently carry 91 of its stories.
Story provenance
Source · retrieval · rights · ranking — open for full record
inspect →
Story provenance
Attribution is not the same as permission. This drawer separates discovery metadata, excerpts, WeSearch-generated summaries, reuse status, and whether the publisher receives the visit. Nothing here claims a legal grant the publisher has not made.
Record
| Original publisher | 404 Media |
| Canonical URL | https://www.404media.co/inside-project-lily-the-humans-reading-your-chatgpt-chats/ |
| Publication time | Mon, 14 Sep 2026 14:22:43 GMT |
| Retrieval time | 2026-09-14T14:26:51.290Z |
| Last seen | 2026-09-14T14:26:51.290Z |
| Headline source | Publisher (no WeSearch rewrite) |
| Excerpt source | publisher body |
| Excerpt method | First ~120 words (~800 chars) of extracted publisher body, fair-use limited. |
| Summary | WeSearch · cerebras-chat (WeSearch summarizer) |
| Summary source text | contentText |
| Citation coverage | Summary is a WeSearch-generated derivative; primary citation is the original publisher URL. |
| Cluster | rd-EqpLFcAyk · 1 stories |
| Cluster logic | Grouped by semantic title/content similarity across sources within a rolling window. Same-publisher template collisions are excluded from coverage comparison. |
| Ranking reason | Story pages are not engagement-ranked. Hub feeds use recency, with optional source-diversified chronological ordering (cap consecutive stories per source). No personalized ranking. |
| Publisher visit | Yes — open original |
| Substitutes article? | No — link-out required for full text |
Rights status (four layers)
WeSearch handling by dimension
| Indexing | May the item be indexed (stored, ranked, made findable)? | Allowed |
| Snippet | May a short excerpt of the publisher's text be shown? | Allowed |
| AI summary | May WeSearch generate its own short summary of the article? | Limited |
| Retrieval / RAG | May the content be exposed for third-party retrieval-augmented generation? | Not asserted |
| Model training | May the content be used to train AI models? | Not asserted |
| Commercial reuse | May the content be reused commercially? | Not permitted |
Basis: Derived from the published RSS/Atom feed. Contact: [email protected]. Reviewed: 2026-07-24.
Opening excerpt (first ~120 words) tap to expand
OpenAI is hiring hundreds of contractors who read a massive stream of real users’ ChatGPT prompts, with the prompts sometimes including sensitive personal information, 404 Media has learned. The prompts these people review can include whole conversations between users and the chatbot, conversations that most of ChatGPT’s more than 900 million users probably don’t realize may be read by actual people.The goal of these prompt review teams is to improve the responses ChatGPT gives to its users, with the contractors rating and critiquing the chatbot’s generated replies.
…
Excerpt limited to ~120 words for fair-use compliance. The full article is at 404 Media.