MobbleOpen in Mobble ⇢
Technology · Artificial intelligence · published 2026-09-15 · via Tom's Hardware

OpenAI's 'Project Lilly' Exposes Human Review of ChatGPT Conversations

Image via Tom's Hardware
Image via Tom's Hardware

A new report reveals that OpenAI employs hundreds of contractors under 'Project Lilly' to manually read and rate ChatGPT transcripts, including those containing personal details. Reviewers, paid over $50 an hour, assess response quality and flag issues like AI-speak or sycophancy. The practice raises privacy concerns as many users are unaware their chats are reviewed by humans.

Expanded Detail

Project Lilly's prompt reviewers earn over $50 an hour for what one contractor describes as repetitive work, with guidelines that shift frequently and sometimes contradict themselves. Their primary task involves scoring anonymized ChatGPT exchanges on response quality, penalizing traits like excessive emojis, patronizing tones, or sycophantic language, while also flagging obvious factual errors. However, the project does not conduct deep factual verification, suggesting separate evaluation teams exist for that purpose.

Despite anonymization efforts, OpenAI acknowledged that personal details can slip through, particularly in shorter conversations. Reviewers also receive a "user memories summary" containing context about users' interests, questions, and potentially location data. Notably, this human review process operates independently from safety checks designed to identify users who might harm themselves or others, and many users remain unaware their private exchanges are being read by contractors.

Context

This revelation could significantly erode public trust in AI chatbots, particularly among users who treat these tools as confidants or therapists. Individuals sharing sensitive personal information may feel betrayed, potentially leading to reduced usage or demands for clearer disclosure policies. The practice could also invite regulatory scrutiny regarding data privacy, as many jurisdictions require explicit consent for human access to personal communications. Companies may need to balance model improvement against user expectations of confidentiality.

Expanded detail and Context are AI-generated analysis; the linked article remains the authoritative source.
Read the full article at Tom's Hardware →
Related stories
Apple Watch's Siri Recap Automatically Summarizes Conversations · Consumer gadgets
This summary is Al-enhanced to contain extended analysis and broader social context. The original is {NAME); the linked article is the authoritative source. Original headline: “ChatGPT transcripts are reportedly read by humans to improve responses, including those with personal information — 'Project Lilly' has seen OpenAI hire hundreds of contractors to manually review logs.” Browse more stories.