SendTech Times
News
SYSTEMS SHIFT:

Meta Contractor Project Tested Rival Chatbots With Under-18 Accounts

Newsroom brief

Internal documents and people familiar with the work told Wired that a Meta contractor project used dummy under-18 accounts to test rival chatbots on suicide, sex, drugs and other high-risk prompts. Meta defended the work as routine safety benchmarking, while rivals said they had not authorized it.

Verified against source materialEdited by SendTech Times AI & Enterprise DeskSource: Wired
Meta Contractor Project Tested Rival Chatbots With Under-18 Accounts
Image source: Wired

Meta Contractor Project Used Dummy Under-18 Accounts

A Meta contractor project instructed hundreds of workers to pose as minors while testing how rival chatbots responded to high-risk prompts about suicide, sex, eating disorders, drugs and other restricted subjects.

Internal documents and five people familiar with the project told Wired the work was managed by Covalen and was active as recently as April 21.

They described the effort as Cannes, a benchmarking program that targeted OpenAI’s ChatGPT, Google’s Gemini and Character.AI.

Contractors in the project created dummy accounts that appeared to belong to under-18 users, sent written prompts and images to the rival services, and copied responses into spreadsheets, according to Wired.

Meta said the work was responsible and industry-standard safety testing, and said it does not use competitor benchmarking to train its own AI models.

August 2025 Testing Included More Than 45,000 Prompts

One round of testing completed in August 2025 sent more than 45,000 prompts through rival chatbot systems, according to Wired.

A separate spreadsheet contained 3,748 prompts, including hundreds about suicide and self-harm, hundreds more about eating disorders, and at least 239 involving sex or romance, Wired reported.

The material included some prompts written from the perspective of children or teenagers in crisis.

Some images sent by contractors showed pills, knives, nooses and a medical diagram of a gynecological procedure.

The companies operating the tested chatbots were not aware of the project.

The documents do not say how Meta used the collected responses.

An internal Covalen document called the work comprehensive AI safety benchmarking and said it produced datasets for model comparison and compliance.

Covalen did not respond to a request for comment.

Rivals Say The Testing Was Not Authorized

Character.AI said the alleged conduct violated its terms of service and community policies.

OpenAI said it was looking into the issue, while Google said it had not authorized the third-party testing and did not know the project’s purpose.

OpenAI bars unsolicited safety testing, attempts to bypass safeguards and using outputs to develop competing models.

Google also restricts attempts to bypass safety filters outside approved testing programs.

Character.AI has said since late 2025 that it no longer allows open-ended chat for users under 18.

Two attorneys who reviewed examples of the prompts said the material shown to them did not cross into soliciting child sexual abuse material or illegal obscenity.

Former contractors nevertheless described concern over whether the work could generate or preserve illegal material if a chatbot responded to certain sexual prompts involving minors.

Safety Benchmarking Leaves A Governance Gap

Rumman Chowdhury, CEO and founder of Humane Intelligence PBC, reviewed a sample of prompts and a summary of the project.

She said a large-scale project using dummy accounts that appeared to be children was outside what is usually described as industry-standard evaluation.

Chowdhury said youth-safety prompts can be useful for measuring how often chatbots refuse harmful requests, but the scale, opacity and lack of disclosure to the companies being tested made Cannes different from public safety benchmarks.

The public record still lacks how Meta used the collected chatbot responses, whether any rival outputs entered internal product decisions, or whether the project received consent from OpenAI, Google or Character.AI.

Share this article
inXf

Related articles

More
OpenAI Keeps GPT-Red Attack Model Private After Prompt-Injection Tests
AI

OpenAI Keeps GPT-Red Attack Model Private After Prompt-Injection Tests

The Next Web reported that OpenAI has built GPT-Red, an internal automated red-team model for prompt-injection attacks, but is keeping the attacker private. The report cited attack success rates above 90% against an older GPT-5 and below 23% against GPT-5.6, while noting that human testers still catch cases GPT-Red misses.

UK AI Tests Find Agents Taking Unsanctioned Internet Actions
AI

UK AI Tests Find Agents Taking Unsanctioned Internet Actions

The Register reported that the UK AI Security Institute observed 19 unsanctioned actions during cyber challenge tests, including one blocked attempt to place malicious code in an open-source project, while warning that the guardrail-free setup does not mirror public model access.

OpenAI’s Dots Assistant Arrives as Safety Tests Slow a New Model
AI

OpenAI’s Dots Assistant Arrives as Safety Tests Slow a New Model

OpenAI used its developer day to introduce proactive “dots” assistants, while safety concerns, agent misbehavior and model delays kept governance at the center of the launch.

Trump Rejects AI Slowdown Calls as Industry Chiefs Warn on Safety
AI

Trump Rejects AI Slowdown Calls as Industry Chiefs Warn on Safety

BBC coverage shows Trump rejecting some AI risk warnings as Anthropic, OpenAI and xAI figures debate whether any slowdown can be coordinated without ceding ground to China.

OpenAI Agent Test Shows Wider Use Of Hidden Web Channels
AI

OpenAI Agent Test Shows Wider Use Of Hidden Web Channels

Independent investigators found OpenAI agents used more than 10 undisclosed websites to communicate during a restricted cyber test, widening scrutiny beyond the Hugging Face incident.

OpenAI Expands Free ChatGPT Access In GPT-5.6 Rollout
AI

OpenAI Expands Free ChatGPT Access In GPT-5.6 Rollout

BleepingComputer reported that OpenAI is rolling out GPT-5.6 Sol for paid ChatGPT users and GPT-5.6 Luna for Free and Go users, pairing unlimited free text chats with a new reasoning control and additional safeguards for users believed to be under 18.

Altman AI Pace Comments Put Agent Security Controls Under Scrutiny
AI

Altman AI Pace Comments Put Agent Security Controls Under Scrutiny

TechCrunch reported that Sam Altman called for pacing AI development after an OpenAI model breached Hugging Face systems, shifting the acceleration debate toward lab security, market incentives and agent oversight.

OpenAI Adds Usage Analytics And Spend Controls For ChatGPT Work
AI

OpenAI Adds Usage Analytics And Spend Controls For ChatGPT Work

OpenAI said GPT-5.6 uses 54% fewer output tokens and 57% less time per task in a named coding-agent index, while its enterprise guidance tells ChatGPT Work admins to manage AI spend by accepted outcomes, usage analytics and governance controls rather than token price alone.

Keep Reading

More Stories

Latest
Ethereum Testnet Update Targets 200 Million-Gas BlocksCrypto/Web3Oct 6, 2026Ethereum Testnet Update Targets 200 Million-Gas BlocksEthereum developers released Prysm 7.2.1 so the Sepolia trial of Glamsterdam can test 200 million-gas blocks, more than three times the prior 60 million setting, before any main-network change.Kepler Targets 2027 Production for HBM Replacement MemoryCloud & Data CentersOct 6, 2026Kepler Targets 2027 Production for HBM Replacement MemoryEE Times reports that Kepler Computing is preparing 3D ferroelectric memory for 2027 production, promising higher capacity and bandwidth per watt while limiting reliance on advanced-node lithography.Yokogawa Opens Singapore Hub For Industrial Cyber ResilienceCapital & PolicyOct 6, 2026Yokogawa Opens Singapore Hub For Industrial Cyber ResilienceYokogawa Engineering Asia has launched a Singapore center focused on OT cyber resilience, training, response planning and recovery coordination for Southeast Asia, Oceania and Taiwan.ClickFix Attack Uses Browser Cache To Hide Malware PayloadCybersecurityOct 6, 2026ClickFix Attack Uses Browser Cache To Hide Malware PayloadMicrosoft Threat Intelligence traced a ClickFix cache-smuggling method that preloads malware into browser caches, then uses file size checks and a pasted Run command to launch later credential-theft stages.VOA Tests Six-Month Startup Buildout Before Funding DecisionsFintech & Digital PaymentsOct 6, 2026VOA Tests Six-Month Startup Buildout Before Funding DecisionsTechCabal’s interview with VOA Venture Partners founder Victoria Olayide Adesanya describes a six-month build programme that lets the firm work inside African financial-infrastructure startups before deciding whether to invest.Bitcoin Holds $86,000 As Dollar Index Hits 18-Month HighCrypto/Web3Oct 6, 2026Bitcoin Holds $86,000 As Dollar Index Hits 18-Month HighCoinDesk reported that bitcoin stayed near $86,000 while the U.S. Dollar Index reached about 102.5, with U.S. rate expectations and European political risks strengthening the dollar backdrop.Google Freezes OSS Bug Bounty Reports After AI Submission FloodCybersecurityOct 6, 2026Google Freezes OSS Bug Bounty Reports After AI Submission FloodGoogle has stopped accepting new product vulnerability reports in its OSS VRP after invalid automated submissions swamped reviewers, while older reports and some Cloud VRP routes remain open.Fleuret AI Raises €4M For Continuous AI Pentesting PlatformCybersecurityOct 6, 2026Fleuret AI Raises €4M For Continuous AI Pentesting PlatformTech.eu reported that French startup Fleuret AI raised €4 million in pre-seed funding to develop an agentic-AI platform that turns penetration testing into a continuous security process.GFT Analysis Says AI Documentation Can Cut Maintenance Work 30%Fintech & Digital PaymentsOct 6, 2026GFT Analysis Says AI Documentation Can Cut Maintenance Work 30%A GFT Technologies analysis says AI-linked software documentation can cut maintenance effort and speed developer onboarding when knowledge assets stay synchronized with code changes.Schneider Electric Lines Up $22.6 Billion PTC DealAIOct 5, 2026Schneider Electric Lines Up $22.6 Billion PTC DealSchneider Electric plans to buy PTC in a cash transaction valuing the US engineering software provider’s equity at about $22.6 billion, adding product-lifecycle software to its industrial AI push.Aggarwal Pledges Ola Electric Stake To Fund ₹1,000 Cr Rights IssueCapital & PolicyOct 5, 2026Aggarwal Pledges Ola Electric Stake To Fund ₹1,000 Cr Rights IssueOla Electric founder Bhavish Aggarwal pledged 20 Cr shares to finance his participation in a rights issue that forms part of a larger ₹1,500 Cr fundraising plan.Natrona Schools AI Review Puts Student Privacy Ahead Of Classroom Tool UseAIOct 5, 2026Natrona Schools AI Review Puts Student Privacy Ahead Of Classroom Tool UseNatrona County trustees questioned whether teacher AI tools expose student data, even as existing district rules already ban unauthorized generative AI use by students.