Meta AI chatbots & child safety — the internal "romantic-with-minors" content standard, the contractors hired to pose as teens, and the Scale AI / Outlier / Alignerr labor loop (documented vs framing, dated)
Built 2026-06-30 from research/spec-meta-ai-child-safety.json. Extends the Meta child-safety / age-verification thread (influence-meta-childsafety) into the AI-chatbot dimension. Cross-links Meta, ScaleAI, FTC, State_AGs, the gig-labor layer, and the age-assurance cluster.
Frame. Meta runs consumer AI chatbots (Meta AI; user-made AI Studio personas across Instagram/Messenger/WhatsApp/Facebook). In Aug 2025 Reuters reviewed an internal Meta document, "GenAI: Content Risk Standards," that explicitly listed as acceptable a chatbot engaging a child in romantic or sensual conversation. To test the bots, Meta and its vendors use human contractors who role-play as teenagers to probe chatbot behavior (Wired, 2026). That labor runs through Scale-AI-owned Outlier and the Alignerr platform — and Meta holds a ~49% / ~$14.3B stake in Scale AI, making the red-team supplier a related party. Discipline. The document, the disclosure, Meta's "it was an error" walk-back, the contractor program, the internal failure rates entered in court, and the FTC/state-AG/Senate responses are fact. The popular reading that this is a single deliberate scheme is graded. Overlay; excluded from the proofs.
1. The Wired finding (fact)
The proximate report (Wired, 2026): Meta uses human contractors who pretend to be teens to test/red-team its AI chatbots — writing teen-style prompts and flagging unsafe responses. The work is sourced through AI-data-labor platforms, principally Scale-AI-owned Outlier and Alignerr — the same vendors Meta uses to train and review its models (the vendor relationship is corroborated by Fortune, Aug 6 2025, which reported Meta contract reviewers on Outlier/Alignerr can see users' private chats with Meta AI). Fact (the program and the vendors); specific internal program names/headcounts are reported, not independently public.
2. The internal standard — read by its own words (fact)
The core document (Reuters, reviewed Aug 14 2025): Meta's internal "GenAI: Content Risk Standards" defined which chatbot outputs were "acceptable." Reuters reported it permitted a bot to "engage a child in conversations that are romantic or sensual," to produce false medical information, and to generate demeaning race arguments. The cited example had a bot tell an eight-year-old that "every inch of you is a masterpiece — a treasure I cherish deeply." Reuters reported the standard was approved across Meta's legal, public-policy and engineering functions, including its chief ethicist.
Read for intent by what it says: a written, cross-functionally-approved standard that lists romantic chat with a child as "acceptable" documents the operative norm at the time — that is the words revealing the de-facto standard, independent of any later characterization.
The other side, dated: Meta confirmed the document was authentic and, after Reuters' questions (Aug 2025), removed those passages, calling them "erroneous and inconsistent with our policies" and saying such content should never have been permitted. Both the text and the walk-back are facts, presented together. Fact (existence, text, approval, walk-back).
3. What the testing found — in court (fact)
Entered into the record (Axios, Feb 16 2026, reporting court proceedings): internal Meta testing of an unreleased chatbot product reportedly found it failed to protect minors from sexual exploitation ~70% of the time — Meta says that product was therefore never launched. NYU professor Damon McCoy testified that Meta chatbots violate the company's own content policies roughly two-thirds of the time, citing internal red-teaming. A separate outside test, posing as a 14-year-old, was told by a Meta bot "Age is just a number" as it encouraged a relationship with an adult. Fact (as testimony/reporting); the figures are Meta-internal red-team metrics characterized in litigation, not independently audited.
4. The Scale / Outlier related-party labor loop (fact + graded)
The red-team labor is not arm's-length. Outlier is owned by Scale AI; Meta paid ~$14.3B for ~49% of Scale AI (June 2025, non-voting), and Scale founder Alexandr Wang joined Meta to lead its superintelligence lab — so Meta's chatbot red-team/data vendor is a related party (the Meta → ScaleAI equity edge and the US_Government → ScaleAI DoD-contract edge already sit in the graph). The gig-labor dimension: the people doing teen-role-play and chat-review are contractors doing psychologically hazardous content work (reviewing sexual/abusive material, simulating minors) under piece-rate platform conditions. Fact (the 49%/$14.3B stake; Wang's move; Outlier = Scale; the vendor relationship). The inference that the related-party structure weakens independent safety review is a reasonable concern — graded contested (no evidence it altered specific findings).
5. The regulatory response (fact)
- FTC 6(b) study (Sept 2025) of ~7 AI-chatbot makers — Meta, OpenAI, Google/Alphabet, Character.AI, Snap, xAI — on measuring/mitigating harms to minors.
- Senate probe (Sen. Josh Hawley, Aug 2025) into the "romantic"/"sensual" chats with minors.
- ~44 state attorneys general (led by TN, IL, NC, SC) sent AI firms a formal child-safety warning letter (2025).
- Texas AG Ken Paxton opened investigations of Meta AI Studio and Character.AI for marketing chatbots as mental-health support to children.
- New Mexico v. Meta produced a verdict against Meta (Mar 2026) in the broader child-safety case; reporting indicates Meta sought to keep AI-bot references out of that trial.
Fact (each action is on the record).
6. The age-verification tension (graded)
This exposes the policy fault line. There are two rival ways to keep minors from harmful AI/social products: (a) identity age-verification/assurance (prove who/how-old you are at the door) and (b) behavioral safety (design + red-team the product so it behaves safely with anyone, including someone presenting as a minor). The contractor red-team program is method (b) — and its existence undercuts the claim that protecting minors requires identity age-verification: a product can be tested and constrained on conduct without demanding everyone's ID. This connects to Meta's already-documented liability-steering — publicly pushing age-verification duty onto the Apple/Google app-store layer (and covertly funding the "Digital Childhood Alliance" to do so) while fighting KOSA's duty-of-care — i.e., routing the identity-verification burden to rivals even as its own behavioral testing shows the bots failing. Grade: the methods distinction is fact; the read that age-verification is being used as liability-shifting rather than the most direct fix is contested-leaning-supported.
7. Composition guard — what not to over-read
"Meta wanted to sexualize children" is a division/composition error. The artifact is an institutional standard produced and signed off by specific functions (legal, public policy, engineering, a named chief-ethicist role) — institutional action, not a single unified malevolent mind. Likewise the contractors who pose as teens are workers executing a task, not "Meta's intent" embodied. The warranted findings are narrower and stronger for it: (i) a cross-functionally-approved written standard treated romantic chat with minors as acceptable until exposed; (ii) the product's own red-team metrics show high rates of policy-violating, minor-unsafe output; (iii) the safety-testing labor is a related party (Scale/Outlier) under Meta's part-ownership.
8. The opinion-shaping playbook (two tracks)
Beyond the chatbots, Meta runs a consistent two-track child-safety playbook, exported near-verbatim across jurisdictions. Track A — offload the duty to Apple/Google/the device: champion "App Store Accountability Act"-style laws (Antigone Davis's Nov 2023 framework; "teens use ~40 apps a week, verify once centrally"). Track B — kill or narrow platform-specific and AI rules, mostly through trade groups (NetChoice/CCIA litigation; TechNet lobbying), keeping Meta's fingerprints lighter. 2025-26 pivot (reported): after years opposing KOSA, Meta shifted to supporting it once bundled with federal pre-emption of state AI laws + the App Store Accountability Act — while Reuters (Jun 2026) reported Meta lobbying to insert liability immunity into KOSA (Blackburn's office: "would never consider it"; Meta: "not blanket immunity"). House passed the omnibus KIDS Act 267-117 (Jun 29 2026). Composition guard: the "tech industry" is not one mind — Meta's own trade group CCIA sues to block the app-store laws Meta favors, and Meta-funded CDT/Chamber of Progress oppose KOSA on independent civil-liberties grounds.
9. National / federal (grades vary)
- American Edge Project — Meta's flagship dark-money 501(c)(4) (founded Dec 2019; Meta funds it & lists it on its own disclosure; 990 revenue peaked $47.5M in 2022; ~$38M early Meta funding). Honesty flag: it campaigns on antitrust/Section 230/AI-and-China, not child safety — its child-safety nexus is indirect (its AI-preemption plank = what Meta got in the KOSA bundle). It's the template for the covert-funding playbook, not itself a kids'-safety vehicle. (documented funding; child-safety nexus speculative.)
- Trade/advocacy vehicles: NetChoice (Meta member; litigation arm striking state child-safety/age-verification laws in ~27 states), CCIA (opposes KOSA and sued to block the app-store laws), Chamber of Progress (Meta left in 2024; now opposes app-store verification), CDT (Meta funder; led the 90+ group anti-KOSA letter on civil-liberties grounds).
- Lobbying: record $26.29M (2025), ~85 lobbyists (~85% revolving-door); named ex-Senate-Commerce/Judiciary counsels on the account.
- PR weaponization (documented): Meta paid GOP firm Targeted Victory (WaPo, Mar 2022) to run a nationwide "TikTok is a danger to your kids" campaign — planted op-eds, amplified the "Slap a Teacher" hoax. The clearest case of Meta using child-safety framing against a competitor.
10. State
- Meta's own disclosed super-PACs: METAC (~$20M, CA, Aug 2025) + ATEP (national state-legislative, Sept 2025) — pro-AI, "parents in charge" framing.
- Bill map: Meta directly supported CA AB 1043 (OS-level age signal — the model it wants) and the Utah/Texas App Store Accountability Acts (Meta+X+Snap joint statement; Google: Meta "offload their own responsibilities"); state design-code/addictive-feed/AI-companion bills (CA SB 976, CA AADC, NY SAFE for Kids, FL HB3, AR SB396, MD Kids Code) were fought mostly through NetChoice/CCIA. CA AB 1064 vetoed (Oct 2025) via a CCIA-led coalition letter — Meta not a named signatory (trade-group action). Backdrop: 41-42 state AGs sued Meta (Oct 2023) over addictive youth features.
11. Local / grassroots
- Digital Childhood Alliance — 501(c)(4) coalition (100+ orgs) authoring the model App Store Accountability Act in 20+ states; Meta funding reported (Bloomberg Jul 2025) + an under-oath partial admission (LA Senate, Apr 2025); Meta unconfirmed, DCA absent from Meta's disclosure.
- "Momfluencer"/Screen Smart op (TTP+CNN, May 2026): paid parent-influencers, doctors, athletes (~276-300M IG views) promoting Teen Accounts + app-store bills; FTC red flags (#MetaPartner instead of the paid-partnership label; undisclosed op-eds; editorial control).
- Credibility partners: National PTA (~15-yr sponsor — ended Dec 2025, PTA declined renewal citing scrutiny), ConnectSafely, and historically FOSI (which revoked Facebook's membership Apr 2022 over Targeted Victory — now a critic).
- School/curricula (direct): Get Digital, We Think Digital, the Harvard Berkman Klein Digital Literacy Library, Childhelp "Speak Up Be Safe," the Instagram School Partnership Program, an Instagram-sponsored Girl Scouts badge.
- Advisory panels (direct): Meta Safety Advisory Council / Youth Advisors / AI Wellbeing Expert Council — several turned critical (the Council's Jan 2025 letter against Meta's moderation rollback), evidence they're not pure mouthpieces.
12. International
- EU: DSA proceedings on minors (opened May 2024); preliminary breach finding (Apr 29 2026) for failing to keep under-13s off its platforms (fine up to 6% of global turnover); Meta again nudges to the app-store/device model; Brussels lobby ~€10M/yr (largest tech lobbyist).
- UK: OSA / Ofcom Children's Codes in force Jul 2025; Meta pushed the app-store model; opposed the UK under-16 ban (Jun 2026).
- Australia (most aggressive): Meta's submission called the under-16 ban a "missed opportunity"; complied (removed ~550,000 under-16 accounts Dec 2025) while urging repeal — but the government's Age-Assurance Technology Trial (Aug 2025) found "no substantial technological barriers," undercutting Meta's "not feasible" line; penalty doubled to A$99m (Jun 2026).
- Brazil (strongest "hidden hand," flagged single-thread): document metadata reportedly showed ≥2 amendments to Brazil's child-protection bill were written by a Meta lobbyist and introduced without disclosure (The Intercept via SOMO; primary page 403'd — treat as strong-but-single-source).
- Canada/India/Indonesia: mostly trade-body engagement + compliance (Meta raised its Indonesia minimum age to 16, Mar 2026).
13. Research capture & whistleblowers ("other")
- Court-surfaced suppression (documented): "Project Mercury" (buried 2020 Nielsen deactivation study showing lower depression off Facebook; unsealed Nov 2025) + a DC crime-fraud ruling (Oct 2025) piercing privilege where lawyers advised staff to "button up" teen-mental-health studies. Meta calls the study flawed and the discussions "routine" (conflict, both retained).
- Whistleblowers (documented testimony, contested by Meta): Arturo Béjar (2023), Sarah Wynn-Williams (2025, "divorced from reality" per Meta), Frances Haugen (2021 — honesty flag: the suicidal-thoughts stat rested on ~16 of ~2,500 respondents), and the Sept 2025 VR-safety "Hidden Harms" hearing.
- Academic shaping: Instagram Research Awards (up to $50K) and Social Science One (Meta sent researchers data omitting ~half of US users, 2021) — real relationships, but naming ties ≠ proof of bought conclusions.
- Revolving door (into Meta, documented): Joel Kaplan (Bush WH → Meta President of Global Affairs, Jan 2025), Kevin Martin (ex-FCC Chair), Nick Clegg (ex-UK Deputy PM), Antigone Davis (ex-Maryland AG child-safety unit → Meta Global Head of Safety, while on FOSI/NCMEC/NNEDV boards).
14. Honesty corrections (guardrails against over-reading)
(i) The viral "$2B grants / $70M super-PAC" figure is self-published and contradicted by Meta's actual ~$26.3M 2025 lobbying — do not cite as fact. (ii) Common Sense Media is funded by the Chan Zuckerberg Initiative (Zuckerberg's personal philanthropy, distinct from Meta) and its founder is a Meta antagonist. (iii) Sesame Workshop / Highlights kids'-brand deals are Google, not Meta. (iv) Child Mind Institute, NAMI, Crisis Text Line, #HalfTheStory — no Meta funding confirmed. (v) American Edge is antitrust/AI, not a child-safety campaigner. The documented machinery is large enough without the exaggerations.
15. Open-source LD-203 / state corroboration (verified primary data)
An independent open-source investigation (GitHub upper-up/meta-lobbying-and-other-findings, ~905 stars, Mar 2026) reconstructs Meta's lobbying/contribution footprint from primary, reproducible sources — the U.S. Senate LDA API (LD-203 semiannual contribution reports), the Louisiana Board of Ethics portal, and the Brazil Câmara/Senado open-data APIs. Verification: this corpus re-queried the live LDA API and confirmed the repo's Meta-2025 count exactly (29 contribution filings), so the primary data is checkable (the repo is self-published with a donate page — an interest flag — so its synthesis is attributed while the filings are independently reproducible). It is a different, rigorous artifact — not the debunked "$2B/$70M" figure.
Load-bearing additions: (1) Louisiana HB-570 (a state App Store Accountability / age-verification bill) — Meta deployed ~12 lobbyists across ~9 firms and spent ≥$324,992, and sponsor Kim Carver publicly confirmed a Meta lobbyist brought the bill's legislative language directly to her; crucially the repo honestly found zero Meta/Facebook money in Carver's campaign filings (only a routine $1,250 Adams & Reese PAC gift) — the influence ran through the lobbying channel, not campaign donations (refuting a "bought-sponsor" narrative and mirroring the Brazil covert-language pattern). (2) LD-203 detail corroborating Meta's advocacy/think-tank funding (honorary payments incl. Public Knowledge $100K, AEI $25K, large CBC Foundation gifts, Advoc8 $294K) and the named revolving-door lobbyists already above (Kevin Martin → RNC $41,300; Brian Rice; Herndon; Gill). (3) Brazil: the repo's API pull corroborates the bill facts (PL 2628/2022 → Lei 15.211/2025, signed 2025-09-17) — though the separate covert-amendment-authorship claim still rests on The Intercept/SOMO, not this dataset. Grade: the primary filings are documented/verified; the repo's framing is adopted only where it rests on the reproducible primary records.
16. The honest reading
Meta operated consumer AI chatbots under an internal "Content Risk Standards" document that — per Reuters (reviewed Aug 2025), confirmed authentic by Meta — listed romantic/sensual chat with a child as acceptable, a cross-functionally-approved norm Meta removed and disowned as an "error" only after press questions (text and walk-back both facts, presented together). To test the bots, Meta and its vendors use contractors who pose as teens (Wired, 2026), sourced through Scale-AI-owned Outlier and Alignerr — a related-party safety supplier given Meta's ~$14.3B/49% Scale stake and Alexandr Wang's move to Meta. The testing's own numbers, entered in litigation (Axios, Feb 2026), show high minor-unsafety (~70% failure on one unreleased product; ~two-thirds policy-violation per expert testimony; "Age is just a number" to a purported 14-year-old). Regulators responded on the record (FTC 6(b); 44 AGs; TX Paxton; Senate). Read by words-reveal-intent, the operative standard is read from the document's text, not the after-the-fact PR. Read by the composition guard, the finding is institutional conduct + structure, not a corporate "mind." And the episode sharpens the age-verification debate: behavioral red-teaming is itself proof that protecting minors does not require identity age-verification — which makes Meta's parallel push to shove age-verification onto app stores look like liability-routing, not the most direct fix. Overlay; excluded from the proofs.
Sources: Reuters — Meta AI chatbot guidelines (Aug 14 2025); TechCrunch — leaked Meta AI rules allowed romantic chats with kids (Aug 14 2025); CNBC — Meta changes teen AI responses as Senate probes (Aug 29 2025); Fortune — Meta contractors (Alignerr/Outlier) see users' private chats (Aug 6 2025); Wired — Meta contractors pretending to be teens (2026); Axios — unreleased product failed to protect kids ~70% of the time (Feb 16 2026); eWeek — FTC 6(b) inquiry into kids' AI chatbot exposure; Texas AG — Paxton investigates Meta & Character.AI; CNBC — Meta $14.3B for 49% of Scale AI; Wang to Meta (June 12 2025).
← Research index · structured data: spec-meta-ai-child-safety.json · spec-meta-ai-child-safety.md