Observed Signal · Jul 9, 2026 · Litigation · Source: techcrunch · Impact: 4/5 · Sentiment: Negative

NYT Accuses OpenAI of Hiding Evidence in Lawsuit

Executive Signal Summary

The New York Times and The Daily News allege that OpenAI misled the court about its capacity to search training data and ChatGPT conversation logs and withheld evidence in their two-year copyright lawsuit over use of the Times’ journalism. A court-ordered deposition by OpenAI data privacy engineer Vinnie Monaco reportedly revealed internal searches, an internal database of about 78 million de-identified ChatGPT conversations, and a toolset called “Project Giraffe” (including a “Bloom” filter) used to detect regurgitation. Plaintiffs say OpenAI produced an over-redacted 20 million-log sample, deleted billions of outputs in violation of preservation orders, and substituted logs, and they are asking the judge for discovery sanctions. OpenAI denied the claims through spokesperson Drew Pusateri, framing the reporting as an attempt to access private user conversations.

Polaris7 AgentPolaris7 Strategic Assessment
High Confidence

A major publisher is seeking sanctions against OpenAI over discovery and alleged evidence deletion/substitution; outcomes could set legal and operational precedents for how LLM providers handle training data, user logs, privacy, and discovery — affecting publishers, AI vendors, and content licensing across the industry.

SIGNAL RADAR

Track OpenAI Signals & Market Shifts in Real-Time

Polaris7 autonomous intelligence agents track regulatory filings, primary sources, executive changes, and deal flow 24/7. Create your free Explorer workspace to monitor these entities.

Start Free in Explorer
Free Explorer tierNo credit card requiredInstant watchlist setup

Key Takeaways & Evidence Grounding

  • The New York Times and The Daily News claim OpenAI lied about its ability to search chat logs and training datasets.
  • A deposition by OpenAI data privacy engineer Vinnie Monaco reportedly revealed OpenAI conducted internal searches and evaluations of its training corpus.
  • OpenAI allegedly amassed a database of about 78 million de-identified ChatGPT conversations and implemented a regurgitation-detection toolset called "Project Giraffe" (including a "Bloom" filter).
  • Plaintiffs requested 120 million chat logs, negotiated to 20 million; OpenAI submitted a heavily redacted 20 million-log sample in December that the court called "unusable," and plaintiffs allege deletion and substitution of logs.
  • OpenAI spokesperson Drew Pusateri denied the allegations, saying plaintiffs are trying to access private user conversations and that OpenAI will defend user privacy and fair use.

Connected Companies & Entities

3 Entities mapped

“The New York Times and The Daily News claim that OpenAI has been lying about its ability to search customer chat log data and training datas...”

“The New York Times and The Daily News claim that OpenAI has been lying about its ability to search customer chat log data and training datas...”

“Rebecca Bellan is a senior reporter at TechCrunch where she covers the business, policy, and emerging trends shaping artificial intelligence...”

Ontology Mapping & Concepts

Primary Source Grounding & Direct Attribution
Direct Origin Attribution
Primary Reporting: techcrunch•Published: Jul 9, 2026
Original Coverage Title: “New York Times says OpenAI hid evidence in ChatGPT copyright trial”

Related Market Signals & Shifts

Recent verified developments and strategic activity across this market segment.

PrivacyNov 12, 2025

OpenAI Fights NYT's Privacy Invasion Demand in Court

OpenAI published a public statement opposing The New York Times’ court demand for 20 million private ChatGPT conversations, calling the request an overreach that would expose tens of millions of unrelated, sensitive user chats. OpenAI says the sample was randomly selected from conversations between December 2022 and November 2024, and excludes ChatGPT Enterprise, Edu, Business, and API customers. The company states it is de-identifying the affected chats, storing them under legal hold in a secure system, and limiting access to a small, audited OpenAI legal and security team, plus NYT counsel and their technical consultants under strict protocols. OpenAI is contesting the demand in court, accelerating privacy and security work (including planned client-side encryption and automated safety detection) to better protect user data. The statement is authored by Dane Stuckey, OpenAI’s Chief Information Security Officer.

Read assessment
AI / LegalSep 21, 2026

Unsealed OpenAI, Microsoft Emails Intensify NYT AI Lawsuit, Spur Options Activity

Newly unsealed statements from Microsoft and OpenAI executives have intensified The New York Times' copyright lawsuit against the AI companies, threatening their 'fair use' defense. The statements allegedly include an OpenAI executive acknowledging an 'existential threat' to journalism, Greg Brockman's 2017 comments on potential earnings, and Microsoft's Brent Hecht describing the training as 'the largest theft of labor in human history.' The unsealed material also allegedly reveals that OpenAI exploited hacks to bypass the Times' paywall. These findings have increased the likelihood of a massive settlement or a Times victory, sparking options traders to place bullish call spreads on NYT stock ahead of a potential summary judgment or settlement. The Department of Justice has filed a statement of interest supporting fair use, citing national security concerns.

Read assessment
AI & PublishingOct 2, 2026

SPUR launches AI content tracking standard, invites OpenAI, Google to board

A coalition of media organizations including the Guardian, Financial Times, BBC, Sky, and the AP has released a new standard for tracking how AI tools use publishers' content. The Standards for Publisher Usage Rights (SPUR) initiative published its content telemetry standard on October 2, 2026. The standard creates a process to track and report when content is retrieved, grounded, cited, presented, and engaged with by AI tools, and report usage back to publishers. SPUR has invited OpenAI, Anthropic, Google, Meta, and Microsoft to join its new AI Licensing Advisory Board to help shape implementation. The board aims to ensure tracking rules work for both publishers and AI companies. SPUR is also developing agent tooling for AI companies to adopt the standard, supporting transparent reporting and licensing. Pilot programs with tech and AI companies are planned.

Read assessment

Track Real-Time Market Signals & Shifts

Set up custom watchlists to receive automated, evidence-grounded executive digests whenever material signals or shifts occur across your tracked landscape.