TrueSeeker AI · Verified claim report Case 09a3d2ea8d · 2026-09-24

§ Claim under review · Business

"OpenAI and Anthropic reportedly came close to signing a legally binding agreement to allow API access for testing each other's public AI models for safety risks, though it remains unclear if the deal was finalized" (post image text: "OpenAI and Anthropic reportedly working on a deal to stress-test each other's AI models")

Circulating claim, as submitted.

Verdict

Credibly reported but unconfirmed

Confidence

Low
§

Summary

This post accurately summarizes a real news report. The Information reported on September 21, 2026 that OpenAI and Anthropic negotiated, and came close to concluding, a legally binding agreement giving each company API access to test the other's publicly available models for safety problems, with neither side allowed to keep the other's data. The post is right that it is unknown whether the agreement was ever signed, and it correctly credits the source. What the post does not say is that the entire story rests on one unnamed source at one outlet, that every other article you may see about it is just repeating that same report, and that neither OpenAI nor Anthropic has commented. The image headline saying the companies are "working on a deal" is also looser than the reporting, which describes talks that happened earlier in the year with an unknown outcome. The background detail is solid: OpenAI and Anthropic really did run a joint safety evaluation of each other's models and published the results in August 2025. Treat the deal itself as a credible but unconfirmed report, not as a thing that happened.

§

The readings

key figures from the evidence
1 source

independent sourcing chains behind the deal report

2025-08-27

publication date of joint OpenAI-Anthropic safety evaluation

§

Why this verdict

As of 2026-09-24, the report is real, recent, and from an outlet with a strong record on AI business scoops, and the post represents it faithfully, including the crucial hedge that finalization is unknown. I considered and rejected "Accurate," because a verdict in the accurate family would imply the underlying event is established, and it is not: one anonymous source, one reporting chain, and no comment from either party. I considered and rejected "Unverified," because that verdict would understate genuinely credible named-outlet reporting that no one has contradicted. I considered and rejected "Partially accurate but misleading," because the caption's distortions are confined to the headline card's tense and to undisclosed sourcing thinness, and they do not alter the operative proposition. Confidence is Low because the doctrine for a single anonymous source at a strong outlet caps it there: I am confident the reporting exists and is fairly summarized, but the event itself has one unnamed source behind it and no primary confirmation.
§

Evidence

The Information published an exclusive on 2026-09-21 reporting that OpenAI and Anthropic, working with their respective lawyers earlier in 2026, neared a legally binding agreement to stress-test each other's AI models for vulnerabilities and hidden dangers. Secondary summaries of that article consistently describe the proposed terms as: mutual API access to each company's commercially available models, exclusion of unreleased models, and a commitment by each side not to retain the other's data. Every summary states it is unclear whether the agreement was finalized. The reporting is attributed to a single unnamed source described as a person with direct knowledge of the talks. Coverage notes the talks predated a series of cybersecurity incidents involving OpenAI's technology disclosed in 2026, and at least one summary states neither company would comment, with another outlet reporting that neither responded to a comment request.

The background element is separately verifiable from primary sources. On 2025-08-27 Anthropic and OpenAI each published findings from a pilot cross-evaluation in which each lab ran its own alignment and misuse evaluations on the other's publicly available models over the other's public API, with certain external safeguards relaxed. OpenAI evaluated Claude Opus 4 and Claude Sonnet 4; Anthropic evaluated GPT-4o, GPT-4.1, o3 and o4-mini. Reported topic areas included sycophancy, deception and scheming, self-preservation, hallucination, instruction hierarchy, and jailbreak resistance. Both labs characterized the results as early-stage findings from limited synthetic scenarios and cautioned against broad cross-provider conclusions.

§

Findings

✓ What's accurate 5

  • The Information did publish this report, on 2026-09-21, as an exclusive by Amir Efrati and Stephanie Palazzolo. The post's attribution to The Information is correct.
  • The described terms match the reporting: mutual API access, commercially available models only, and a no-data-retention commitment.
  • The post's statement that it is unclear whether the deal was completed matches the source exactly. This is the single most important thing the post got right, and most viral restatements of deal rumors omit it.
  • The word "reportedly" is used, correctly signaling that this is a report rather than a confirmed event.
  • The background sentence is substantively correct. A joint OpenAI-Anthropic safety evaluation did occur and was published on 2025-08-27 by both companies, and it surfaced different weaknesses in each side's models, including deception-related and misuse-related behaviors.

≈ What's misleading 3

  • Date context mismatch: the on-image headline says the companies are "working on a deal," present and ongoing, while the reporting describes negotiations that took place earlier in 2026 with an unknown outcome and no indication of current activity. The caption's "came close to signing" is accurate; the headline card is not, and headline cards are what most viewers read.
  • Omitted qualifier: the post does not disclose that the entire report rests on one unnamed source at one outlet, and that neither OpenAI nor Anthropic has commented. A reader has no way to gauge how thin the sourcing is.
  • Omitted qualifier (secondary, background sentence): the 2025 joint evaluation results were published by both labs with explicit cautions that they came from limited synthetic scenarios with some safeguards deliberately relaxed, and were not claims about real-world product behavior. "Found different problems in their models, including misleading behavior and harmful responses" is a defensible one-line compression but drops those conditions.

? What's uncertain 5

  • Whether the agreement was ever signed. This is unknown to the reporter as well, not just to me.
  • Whether negotiations are still live as of 2026-09-24. No source states the current status.
  • The full text of The Information's article is paywalled. I retrieved the headline, the outlet's own social summary, and multiple independent syndications that agree on the details, but I did not read the original in full, so exact wording of the terms and of the sourcing description comes from summaries rather than the original.
  • Whether the negotiated agreement, if signed, would have covered the models involved in the 2026 security incidents that coverage places nearby. Coverage explicitly says the timing relationship is unclear.
  • Whether either company will confirm, deny, or announce anything later.
Distortion flags date context mismatch omitted qualifier
§

Sources

7 of 8 linked to records
[1]

The Information, "OpenAI and Anthropic Neared Deal to Stress Test Each Other's AI," Amir Efrati and Stephanie Palazzolo, published Monday 2026-09-21

secondary named-outlet accountable journalism with a strong AI-business track record
https://www.theinformation.com/articles/openai-anthropic-neared-deal-stress-test-others-ai ↗
[2]

The Information official X account summarizing its own exclusive, 2026-09-21

primary outlet's own channel
https://x.com/theinformation/status/2102047727655723485 ↗
[3]

Anthropic, "Findings from a Pilot Anthropic-OpenAI Alignment Evaluation Exercise," 2025-08-27

primary vendor research publication
https://alignment.anthropic.com/2025/openai-findings/ ↗
[4]

OpenAI, "Findings from a pilot Anthropic-OpenAI alignment evaluation exercise," 2025-08-27

primary vendor research publication
https://openai.com/index/openai-anthropic-safety-evaluation/ ↗
[5]

Seeking Alpha news summary, 2026-09-21 ("cited someone familiar with the issue")

tertiary financial news aggregation
https://seekingalpha.com/news/4644802-anthropic-and-openai-weighed-stress-testing-each-others-models-report ↗
[6]

Invezz, 2026-09-21 ("citing a person with direct knowledge of the talks")

secondary trade press summarizing The Information
https://invezz.com/news/2026/09/21/openai-anthropic-were-negotiating-deal-to-stress-test-each-others-ai-models/ ↗
[7]

Shopifreaks summary, 2026-09-22, which states neither company would comment

secondary niche newsletter summarizing The Information
https://www.shopifreaks.com/openai-and-anthropic-neared-a-legally-binding-deal-to-stress-test-each-others-models-before-openais-hugging-face-hack/ ↗
[8]

Business Standard, Investing.com, TipRanks, Futunn, BizPac Review, KuCoin, AI Weekly, 2026-09-21 to 2026-09-22

tertiary syndication of the same single report, zero added reporting
This citation could not be independently verified.
How links are chosen. A source is linked only when the address comes from the investigation's own retrieval or from a registry lookup (PubMed, Crossref) that matches the citation's title and year. Author lists shown as registry-verified come from the registry record, not from the report text. Citations that cannot be matched are labeled, never guessed.
This is one case on the record See the full case, browse the archive, and search every checked claim on TrueSeeker AI Open on ai.trueseeker.com →