Ground Truth

One claim checked against evidence

Agents & softwareGround Truth

JSON Schema makes agent tools production-ready

NIST AI 800-5 finds commenters agree agent threats are novel and existing controls need adaptation. Schema syntax is not the bar.

TLDR

Claim: valid JSON against a schema is enough to ship side-effecting agent tools. NIST Trustworthy and Responsible AI 800-5 (May 18, 2026) summarizes CAISI's agent-security RFI: commenters widely agreed agents present novel threats and fundamental cybersecurity practices require adaptation. The January 2026 RFI foregrounded indirect prompt injection and misaligned objectives. Verdict: overstated for workflows with side effects.

ModelsGround Truth

One million tokens fixes long-document recall

DeepSeek-V4 reports MRCR 1M MMR 83.5 and CorpusQA 1M ACC 62.0 at Pro Max. Strong on paper. Not a buyer corpus guarantee.

TLDR

Claim: stuffing full documents into a 1M-token window reliably surfaces buried constraints without retrieval. DeepSeek-V4 (arXiv:2606.19348, April 2026) scores MRCR 1M MMR 83.5 and CorpusQA 1M ACC 62.0 at Pro Max on HuggingFace, but Non-Think mode is 44.7/35.6, Claude Opus 4.6 leads both benchmarks, and the paper notes degradation beyond 128K. Verdict: overstated for buyer corpora without placement tests on the shipped checkpoint.

Work & moneyGround Truth

Enterprise AI is paying off at scale

McKinsey's 2026 AI Trust survey finds average RAI maturity 2.3 and only ~30% at level 3+ on governance. Not an EBIT proof point.

TLDR

Claim: enterprise AI is delivering material EBIT impact at scale across the market. McKinsey's State of AI Trust in 2026 (March 25, 2026) surveyed ~500 organizations: average RAI maturity rose to 2.3 from 2.0, but only about 30% reach level 3+ on strategy, governance, and agentic AI controls. Security and risk, not regulation, top barriers to scaling agents. Verdict: overstated as a market-wide EBIT claim without firm-level evidence.