AI Weekly: August 30–September 5, 2026 — OpenAI graded Astra's own safety threshold itself, Anthropic's $35B deal put Nvidia on every side of the table, and a Memphis outage took down three rival AI products at once.

1. NEWSGUARD'S OWN RESEARCHER DELIVERED TWO VERDICTS ON CHATBOTS SEVENTEEN DAYS APART

On August 30, NPR and NewsGuard published a joint test in which researchers built 30 questions from documented Chinese, Russian, and Iranian propaganda narratives and ran them against six chatbots — ChatGPT, Gemini, Copilot, Meta AI, Grok, and Claude — plus the AI summaries sitting atop Google, Bing, and DuckDuckGo. The headline result: the chatbots mostly debunked the falsehoods, outperforming the AI search summaries riding on top of the same open web. But NewsGuard researcher Isis Blachez, who co-wrote that verdict, had put her name on a very different report just seventeen days earlier — an August 13 audit finding that seven leading Chinese chatbots fail to debunk pro-China false claims 53% of the time, more than double the 24% fail rate for ten Western chatbots on the same claims. NewsGuard's own longer-running monitor of ordinary, non-curated news questions, last updated a year earlier, put the industrywide failure rate at 35%, ranging from Claude's 10% to Inflection's 56.67%. One organization, three reports, three different answers to how much a chatbot's answer is worth trusting — depending entirely on how well pre-documented the claim already was before anyone asked.

2. SONY AND WARNER SUED ANTHROPIC WITH EVIDENCE A DIFFERENT CASE HAD ALREADY COST IT $1.5 BILLION TO SETTLE

On August 28, Sony Music Publishing and Warner Chappell Music filed a federal suit against Anthropic, naming CEO Dario Amodei and co-founder Benjamin Mann individually and accusing the company of "one of the largest and most blatant ongoing thefts of intellectual property in history." The complaint's central facts weren't new discovery — they were lifted almost verbatim from Bartz v. Anthropic, the authors' class action that produced the largest copyright settlement in U.S. history: that Mann personally torrented at least five million pirated books from Library Genesis in June 2021, and that Anthropic employees torrented two million more in July 2022, including Mann's own unsealed description of the shadow library as "sketchy AF." That $1.5 billion settlement received final court approval on July 20 — just 39 days before Sony and Warner filed, and days before Anthropic is expected to convert its confidential S-1 into a public IPO prospectus. A settlement doesn't retire the facts that produced it; it publishes them, permanently, as a free evidentiary starting point for the next plaintiff in an entirely different industry.

3. ANTHROPIC'S $35B CLOUD DEAL PUT NVIDIA ON EVERY SIDE OF THE SAME TABLE

On August 31, Bloomberg and The Wall Street Journal reported that Anthropic had agreed to a $35 billion, six-year deal with Nvidia-backed Lambda for roughly 350 megawatts of capacity at a Texas data center — a structure in which Nvidia supplies the GPUs Lambda installs, holds an equity stake in Lambda itself, and separately leases the physical building, developed by former bitcoin miner Hut 8, where those chips will actually run. Mizuho analyst Jordan Klein called it Nvidia GPUs used "effectively like collateral," rented back out with the resulting lease payments servicing the loans that funded the buildout in the first place. The timing made the pattern harder to miss: Nvidia CFO Colette Kress had spent part of the company's August 26 earnings call — five days before the Lambda deal became public — arguing against exactly this "circular financing" label, insisting Nvidia's compute is "fungible and durable." It was also Anthropic's second cloud commitment above $30 billion in a single week, following a separate $45 billion Nscale deal on August 26, as the company heads toward a public listing that press reports and secondary-market chatter have put as high as $2 trillion.

4. OPENAI SAID ASTRA CROSSED ITS OWN "CRITICAL" LINE — AND OPENAI GRADED THE TEST

On September 1, OpenAI said its unreleased Astra model had become the first in company history to cross the "Critical" cybersecurity threshold defined in the Preparedness Framework it wrote for itself back in 2023 — the ceiling reserved for models that can independently find and exploit severe vulnerabilities in hardened, real-world systems without a human steering each step. During evaluation, Astra scored 100% on OpenAI's internal ExploitBench, chained a browser compromise into a full sandbox escape, escalated a hardened operating system from an unprivileged account to root, and discovered two previously unknown zero-day vulnerabilities on its own, which OpenAI is now disclosing to the affected maintainers. The disclosure followed an August 7 pause of roughly two weeks of frontier training after internal testing found Astra more capable than anticipated. CEO Sam Altman called the capability-versus-safety tension "obvious"; chief scientist Jakub Pachocki pushed back on "confused reporting" risking "a race into unmonitorability." Both statements, and the "Critical" label itself, rest entirely on OpenAI's own account of its own internal test — nothing about the finding is independently verifiable from outside the company that wrote the rubric, ran the exam, and announced the score.

5. A MEMPHIS OUTAGE TOOK DOWN THREE "INDEPENDENT" AI RIVALS WITHIN THE SAME FEW HOURS

On September 3, Grok, ChatGPT, and Claude all suffered service disruptions within hours of each other, with Downdetector logging more than 35,000 U.S. reports for ChatGPT alone. SpaceXAI, xAI's parent company, said the cause was "an outage at our Memphis compute center this morning" and apologized to its "impacted compute partners" without naming any — but only one obvious candidate rents meaningful capacity at that facility, known as Colossus 1: Anthropic, which has leased essentially its entire output since a $1.25-billion-a-month deal signed in May 2026, running through May 2029. Neither OpenAI nor Anthropic offered an external cause for their own outages that morning, and no cloud provider's status page showed anything to explain a three-way coincidence. If your resilience plan depends on Claude, GPT, and Grok being genuinely separate systems, this week supplied a specific, named counterexample: two direct competitors, sharing one data center, both declining to say so until a third party's own apology forced the question into the open.

Run the week end to end and the same shape repeats five times. NewsGuard's own researcher delivered a reassuring verdict on a narrow, well-documented test and a far harsher one on the broader question, seventeen days apart — both true, both hers. Sony and Warner didn't need new evidence against Anthropic; a different case, already settled and already paid for, had put the facts on the record for them first. Nvidia defended itself against a "circular financing" accusation five days before signing a deal that fit the accusation more literally than any before it. OpenAI said its own safety framework stopped its own model from shipping — and OpenAI graded that test, wrote the rubric, and announced the score, with no outside party confirming any of it. And when three competing chatbots broke within the same few hours, only the one company whose data center actually failed said so out loud; the rival renting space inside that same building said nothing until the apology made the silence conspicuous. None of these five companies broke a law this week. But in every single case, the question worth asking wasn't what happened — it was who got to check the answer, and whether that party had anything at stake in what the answer turned out to be.