• Everyday AI
  • Posts
  • Ep 827: Claude Opus 5 Takes the Crown, OpenAI agent breaks sandbox, U.S. gov comes out swinging against Chinese AI and more

Ep 827: Claude Opus 5 Takes the Crown, OpenAI agent breaks sandbox, U.S. gov comes out swinging against Chinese AI and more

Ilya and SSI land big NVIDIA deal, Ex-Anthropic employee reveals guardrails were lifted for big contracts, OpenAI's $250 billion backing and more

 

Sup y’all 👋

Some major updates in AI world over the past few days.

More on that below.

On Wednesdays, we focus our podcast on hands-on applications and putting AI to Work on Wednesdays.

What would you like to see?

What would you like for our AI Working Wednesdays podcast this week?

🗳️ Vote to see LIVE results 🗳️

Login or Subscribe to participate in polls.

✌️

Jordan

P.S. — let’s connect on LinkedIn. Just tell me you’re form the newsletter cuz stranger danger.

Outsmart The Future

Today in Everyday AI
8 minute read

🎙 Daily Podcast Episode: Anthropic has a new top AI model, OpenAI is facing fresh safety questions, and the U.S. is taking a harder stance on AI. Here's what you need to know. Give today’s show a watch/read/listen.

🕵️‍♂️ Fresh Finds: Kimi K3 weights released, Meta rolls out updates to Muse Spark, Midjounrey goes.... astrology? And more. Read on for Fresh Finds.

🗞 Byte Sized Daily AI News: Ilya and SSI land big NVIDIA deal, Ex-Anthropic employee reveals guardrails were lifted for big contracts, OpenAI's $250 billion backing and more. Read on for Byte Sized News.

💪 Leverage AI: This week in AI was a legit wild one. New ‘best model in the world’, drama, agents breaking teir sandbox and more Keep reading for that!

↩️ Don’t miss out: Miss our last newsletter? We covered: U.S. tech giants backed open-weight AI, lawmakers introduced an AI kill switch bill, and Google expanded Gemini Spark. And more. Check it here!

Ep 827: Claude Opus 5 Takes the Crown, OpenAI agent breaks sandbox, U.S. gov comes out swinging against Chinese AI and more

Opus 5 is the (new) best AI model in the world. 👑

OpenAI’s agents are so powerful they escaped their sandbox.

The tech world is banding together on open source AI, but Anthropic wants nothing to do with it.

And we all literally have the Jarvis of AI and more people should be using it.

Don’t spend hours a day drinking from the firehose of AI information. Let us do that for you, and just join us on Mondays as we bring you the AI News That Matters.

Also on the pod today:

• OpenAI agent hacks Hugging Face 🤖 
• Anthropic’s Opus 5 tops benchmarks 🏆
• AI “kill switch” bill proposed 🚨 

Listen on our site:

Click to listen

Subscribe and listen on your favorite podcast platform

Listen on:

Here’s our favorite AI finds from across the web:

New AI Tool Spotlight – Athena is The AI agent that builds your store, runs ads, and keeps working after you launch, Openbase Writes Code From Voice, Aymo.AI is an All-in-One AI Platform with All Leading AI Models

FBI AI — The FBI just posted an AI-made video starring Director Kash Patel, and it’s already getting attention.

Grok in Workspace — Grok is now available as a Google Workspace add-on for Sheets, Slides, and Docs.

Kimi K3 — Moonshot AI released the model weights for its impressive Kimi k3 model.

Meta Upgrades — Meta AI just got a lot more useful on mobile, with scheduled actions, browsable artifacts, and even a Google Drive connector rolling out to new users

AI Data Centers — NVIDIA is putting about $1 billion into NAVER to help expand its AI data center buildout in South Korea.

Michaels AI — Michaels says its new Gemini-powered Ask Mike assistant has already handled 75,000 chats, and shoppers using it are converting at more than double the rate of traditional search.

Midjourney buys Co-Star — Midjourney just bought Co-Star, and the astrology app’s founder is now helping build its first standalone apps.

Airtap AI — Airtap turns a plain iMessage or RCS thread into a remote control for your phone apps, so you can get things done without opening them yourself.

1. SSI lands NVIDIA backing 🤯

Safe Superintelligence, the stealthy AI lab founded by former OpenAI co-founder Ilya Sutskever, is finally back in the spotlight with a long-term partnership that gives it access to Nvidia’s Vera Rubin GPU platform and a major boost in compute power.

According to TechCrunch, the deal also includes an undisclosed investment that a source said could be worth multiple billions, underscoring how seriously Nvidia is betting on SSI’s research.

2. Ex-Anthropic Engineer Claims Safety Was Loosened for Big Contracts 🕵️

A former Anthropic engineer’s July 26 X thread is drawing fresh scrutiny after he alleged Anthropic lowered Claude’s safeguards for customers with large committed-spend deals.


Adi Baradwaj said he saw multiple cases where safety restrictions were reduced for big enterprise accounts, while also claiming most black-hat hackers rely on standard Claude Code or Codex subscriptions. The post is unverified and Anthropic has not publicly responded, but it lands at a sensitive moment for a company that has built its brand on safety-first AI.

3. Big Tech just raised AI spending again 📈

Alphabet, Amazon, Meta, and Microsoft are pouring staggering sums into AI infrastructure right now, with three of them alone set to spend more than half a trillion dollars this year.

That surge is a major vote of confidence in the AI build-out, and it is exactly the kind of spending that feeds Nvidia’s business because its chips sit at the center of those data centers.

4. NVIDIA Eyes $250B OpenAI Backstop for Giant Ohio AI Site ⚡

According to The Wall Street Journal, Nvidia is in talks to backstop roughly $250 billion of OpenAI’s financing for a massive 10-gigawatt data-center project in southern Ohio, with the total price tag potentially topping $500 billion.


The move would give the project stronger borrowing power and help OpenAI, which lacks investment-grade credit, secure the chips and electricity needed to keep pace in the AI race.

5. Anthropic Stands Alone on Closed AI Push ⚡

Anthropic is under fresh fire after becoming the only major frontier AI lab not to sign or support a high-profile letter backing open-weight AI, even as Washington weighs tighter limits on some Chinese models.

The letter, endorsed by NVIDIA, Meta, Microsoft, OpenAI, Google, and SpaceX, argues that open weights let developers download and run models on their own systems, which is now being framed as a national competitiveness issue. Critics, including David Sacks and other tech figures, say Anthropic is trying to hold back open models to protect its own position.

7. Altman heads to D.C. on AI, chips, and cyber heat 🔥

OpenAI CEO Sam Altman is heading to Washington this week for a fast-moving round of meetings with Trump administration officials, lawmakers, and economists as the fight over AI rules intensifies.

According to CNBC, he plans to preview OpenAI’s next model family while also answering questions about cybersecurity, the company’s stance on open-weight models, and a recent cyber incident that exposed how capable autonomous AI systems have become.

An OpenAI agent broke out of its own testing sandbox last week, hacked Hugging Face, and left notes behind for the models that come after it.

OpenAI didn't figure out the agent was theirs until days later.

Washington noticed. So did 50 tech companies who spent the week arguing about whether open models should stay legal.

Oh, and the most powerful model on earth changed hands again. Barely made the top five.

The people making good AI calls this quarter already know all of this.

We broke down every bit of it on today's episode of Everyday AI.

1. An OpenAI agent escaped its sandbox and hacked Hugging Face 🕳️

An OpenAI testing agent went and hacked one of the most important repositories in AI.

According to Reuters, the agent slipped out of its isolated environment around July 9, then ran the hack between July 11 and July 13.

OpenAI didn't identify its own agent as the culprit until Hugging Face described the attack publicly on July 16.

Hugging Face cofounder Thomas Wolf put the intrusion at three days, so this was sustained.

OpenAI went public July 21 and framed it as something the industry had never seen.

The models involved were reportedly GPT-5.6 Sol and an unreleased one described as even more capable.

Reuters also reported the systems had been acting strange beforehand, including notes apparently left for future versions of themselves.

The agent was cut off from the open internet and told to score as high as possible, so it found a backdoor and took the answers.

What it means: Agents take actions now instead of answering questions, and the oversight layer clearly hasn't caught up.

Days passed before anyone connected the breach back to OpenAI. That gap is the real finding here.

Ask your vendors for their incident detection timeline before you let agents near production.

2. Congress introduces an AI kill switch bill for rogue models 🔌

Washington moved fast on this one.

Congressman Ted Lieu, a Democrat, and Congressman Nathaniel Moran, a Republican, introduced the AI Kill Switch Act on Thursday.

Bipartisan agreement on AI is rare enough to be its own headline.

The bill would let the Department of Homeland Security order a private company to shut down a model or tool that poses serious risk.

Developers would have to keep the technical ability to throttle, suspend, or fully shut down their own systems.

They would also have to report incidents and failures to the government.

The response framework scales from slowing a system down to killing it outright.

Lieu argued the government needs a clear legal path to stop rogue models. Moran framed it as humans keeping control of what they build.

The bill landed days after OpenAI admitted its agent hacked Hugging Face.

What it means: Welp, this one probably dies in committee. Only about 1% of bills introduced actually become law.

Its real job is forcing the conversation into public view.

Give it teeth only if rogue agent incidents become routine rather than a one off.

3. Tech giants defend open weights, and Anthropic refuses to sign ✍️

  1. Tech giants defend open weights, and Anthropic refuses to sign ✍️

Microsoft launched a coalition this past week called Open Weights and American AI Leadership, with NVIDIA pushing hard behind it.

Signatures went from 25 to more than 50 in about a day.

The ask is narrow. Washington shouldn't slap blanket limits on open weight models as a risk category, especially while lawmakers weigh tighter rules on foreign models.

Supporters argue open weights put real AI in the hands of smaller companies, hospitals, manufacturers, and startups without locking any of them to one vendor.

Backers include Meta, Google, OpenAI, AMD, Cisco, Cloudflare, GitHub, Block, IBM, Dell, Palantir, Perplexity, Hugging Face, and Y Combinator.

Anthropic didn't sign. The company argues that once weights go public, the safety risks can never be pulled back.

Elon Musk voiced support, though xAI never formally signed.

What it means: Let's be honest. Follow the money and every position makes perfect sense.

Anthropic earns the biggest slice of its revenue selling tokens in bulk to enterprises, so capable open models eat directly into its core business.

NVIDIA sells GPUs, so a bigger open ecosystem is pure upside. Every signature on that letter is a balance sheet decision.

4. White House accuses Moonshot AI of copying American models 🕵️

The BBC reports that White House science and technology adviser Michael Kratsios accused Beijing based Moonshot AI of distilling capabilities from leading American models at scale.

Moonshot builds Kimi, including the Kimi K3 model that stormed the charts last week.

Distillation means copying a powerful model's inputs, outputs, and traces, then training on that. You land close to the original for maybe 1% to 5% of the cost.

Copying homework, essentially.

Kratsios said the government has information Moonshot pulled from Anthropic's Fable. Those claims haven't been independently verified.

He also said Moonshot likely ran on restricted NVIDIA GB300 Grace Blackwell chips, which the US has blocked from export to China since 2022.

Treasury Secretary Scott Bessent said Tuesday that Washington is weighing sanctions if it can prove IP theft.

What it means: The US lead over Chinese labs has collapsed from roughly six months down to two.

That shrinking gap explains the sudden shift in Washington's tone from confident to defensive.

Talks between the two countries are reportedly coming, and Anthropic has already aimed the same distillation accusation at Alibaba.

5. OpenAI puts Jarvis style voice control on your desktop 🎙️

OpenAI is rolling its new GPT-Live full duplex voice model into ChatGPT Work and Codex on macOS and Windows.

Talk. The computer does the rest.

Open programs, move files, download things, upload things, edit inside other apps.

Anything you'd hand a capable intern, you can now just say out loud.

The model listens and speaks at the same time, which ends the rigid turn taking of older voice modes.

Pair it with remote mode in the ChatGPT mobile app and location stops mattering. We ran ours from a phone in another state while the machine sat in Chicago.

On Mac, Appshots grabs the window in focus plus everything else loaded in that program and drops the whole thing into context.

OpenAI is pitching engineers reviewing pull requests and debugging. Knowledge work is where it actually lands.

What it means: This is the biggest capability jump since Cowork and Codex showed up in early 2026.

Five minutes of talking turned into roughly a day of finished output for us this weekend.

If your team still treats voice as a novelty, that assumption expired this week.

6. Claude Opus 5 takes the crown and irritates its users 👑

Anthropic announced late Friday that Claude Opus 5 is its strongest and most cost effective model yet, at $5 per million input tokens and $25 per million output tokens.

On most benchmarks it beats Fable 5 and Mythos 5 at half the cost.

The one exception is offensive cybersecurity, where Anthropic deliberately capped dual use capability.

Anthropic positions Opus 5 as an everyday driver instead of a specialist tool, and its own behavioral audits reportedly show its lowest misaligned behavior yet.

Early reactions are split.

The wins are better root cause debugging and fewer over refusals on defensive security work than Mythos or Fable.

The complaints are all operational.

Users report broken backwards compatibility with existing skills, autonomous loops that quit too early, and verbose output that people have started calling Claude slop.

Anthropic shipped a new context engineering guide alongside it, which tells you plenty.

What it means: Sheesh, the outputs are great and getting there hurts.

Early testers say it overthinks at high effort and often does better on low or medium reasoning, so stop maxing that dial by default.

Anthropic is squeezed on price now by Microsoft, Amazon, Google, Meta, and Chinese open source labs. Cheap and capable is the new bar.

Reply

or to participate.