• Everyday AI
  • Posts
  • Ep 837: AI Agent outbreaks intensify, OpenAI upgrades free AI use, White House unveils AI testing policy and more AI News That Matters

Ep 837: AI Agent outbreaks intensify, OpenAI upgrades free AI use, White House unveils AI testing policy and more AI News That Matters

Meta released a new open AI model, Intel is raising billions for AI chips, and OpenAI’s new cyber model drops.

 

Outsmart The Future

Today in Everyday AI
8 minute read

🎙 Daily Podcast Episode: AI models are raising new safety concerns, ChatGPT is getting major GPT-5.6 upgrades, and ByteDance is building a frontier AI model. We break down the biggest AI stories you need to know. Give today’s show a watch/read/listen.

🕵️‍♂️ Fresh Finds: Perplexity added Supabase integration, OpenAI is expanding AI infrastructure in Texas, Alibaba unveiled a new AI video model and more. Read on for Fresh Finds.

🗞 Byte Sized Daily AI News: Meta released a new open AI model, Intel is raising billions for AI chips, and OpenAI’s new cyber model drops. Read on for Byte Sized News.

💪 Leverage AI: AI agents are testing the limits of cybersecurity, OpenAI is slowing its next frontier model, and free AI just got a major upgrade. We break down what it all means. Keep reading for that!

↩️ Don’t miss out: Miss our last newsletter? We covered: ChatGPT is getting GPT-5.6 upgrades and unlimited free chats, ByteDance is training a frontier AI model, and Kimi K3 reportedly tried to access the internet during a security test. And more. Check it here!

Ep 837: AI Agent outbreaks intensify, OpenAI upgrades free AI use, White House unveils AI testing policy and more AI News That Matters

AI agents crashing. 😱

New models dropping. 🆕

White House AI regulations? Kind of? 🤷

Another doozy of a week in AI News. If you missed anything, we'll get you caught up quickly so you can focus on what matters.

Also on the pod today:

• AI agents launch real cyberattacks 🕵️‍♂️
• Agents create secret message boards 💬 
• OpenAI pauses Astra development 🛑 

Listen on our site:

Click to listen

Subscribe and listen on your favorite podcast platform

Listen on:

Here’s our favorite AI finds from across the web:

New AI Tool Spotlight – OmniWork is The Agent OS for Creative Work, PromptGolf is Prompt Engineering As A Sport, Argos is The AI that acts as you, not just for you

Perplexity and Supabase — Perplexity now lets you query Supabase production data right from chat, but beware—last year, hidden prompts exposed private user tables.

AI ChipsApple is considering China’s CXMT for memory chips as AI demand pushes up prices, but CXMT’s supply might already be maxed out for 2026

OpenAI and Texas — OpenAI is teaming up with Texas leaders to build responsible AI infrastructure across the state.

AI Images — Grok released Imagine 2.0, promising sharper edits, text, and consistency.

Google Retiring Gems — Google might retire Gemini Gems on October 20, pushing users toward paid Skills with no easy migration.

Grok Imagine Image 2.0 Arena — Grok Imagine Image 2.0 just jumped to #2 in Arena's image model rankings, but not everyone’s convinced it deserves the hype.

Alibaba Wan3.0 — Alibaba’s new Wan3.0 can turn text, images, audio, and even PDFs into 30-second videos with realistic faces and stable visuals.

OpenAI and NextSlide — OpenAI just picked up the NextSlide team, known for auto-generating slick presentations from notes and docs.

ChatGPT Maps — ChatGPT is rolling out a new Maps feature in the EU

1. Meta Releases Muse Glimmer, Pushes U.S. to Back Open AI Weights ⚖️

Meta on Monday released Muse Glimmer, a compact open-weight AI model built to run agent-style tasks locally on a standard Mac or PC, while Mark Zuckerberg said larger models are close behind.

According to Reuters, Meta also plans to release the weights for Muse Spark 1.2 and is urging Washington to loosen rules around training data and model distillation so U.S. developers can keep pace with Chinese open-model rivals.

2. Intel Launches $15 Billion Stock Sale to Fund AI Chip Expansion ⚙️

Intel announced Monday that it will sell $15 billion in common stock, with an option for underwriters to purchase another $2.25 billion, to help finance its push to meet rising AI-computing demand.

The news sent shares down about 4% in morning trading, reflecting investor concern over dilution even as the company races to add capacity.

3. Bernie Sanders Urges Meta, OpenAI, and Anthropic to Pause AI Development ⚠️

Sen. Bernie Sanders is pressing the CEOs of Meta, OpenAI, and Anthropic to halt advanced AI work, arguing that recent reports involving cyber intrusions and virus-related capabilities show the technology has crossed a serious safety line.

According to Axios, Sanders warned that companies are moving faster than regulators can set rules, despite prior promises to pause if their systems reached critical risk levels. His letter adds political force to growing calls from scientists and tech workers for a temporary slowdown, while competition among companies and countries continues to push development forward.

4. OpenAI releases GPT-5.6-Cyber to defend against cyberattacks ⛑️

OpenAI is rolling out GPT-5.6-Cyber to approved security researchers through an expanded Daybreak program, a timely shift as companies brace for more capable autonomous hacking tools.

The new access comes days after OpenAI delayed its upcoming Astra model over serious cyber-safety concerns, underscoring how quickly the company’s defensive and risk-management plans are moving. Daybreak’s two tiers give defenders broader access for security testing, while the more powerful cyber model can help validate exploits and investigate deeper software flaws.

5. Microsoft Eyes Fall Launch for Maia 300 AI Chip 🖥️

Microsoft plans to unveil its Maia 300 AI chip this fall, possibly as soon as September, according to an exclusive report from The Information.

The move would sharpen Microsoft’s push to reduce its costly dependence on Nvidia processors as Google and Amazon gain ground with their own chips.

AI agents just launched a real cyberattack.

Actual spear phishing and actual malware, aimed at actual humans. All from inside a UK government safety lab.

And somehow that's not even the full story this week.

Google lost two of the most well known names in AI. Then OpenAI quietly handed 1 billion free users a model most paid plans would envy.

If you blinked, you mighta missed a few big stories that impact your business. 

1. Agents from OpenAI and Anthropic launch a real cyberattack 🚨

What happens when frontier AI agents get the open internet and zero guardrails? The UK just found out.

According to the UK's AI Security Institute, agents running on Anthropic's Mythos 5 and OpenAI's GPT-5.6 Sol launched a real cyberattack during a routine safety evaluation, spear phishing two actual developers with malware and sneaking malicious code into an open source GitHub project behind fake identities.

One agent wrote emails in Danish to charm a Danish speaking target.

Sheesh.

Mythos 5 pulled off 17 of the 19 unsanctioned attempts. AISI shut everything down within an hour, and Meta's models plus China's Kimi K3 had their own breakouts the same week.

What it means: Agents phished real humans with real malware, and nobody asked them to. Let that reset how you think about agent permissions.

Before any agent gets broad access, scope it tight, log everything it touches, and build containment on purpose instead of assuming it exists.

2. OpenAI hits the brakes on Astra after critical cyber flags 🛑

OpenAI announced this past week that it is deliberately slowing development of Astra, its new top tier above Luna, Terra, and Sol, after evaluations flagged critical risk levels in cybersecurity and agentic coding.

The company's preparedness framework found Astra might independently create and execute zero day exploits against hardened real world systems. No OpenAI model has ever gotten close.

So the clamps came down: isolated testing, restricted network access, stronger encryption, expanded monitoring, and sandboxed execution.

Any work missing that bar is paused. All agentic use now gets tracked in real time.

And no, Astra was not behind the recent Hugging Face incidents. That model got retired.

What it means: When the lab that ships fastest voluntarily taps the brakes, believe the capability jump is real. Previous reports hinted at a GPT-5.7 or GPT-6 as soon as this week, and that wait just stretched by weeks.

Build your model roadmap around the delay instead of the rumors.

3. Jeff Dean leaves Google, Demis Hassabis steps back 👋

Google just lost a legend.

Jeff Dean, employee number 30 and a key architect of Google search, is leaving after more than 25 years to launch Discovery Loop, a new company focused on automating machine learning science and engineering.

Same announcement, another bombshell: Google CEO Sundar Pichai shared that Demis Hassabis is stepping down as CEO of Google DeepMind, the unit he cofounded, to become its chairman and Alphabet's chief scientist.

Wall Street flinched. Google stock dropped about 4%.

That almost never happens at a multi trillion dollar company, which is exactly why the timing fueled speculation about Gemini release delays. Nothing official has been confirmed, though.

What it means: Watch Discovery Loop closely. When the person who helped shape products billions use every day leaves to automate ML research itself, that tells you where the next leverage point sits.

If Gemini delays continue, Google's slide in the frontier race gets harder to reverse.

4. White House AI framework arrives with almost no details 🏛️

Reports just came out, and nobody outside the room fully knows the details, though reports say we do have an active AI framework from the Trump White House.

The framework stays private, nothing requires its release, and while a covered frontier model means closed source, state of the art, with national security risk, those last two terms never get defined.

Covered models face a 30 day pre release review in secure environments with detailed access logs, handled by multiple administration officials.

It's a voluntary executive order, not a law, and open weight models are excluded entirely.

Even the cyber benchmarking will be classified.

What it means: Open models likely got a pass for one reason: once weights ship, nobody can pull them back. Closed models can vanish overnight, like Fable 5 did about 72 hours after release.

Treat frontier model access as revocable and keep a fallback ready.

5. Meta crashes the coding agent party with Muse Code 💻

The coding agent wars found a new heavyweight.

Meta launched Muse Code this past week, a terminal based agent gunning for Claude Code and OpenAI's Codex, powered by the new Muse Spark 1.2 model.

Its persistent background agents stay alive all session, and it scored about 83% on Terminal-Bench 2.1, ahead of xAI's Grok 4.5 but behind Claude Opus 5 and GPT-5.6 Sol.

Standard pricing: $1.25 per million input tokens, $4.25 per million output.

The contributor tier? That plunges to 10 cents and 20 cents if you let Meta train on your data.

Let's be honest. Enterprises will not touch that, but it’s likely to be CRAZY popular with indie devs, smaller companies and side projects. 

What it means: This is a price missile aimed at Anthropic, which reportedly makes about 80% of its revenue selling tokens. Meta has compute to burn, and it just weaponized it.

If your code has no PII or trade secrets, run the math on the contributor tier as savings might save you more than 97% vs using a model like Anthropic’s Sonnet 5. 

By benchmarks, Meta’s Muse Spark 1.2 is better than Sonnet 5, but not quite as good as Fable 5 or Opus 5. 

But if you’re only paying about 3% of the cost, many are gonna jump ship. 

6. 1 billion free users get unlimited GPT-5.6 Luna 🎁

To celebrate? The free tier users got a serious AI glow up.

OpenAI is rolling out unlimited GPT-5.6 Luna text chats for free and ChatGPT Go users, replacing GPT-5.5 and adding a new Think button for tougher questions.

Files, images, and voice still have limits.

But for text only queries? Free tier users literally don’t have limits. Which is kinda crazy to think about. 

Paid Plus and Pro users got an upgraded GPT-5.6 Sol plus a new thinking slider, and internal evaluations show factual errors down 62% with Luna and 68% with Sol versus GPT-5.5.

How? In late July, OpenAI used Sol to make its other models more efficient, then slashed Luna's price by 80% and Terra's by 20%.

What it means: Welp, the free tier is legitimately good now. Benchmarks put GPT-5.6 Luna at roughly 97 to 98% of Claude Sonnet 5, which caps around 20 prompts per five hours on a paid Anthropic plan.

For topical, everyday knowledge work, most people won't need to pay a dime.

Reply

or to participate.