• Everyday AI
  • Posts
  • Ep 826: ChatGPT goes Jarvis Mode, Claude can learn from you, Google unleashes spark agent and 7 more AI updates you can use today

Ep 826: ChatGPT goes Jarvis Mode, Claude can learn from you, Google unleashes spark agent and 7 more AI updates you can use today

U.S. tech giants backed open-weight AI, lawmakers introduced an AI kill switch bill, and Google expanded Gemini Spark. And more.

 

Outsmart The Future

Today in Everyday AI
8 minute read

🎙 Daily Podcast Episode: OpenAI, Anthropic, Google, and Microsoft all dropped major AI updates this week. Here are the 7 worth using today. Give today’s show a watch/read/listen.

🕵️‍♂️ Fresh Finds: Alexa+ is getting smarter, Stripe is reportedly eyeing OpenRouter, and FLUX 3 launched a new multimodal AI model. And more. Read on for Fresh Finds.

🗞 Byte Sized Daily AI News: U.S. tech giants backed open-weight AI, lawmakers introduced an AI kill switch bill, and Google expanded Gemini Spark. And more. Read on for Byte Sized News.

💪 Leverage AI: OpenAI, Anthropic, and Google all made AI more hands-free this week. Here's how to use the latest voice and agent features at work. Keep reading for that!

↩️ Don’t miss out: Miss our last newsletter? We covered: AMD landed a major Anthropic deal, Amazon cut jobs in its AGI division, and Alphabet is boosting AI spending. And more.Check it here!

Ep 826: ChatGPT goes Jarvis Mode, Claude can learn from you, Google unleashes spark agent and 7 more AI updates you can use today

Over 3 hours, OpenAI, Anthropic, Google AND Microsoft all dropped new AI upgrades that are live.

How you use AI in your work literally changes every day, as frontier labs are racing to roll out big quality of life updates between big model drops.

How can you keep up?

This week did not disappoint. You don't want to miss what's now at your fingertips. JARVIS mode, anyone?

Also on the pod today:

• ChatGPT Health syncs Apple data 🍏 
• Microsoft MAI Image 2.5 Pro 🖼️
• Google Gemini 3.6 Flash drops ⚡

Listen on our site:

Click to listen

Subscribe and listen on your favorite podcast platform

Listen on:

Here’s our favorite AI finds from across the web:

New AI Tool Spotlight – Fluree AI is the hosted Fluree platform, built on FlureeDB. No clusters, no setup, HarnessRouter Brings the world's best AI agents into your app, FreeSolo is Post-Training Built for your Agent.

Alexa+ Upgrades — Alexa+ is getting a lot more useful, with new tools that let brands plug in smarter device controls, service integrations, and even checkout.

Stripe and OpenRouter — Stripe may be lining up a $10B buy of OpenRouter, a move that would give it a bigger foothold in AI.

FLUX 3 Model — FLUX 3 is being pitched as more than a content generator, it’s a single backbone that learns the world through video, audio, and actions.

ElevenLabs Music v2 — Upload a reference track, and Music v2 can generate something that matches its style, instrumentation, and feel.

Microsoft MAI Models — Microsoft says its smaller MAI models are matching GPT-5.6 on common Excel tasks and beating similar-size models in GitHub Copilot, while using fewer tokens and less compute.

Grok Build Workflows — Grok Build can now split big, messy tasks across hundreds of agents and bring back one clean report when it’s done.

Runway API — Runway is pitching one API for image, video, audio, and real-time media, with model routing, evaluation, and usage controls built in.

NVIDIA on The Moon — Nvidia’s Jetson chips could soon be running a rover on the moon, helping Lunar Outpost’s robots process lidar data on the fly.

Google Selfie sign-in — Google is rolling out selfie video as a new sign-in option, giving you another way back into your account if you’re locked out.

NVIDIA and KAIST — NVIDIA and KAIST are teaming up on a new AI lab in Seoul, with funding, compute, and internships to push agentic AI built for Korea.

1. US Tech Giants Back Open-Weight AI 🏋️

A major coalition of U.S. tech leaders has just thrown its weight behind open-weight AI, arguing that America’s AI edge depends on models that can be downloaded, inspected, and modified rather than locked away.

Nvidia, Microsoft, Dell, Palantir, and others said open systems speed innovation, improve security, and help keep AI leadership broad instead of concentrated in a few closed models.

2. AI Kill Switch Bill ⚠️

Two House lawmakers introduced the “AI Kill Switch Act” after OpenAI’s recent security incident raised fresh concerns about rogue models.

The bill would require companies to keep the ability to slow, suspend, or shut off AI systems and would let Homeland Security step in if a model could cause catastrophic harm.

3. Google widens Spark access for paid users ⚡

Google is expanding Gemini Spark, its agentic AI assistant, to more paid subscribers, with U.S. Google AI Pro users and Google AI Ultra users in most regions now getting access.

The move matters because Spark is built into Workspace and can handle routine email, calendar, and document tasks, turning Google’s software suite into something closer to a hands-off assistant.

4. Experts Reject Fable-Only Theory for Kimi K3 🫸

Fresh scrutiny over Moonshot’s Kimi K3 has centered on claims that it was built by copying Anthropic’s Fable, but researchers say that theory does not match the timeline or the model’s capabilities.

According to TechCrunch, experts argue that straightforward distillation would be too slow and too limited to explain how Kimi K3 became so strong so quickly.

5. AMD launched Helios, a new AI factory rack system 🧠

AMD launched Helios as its latest push to rewire AI infrastructure around full-rack systems, not just standalone chips, at a moment when token demand, inference, and agentic workloads are surging.

The platform combines MI455X GPUs, EPYC CPUs, Pensando networking, and ROCm software in a 72-GPU design meant to boost throughput, memory capacity, and bandwidth while improving token economics.

6. Cognition buys Poke-maker Interaction ☝️

Cognition said today it is acquiring The Interaction Company of California, the team behind Poke, bringing together a well-known consumer AI texting app with its software engineering agent Devin.

The deal matters now because Poke has already crossed 100 million messages in three months and is the only AI agent approved to text natively on Apple Messages, giving Cognition a ready-made product with real user traction.

Your keyboard just got demoted.

This week, OpenAI, Google, and Anthropic shipped the same idea in different wrappers: stop typing, start talking, let AI do the work.

Tell ChatGPT what you need and it'll run multiple agents at once. Show Claude a task one time and it'll repeat it forever.

Tell Gemini to handle your week and it does.

That's not some future roadmap, it's live today.

Plus a ChatGPT built around your medical history, and a smarter voice mode for Claude with real brains behind it.

Biiiig week.

If your team's still hand-typing every prompt, y'all are prolly behind whoever isn't.

1. ChatGPT Health Now Live For Everyone 🩺

OpenAI rolled ChatGPT Health out to every US user, ending its January waitlist. Connect Apple Health and your medical records, and it lives in a private space away from your regular chats.

It's free on every plan, though you need to be 18 or older.

That's the unlock: one place for health info scattered across different portals, EMRs, and apps, ready to compare lab results, flag medication changes, or prep you for your next appointment.

It won't diagnose anything, but it'll make you the most informed patient in the room.

Try This

Connect Apple Health to ChatGPT and ask it to summarize what's changed since your last physical. You'll walk into your next appointment already knowing what to ask.

2. Claude Voice Mode Adds Opus, Sonnet 🎙️

Anthropic upgraded Claude's voice mode to run on Opus and Sonnet, not just the Haiku model it launched with, and you can now swap models mid-conversation while voice mode works with your connectors.

Paid accounts get the upgrade, free stays on Haiku with one connection, and it covers 11 languages.

It's still walkie-talkie style, listen, pause, respond, not full duplex like ChatGPT's GPT-Live, but pulling live data from Gmail or Slack mid-conversation is a big jump for anyone doing hands-free work between meetings.

Try This

Open Claude voice mode, connect one app like Gmail, and ask it something that needs that app's live data. If it delivers hands free, make voice mode your default for thinking out loud between meetings.

3. Microsoft Ships MAI Image 2.5 Pro 🎨

Microsoft launched MAI Image 2.5 Pro, its highest fidelity image model yet, built for detailed editing and precise in-image text. It's live in Microsoft Foundry, inside Copilot PowerPoints on a Microsoft 365 Copilot plan, and through the API at $5 per million text input tokens and $106 per million image output tokens.

If Copilot is the only AI tool your organization allows, this closes a gap for building decks and marketing visuals without leaving Microsoft's ecosystem. It's a clear upgrade over MAI Image 2.0, though the market's top image models likely hold an edge on raw quality.

Try This

Next time you're stuck with clip art in a Copilot deck, generate the hero image with MAI Image 2.5 Pro instead. See if it's finally good enough to skip your usual image tool.

4. Google Ships Gemini 3.6 Flash, Flash-Lite ⚡

Google released three new Gemini models in one shot: Gemini 3.6 Flash and Gemini 3.5 Flash-Lite for everyone, and a restricted Gemini 3.5 Flash Cyber for governments and trusted partners. All three are live now across Gemini, the API, AI Studio, Android Studio, and Gemini Enterprise on paid plans.

3.6 Flash burns 17% fewer tokens than its predecessor with fewer reasoning steps per workflow, and Gemini 3.5 Pro still hasn't shipped while Google quietly pretrains Gemini 4.

Cheaper wins, and it adds up fast.

Flash-Lite will power Google's search AI overviews, where speed is everything.

Try This

If you're running Gemini 3.5 Flash in production, swap in 3.6 Flash for one week and track your token spend. If costs drop and quality holds, make the switch permanent.

5. Gemini Spark Opens To Pro Users 🤖

Google widened Gemini Spark, its 24/7 personal AI agent, from Ultra subscribers only to Pro subscribers too. It's live for paid Gemini Pro users in the US, with more countries coming soon, though Workspace business plans aren't guaranteed access yet.

Spark runs multi-step tasks on its own, research, planning, monitoring, built around tasks, skills, and schedules, and it can open, read, and edit Docs, Sheets, and Slides directly, including shared team files.

That's a meaningful shift for teams.

Try This

Give Spark one real multi-step task, like planning a trip using your Gmail and calendar. See how much busywork it clears before you even open your inbox.

6. Claude Cowork Learns By Watching You 🎬

Anthropic shipped Record a Skill inside Claude Cowork. Screen-record yourself doing a task once, narrate what you're doing, and Claude turns it into a reusable skill, no SKILL.md file, no prompt engineering required.

It's rolled out to all paid Claude plans, though it's desktop only, tucked inside the Cowork tab, which kills the biggest barrier to building a skill: explaining it in writing.

And because skills are shareable, one you record in Claude can run in Codex or ChatGPT too, which turns this into less of a Claude feature and more of a new way to hand off repeat work.

Try This

Screen-record yourself doing one task you repeat every week and narrate your way through it. Save it as a skill and hand it off next time instead of doing it yourself.

7. ChatGPT Voice Now Controls Your Desktop 🗣️

OpenAI launched ChatGPT Voice on desktop, powered by GPT-Live, so you can control your computer and direct multiple agents in ChatGPT Work or Codex using just your voice. It's rolling out on macOS and Windows for Plus, Pro, Business, Edu, and Enterprise plans, and it works through ChatGPT's remote feature from your phone.

On Mac, Appshots pulls in full context from whatever program you have open, not just what's visible on screen, so you can juggle several apps hands-free.

No hands.

This is the first way to direct AI agents doing real work with your voice, not just chat with one.

Try This

Open ChatGPT Voice on desktop, pull up two apps you use daily, and direct a task in each one using only your voice. Once you see it work hands free, rethink how much of your day actually needs a keyboard.

Reply

or to participate.