workBy HowDoIUseAI Team

What actually happened in AI last week (and why it matters)

Claude Opus 4.5, Nano Banana in Google Search, and talking AI avatars all dropped at once. Here's what's worth your attention and what to skip.

Some weeks in AI feel like watching four different companies play chess against each other while you're just trying to figure out which piece to move. A new frontier model drops, a search engine quietly becomes an image editor, and somewhere in the middle of it all, an AI avatar named Piper is ready to co-host your next video. If you blinked, you probably missed at least two of these.

Here's the thing worth understanding: not every AI headline matters equally. Some releases change how you work starting today. Others are interesting but not urgent. This guide breaks down what actually happened, what it means for how you use AI day-to-day, and where to go to try each thing yourself.

What made Claude Opus 4.5 the model everyone's talking about?

Anthropic didn't just ship an update — it shipped what the company itself is calling a genuine leap forward. According to Anthropic's official Opus page, Claude Opus 4.5 is a hybrid reasoning model that pushes the frontier for coding and AI agents, setting a new standard across coding, agents, computer use, and enterprise workflows.

The coding improvements are the part actually worth paying attention to. Developers who spent years fighting infinite loops and half-fixed bugs in agentic coding tools are reporting a real shift here — the model handles ambiguous instructions without needing hand-holding, and testers found the model handles unclear requirements without extensive guidance.

Pricing stayed reasonable too. Pricing for Opus 4.5 starts at $5 per million input tokens and $25 per million output tokens, with up to 90% cost savings with prompt caching and 50% savings with batch processing. If you're building anything that leans on long, complex reasoning chains — multi-step agents, large codebases, enterprise workflows — this is the model to test first.

To try it yourself:

  1. Head to claude.ai and open a new chat
  2. Select Opus 4.5 from the model picker (Pro, Max, Team, and Enterprise plans all have access)
  3. Give it a genuinely messy, multi-step task — the kind you'd normally break into five separate prompts — and watch how it handles ambiguity on its own
  4. If you're a developer, grab API access through the Claude Developer Platform using the model ID claude-opus-4-5

Worth noting: Anthropic has kept shipping fast since this release, with newer Opus versions already rolling out. But Opus 4.5 remains the moment the "this model actually reasons through hard problems" narrative really took hold.

Why is Nano Banana showing up inside Google Search now?

This is the quieter release of the two, but arguably more disruptive long-term. Google didn't just improve its image generator — it started weaving it directly into the products billions of people already use every day.

Google announced Nano Banana Pro, built on Google's Gemini 3 Pro, which launched two days prior. That alone would be a solid week. But the real story is distribution: Nano Banana Pro became available in the Gemini app, Google's writing assistant, NotebookLM, as well as the company's developer, enterprise and advertising products, and Google AI Pro and Ultra subscribers gained access to the product in Google's search features AI Mode.

Think about what that actually means. You're no longer opening a separate app to generate an image — you're typing a search query and getting a generated, editable visual right there in the results. Google DeepMind's own Nano Banana model page walks through the different tiers now available, from the fast Lite version to the studio-quality Pro model.

Here's how to try it inside Search directly, based on Google's own usage guide:

  1. Open Google Search or the Gemini app and switch to AI Mode
  2. Select the "🍌 Create images" tool from the menu
  3. Choose Fast, Thinking, or Pro depending on how much detail and accuracy you need
  4. Type a specific, detailed prompt — something like "create an image of a cat napping in a sunbeam on a windowsill" works better than vague requests, since the more details you provide, the better Gemini is at following your instructions

The practical takeaway: if your work involves visuals — marketing mockups, quick concept art, product images — you may not need a dedicated design tool for first drafts anymore. The image generator is already where you're searching.

How are AI website builders turning into a speed competition?

One of the more entertaining trends happening right now is watching AI coding agents race each other to build the same project from scratch, live, in front of an audience. Give two agents the same brief — build a small themed website, say a garden site with some playful animated elements — and watch which one finishes first and which one actually looks usable.

This isn't just a stunt. It's a genuinely useful way to stress-test tools before you commit to one for actual client or business work, because it exposes exactly where each agent cuts corners: broken navigation, elements that don't render correctly, or logic that technically runs but looks rough. Tools like Lovable, Bolt, and v0 by Vercel all compete in this exact space — take a plain-English brief and turn it into a working, styled site or app in minutes rather than hours.

If you want to run your own version of this test:

  1. Pick two or three AI app builders you're curious about
  2. Give each the exact same one-paragraph brief — be specific about the theme, the pages needed, and any interactive elements
  3. Time how long each takes to produce something usable
  4. Compare not just speed but whether basic things work — do buttons actually click through, does the layout hold up on mobile, does it look like a real site or a rough draft

The honest conclusion most people land on after running this test themselves: speed is impressive, but "first to finish" and "actually good enough to ship" are two very different races.

What's the deal with talking AI avatars you can host videos with?

The other release worth understanding is the shift toward real-time, interactive AI avatars — not the pre-recorded talking-head videos you've seen for a couple of years now, but avatars that can actually listen and respond live.

HeyGen has been at the center of this shift, rebranding its real-time avatar feature into a dedicated product. According to HeyGen's help documentation, LiveAvatar is HeyGen's real-time AI avatar platform for live, interactive conversations, where avatars can listen and respond instantly through voice, video, or text with minimal delay.

What makes this genuinely useful rather than a gimmick is the flexibility in how much of the "brain" you control. Per HeyGen's developer docs, FULL mode means HeyGen manages the end-to-end conversation including speech-to-text, the LLM, text-to-speech, turn-taking, and memory, while LITE mode lets you keep your own agent or LLM and use LiveAvatar purely for the avatar layer.

Here's how to set up your first interactive avatar session:

  1. Go to app.liveavatar.com and log in with your existing HeyGen credentials — you can log in with the same credentials you use for HeyGen
  2. Toggle between your own custom avatars, public avatars, or built-in example avatars if you're just testing
  3. Set up a Context (previously called a Knowledge Base) to control tone and what your avatar knows — select your avatar, click the Context button, and add instructions, tone preferences, and information your avatar should know
  4. Customize the voice and rename your avatar before going live

Where this actually gets useful: two-person setups where you host alongside a built-in avatar as a co-presenter, running product demos, or building customer support experiences where the "personality" answering questions has a face and voice instead of just text on a screen.

What should you actually do with all of this?

If you only have time to try one thing from this roundup, make it Opus 4.5 for anything code-related — the reliability jump is the kind of thing you notice on the very first real task you throw at it. If you're doing visual work, get familiar with Nano Banana living inside Search now, because that's where the habit-forming distribution advantage really lives. And if you're building anything client-facing that benefits from a human presence, the interactive avatar space is moving fast enough that it's worth a test run even if you don't have an immediate use case yet.

None of these tools are perfect. The website-builder race still produces broken layouts half the time, and avatar controls occasionally glitch in ways that feel more sci-fi-gone-wrong than polished product demo. But that's kind of the point of watching this space closely — the rough edges this month are usually smoothed out by next month, and the people paying attention now are the ones who'll know exactly which tool to reach for when it matters.