New Google AI Updates Are INSANE!
Google announced a sweeping set of AI updates at their Cloud Next event, including new TPU chips, an enterprise agent platform, advanced research agents, and new embedding and training technologies. These updates are positioned as tools that will make AI faster, cheaper, and more accessible for businesses of all sizes. The presenter argues that companies adopting these tools quickly will gain a significant competitive advantage.
Summary
The video covers a series of major AI announcements from Google's Cloud Next event, framed around their potential business impact. The presenter begins with Google's eighth-generation TPU chips — the TPU 8T (for training) and TPU 8I (for inference) — noting that the training chip is nearly three times more powerful than its predecessor. The key takeaway is that cheaper, faster AI hardware reduces costs for end-users across all AI tools, not just those building models.
The next major announcement is the Gemini Enterprise Agent Platform, described as a one-stop shop for deploying AI agents in business workflows. It includes Agent Designer, a no-code tool for building custom agents; long-running agents that can operate autonomously for hours or days; and an Agent Inbox for monitoring all active agents in one dashboard. The platform supports over 200 AI models, including Gemini 3.1 Pro and models from Anthropic, giving users flexibility to choose the best model per task.
Google also launched Deep Research and Deep Research Max — two AI agents designed for automated research. Deep Research prioritizes speed, while Deep Research Max is built for depth, capable of running for hours and producing comprehensive reports with charts and citations. The presenter highlights that Deep Research Max scored 93.3% on a rigorous research benchmark, compared to 66% for the previous version — a significant jump. It can also search private internal files and company data.
Gemini Embedding 2 was made generally available, enabling multimodal search across text, images, video, and audio. The presenter gives examples like searching hours of video footage by typing a natural language query, or allowing e-commerce customers to upload a photo to find matching products without keywords.
Google also open-sourced a format called design.md, which acts as a brand style guide that any AI tool can read and apply consistently — covering colors, fonts, and visual style. Because it's open-source, it works across tools like Claude Code, Cursor, and Copilot, not just Google's ecosystem.
For Google AI Pro and Ultra subscribers, Google expanded access to higher usage limits and the Nano Banana Pro image generation model inside Google AI Studio. Finally, Google DeepMind introduced Decoupled Dial-A-Co, a distributed AI training technology that allows model training to be split across multiple data centers globally, eliminating single points of failure and making AI development faster and more resilient. It was tested on the Gemma 4 model family across four US regions successfully.
Key Insights
- Google's new TPU 8T training chip is claimed to be nearly three times more powerful than its predecessor, which the presenter argues will reduce the cost and time of AI development and ultimately lower prices for end-users of consumer AI tools.
- The Gemini Enterprise Agent Platform includes 'long-running agents' capable of operating autonomously for hours or even days, enabling workflows like overnight lead generation, outreach, and calendar booking without human involvement.
- Deep Research Max scored 93.3% on a rigorous research benchmark compared to 66% for the previous version — a jump the presenter characterizes as massive — and can additionally search a user's own private files and internal company data.
- Gemini Embedding 2 supports multimodal search across text, images, video, and audio in a single model, enabling use cases like natural language search across hours of video footage or photo-based product search in e-commerce without keywords.
- Google DeepMind's Decoupled Dial-A-Co technology decouples AI model training from a single cluster of chips, distributing it across multiple data centers globally so that a failure in one location does not halt the entire training process — validated on a 12-billion parameter Gemma 4 model across four US regions.
Topics
Transcript
[0:00] New Google AI updates are insane. Today I'm going to show you the craziest Google AI updates ever. New chips, new agents, new everything. This stuff is so big, it could change your business overnight. And the best part? Most people don't even know it exists yet. So stick around because by the end of this video, you'll have a huge edge over everyone else. Let me tell you, Google just dropped a bomb at their Cloud Next event. Talking new chips, new agents, [music] new tools, all built for one thing: help you run your business with AI. So let's get into it. First big drop is Google's brand new chips. They're called TPUs eighth generation and [0:30]…
Full transcript available for MurmurCast members
Sign Up to AccessMore from Julian Goldie SEO
NEW Nvidia Autonomous AI is WILD!! 🤯
Nvidia announced Nemo Clo, a new autonomous AI agent system that operates independently without continuous prompting. Powered by Nemotron 3 Ultra (a 550 billion parameter model), the system is five times faster and cheaper than previous versions, with OpenShell providing secure sandboxed execution.
Laguna XS 2.1: New FREE + Opensource Local AI!
Julian reviews Laguna XS 2.1, a new free open-source local AI coding model from Poolside that performs comparably to Qwen 3.6 and outperforms Claude Haiku on benchmarks. He demonstrates its practical capabilities by building landing pages and functional apps, highlighting its speed, offline functionality, and multiple deployment options through local setup, Claude Code, or OpenRouter's free API.
How to Run Hermes FREE Forever!
The video demonstrates how to run the Hermes AI agent for free using Gemma 4, a local open-source model from Google, with significant speed improvements through MLX optimization. The setup works on Apple Silicon Macs or via free APIs on Open Router, enabling autonomous agents to work offline and privately without subscription costs.
This NEW Chinese AI is INSANE! (FREE + Open Source!)
Long Cap 2.0 is a new open-source Chinese AI model from a food delivery app company that offers 1 million tokens of free context memory, beats GPT-4.5 on SWE bench pro benchmarks, and uses efficient parameter activation to reduce computational overhead while maintaining high performance.
Claude Code is now FREE: Here’s how…
Google's new Gemma 4 model running on Ollama is 90% faster on Apple Silicon, enabling free Claude Code usage locally without token costs. The setup requires three simple steps: downloading Ollama, Gemma 4, and installing into Claude Code, with alternatives available via OpenRouter API for non-Mac users.