Agent OS Q&A: Setup, Loops + New AI Models
A Q&A session covering the Agent OS—a modular AI automation system where users can plug in various AI models and tools. The speaker demonstrates setup processes, discusses token optimization techniques, explains looping features in Claude Code, and addresses community questions about integrating multiple AI agents without overwhelming complexity.
Summary
The session begins with an overview of the Agent OS, a customizable system that allows users to integrate multiple AI agents and automate workflows within a single interface. The speaker highlights recent additions including a video agent built on Minimax H3, integration of Deep Seek V4 Flash (released 24 hours prior), and new Google Search Console integration for SEO content creation. Setup time is discussed as taking 20-30 minutes for experienced users and up to an hour for beginners, with API configuration adding 5-10 minutes.
The transcript covers several tools and services: Whisper Flow for voice-to-AI interaction with new note-taking features, Dina AI for 30-second video generation with multilingual support, and Buzz for agent workflow orchestration. A key concern raised is token usage and context bloat when chaining multiple agents; the speaker addresses this by recommending token minimization techniques like RTK Caveman and headroom ponytail, available through an open-source playbook in the AI Profit Boardroom.
For users struggling with tool complexity, the speaker recommends simplification—reducing to Claude Code alone rather than juggling Hermes, Buzz, and other tools simultaneously. Three looping approaches are explained: manual instruction ('keep running until finished'), the /goal command (loops until a judge model confirms goal completion), and the /loop command (time-interval based looping). The speaker notes spending 3-4 hours daily building and updating the system.
VPS deployment is addressed through case studies: Cloudflare tunnels for mobile/desktop access, manual agent-assisted setup on multiple devices, and Tailscale for cross-device access. The speaker uses Claude CLI locally rather than the API for development, leveraging existing context to quickly add features like new model integrations.
Key Insights
- The speaker claims that modular AI systems eliminate the feeling of being overwhelmed by rapid model releases because new models can be plugged in immediately without system overhaul, making the system better with each update rather than causing user anxiety.
- Setup time varies significantly by user experience level—20-30 minutes for experienced users, up to an hour for beginners, with most setup time spent on API configuration rather than the base system installation.
- Token bloat from multi-agent chaining is manageable through open-source token minimization techniques (RTK Caveman, headroom ponytail), and the speaker reports not approaching token limits despite running three separate workflows in one day.
- The speaker recommends simplifying to a single tool (Claude Code) with built-in looping features (/goal and /loop commands) rather than orchestrating multiple specialized tools, particularly for users finding the ecosystem confusing.
- The speaker maintains Agent OS locally for security reasons rather than on VPS, but notes community members have successfully deployed it via Cloudflare tunnels, Tailscale, and manual agent-assisted multi-device setup.
Topics
Transcript
[0:00] So today we're going to be answering the latest questions on the agent OS, which is a powerful system where you can plug all your agents in, build custom automations, have everything inside one tab, plug it into your memory, and basically build and automate anything that you want. To give you some examples over the last 24 hours, we have this video agent that is super powerful for, for example, adding avatar videos with uh fully edited video plugged in. As you can see, this was done literally with one single prompt, and it's on the new Miniax H3, which is pretty amazing in [0:30] itself. We also plugged in Deep Sea Coder. Deep Seek just dropped yesterday.…
Full transcript available for MurmurCast members
Sign Up to AccessMore from Julian Goldie SEO
Claude AI SEO Gets Me 2,680 Clicks a Day
The speaker demonstrates how to generate 2,680 clicks per day using an AI-powered SEO system that leverages social media content ranking on Google. The strategy involves using Google Search Console data to identify target keywords, then creating automated video content across multiple platforms to rank multiple times for the same keyword.
Antigravity 2.0 + Gemini 3.6 Flash is INSANE!
Google has upgraded its Antigravity AI agent platform with Gemini 3.6 Flash, a faster and more token-efficient model that enables multi-agent workflows. The system allows a single prompt to be split across multiple specialized agents (planner, writer, reviewer, critic, auditor) that work in sync through Agent OS, a shared memory layer that maintains consistent brand voice and context across all agents.
DeepSeek V4 Flash 0731 + Hermes Agent is INSANE!
DeepSeek V4 Flash, a newly released Chinese AI model optimized for agentic tasks, integrates with Hermes Agent to enable powerful automation workflows. The model offers significant performance improvements over its predecessor, cheaper pricing than frontier models, and a 1M token context window ideal for multi-step agent operations.
Impeccable Makes Claude, Codex and Kimi 10X Better
Impeccable is a free, open-source design tool that improves AI-generated web designs by eliminating generic templates and providing design direction through product/design files and 23 specific commands. The tool works across Claude, Codex, and other models, detecting and fixing common AI design flaws like repetitive gradients, poor spacing, and accessibility issues before deployment.
NEW DeepSeek V4 Flash Update!
DeepSeek V4 Flash has been officially released with significant performance improvements over its preview version, designed specifically for AI agents. The speaker demonstrates practical applications by building 50+ projects and integrating it into their agent operating system, while emphasizing it's not frontier-level but offers excellent speed and cost efficiency.