Agent OS Q&A: Setup, Loops + New AI Models
A Q&A session covering the Agent OS—a modular AI automation system where users can plug in various AI models and tools. The speaker demonstrates setup processes, discusses token optimization techniques, explains looping features in Claude Code, and addresses community questions about integrating multiple AI agents without overwhelming complexity.
Summary
The session begins with an overview of the Agent OS, a customizable system that allows users to integrate multiple AI agents and automate workflows within a single interface. The speaker highlights recent additions including a video agent built on Minimax H3, integration of Deep Seek V4 Flash (released 24 hours prior), and new Google Search Console integration for SEO content creation. Setup time is discussed as taking 20-30 minutes for experienced users and up to an hour for beginners, with API configuration adding 5-10 minutes.
The transcript covers several tools and services: Whisper Flow for voice-to-AI interaction with new note-taking features, Dina AI for 30-second video generation with multilingual support, and Buzz for agent workflow orchestration. A key concern raised is token usage and context bloat when chaining multiple agents; the speaker addresses this by recommending token minimization techniques like RTK Caveman and headroom ponytail, available through an open-source playbook in the AI Profit Boardroom.
For users struggling with tool complexity, the speaker recommends simplification—reducing to Claude Code alone rather than juggling Hermes, Buzz, and other tools simultaneously. Three looping approaches are explained: manual instruction ('keep running until finished'), the /goal command (loops until a judge model confirms goal completion), and the /loop command (time-interval based looping). The speaker notes spending 3-4 hours daily building and updating the system.
VPS deployment is addressed through case studies: Cloudflare tunnels for mobile/desktop access, manual agent-assisted setup on multiple devices, and Tailscale for cross-device access. The speaker uses Claude CLI locally rather than the API for development, leveraging existing context to quickly add features like new model integrations.
Key Insights
- The speaker claims that modular AI systems eliminate the feeling of being overwhelmed by rapid model releases because new models can be plugged in immediately without system overhaul, making the system better with each update rather than causing user anxiety.
- Setup time varies significantly by user experience level—20-30 minutes for experienced users, up to an hour for beginners, with most setup time spent on API configuration rather than the base system installation.
- Token bloat from multi-agent chaining is manageable through open-source token minimization techniques (RTK Caveman, headroom ponytail), and the speaker reports not approaching token limits despite running three separate workflows in one day.
- The speaker recommends simplifying to a single tool (Claude Code) with built-in looping features (/goal and /loop commands) rather than orchestrating multiple specialized tools, particularly for users finding the ecosystem confusing.
- The speaker maintains Agent OS locally for security reasons rather than on VPS, but notes community members have successfully deployed it via Cloudflare tunnels, Tailscale, and manual agent-assisted multi-device setup.
Topics
Transcript
[0:00] So today we're going to be answering the latest questions on the agent OS, which is a powerful system where you can plug all your agents in, build custom automations, have everything inside one tab, plug it into your memory, and basically build and automate anything that you want. To give you some examples over the last 24 hours, we have this video agent that is super powerful for, for example, adding avatar videos with uh fully edited video plugged in. As you can see, this was done literally with one single prompt, and it's on the new Miniax H3, which is pretty amazing in [0:30] itself. We also plugged in Deep Sea Coder. Deep Seek just dropped yesterday.…
Full transcript available for MurmurCast members
Sign Up to AccessMore from Julian Goldie SEO
How to Run DeepSeek V4 Flash for FREE!
A tutorial demonstrating how to use DeepSeek V4 Flash for free through Open Code and integration with agent operating systems like Hermes Agent. The speaker showcases building websites and apps using this free AI model and explains how it compares favorably to larger models despite being smaller.
Microsoft Fara1.5 27B NEW Browser Automation Model is WILD!
Microsoft released Phi-3.5, a family of three computer use models (4B, 9B, 27B) that automate browser tasks through vision-based clicking rather than HTML parsing. These open-weight models significantly outperform larger closed-source alternatives like OpenAI's Operator and Google's Gemini 2.0 on web automation benchmarks.
Claude Obsidian 2.0 is INSANE (FREE!)
Claude Obsidian 2.0 is presented as a free AI memory upgrade that allows users to upload files into a folder for permanent retention and linking. The system creates a knowledge graph that learns from business documents, provides sourced answers, and can be shared across teams.
Claude Agent OS is INSANE! 🤯
Julian presents a comprehensive Claude-based agent operating system that integrates multiple AI models, automated workflows, and a persistent memory system to automate daily tasks. The system runs 24/7 and uses free or existing subscriptions, combining tools like voice agents, content creation, competitor monitoring, and real-time news analysis into a single unified dashboard.
NEW ChatGPT Update is INSANE!
OpenAI released a major ChatGPT update featuring a Chrome extension called Side Chat and an improved desktop app that work together to streamline SEO research and content creation. The update allows users to analyze multiple browser tabs simultaneously, highlight text for quick answers, and convert research into finished work without constant tab switching.