How I AI

How I AI

YouTube44 episodes summarized

MurmurCast publishes AI-generated summaries of How I AI’s YouTube episodes — 44 summarized so far, covering AI-assisted video editing, Social media content creation, Workflow automation, GPT-5.6 and Codex capabilities, Product manager productivity, AI model comparison (Fable vs Soul/GPT-5.6 Sol). Each summary distills the key insights, topics, and takeaways so you can decide what’s worth your time before pressing play.

GPT-5.6's video editing via Codex is genuinely one of my favorite new workflows

Jul 11, 2026

A product manager describes using GPT-5.6 with Codex to automate video editing for social media clips. By simply dragging a file and providing natural language instructions, the AI generated five polished, fast-paced hype videos from a long conference talk recording, dramatically reducing the time-intensive manual clipping process.

StoryTechnicalAI-assisted video editingSocial media content creationWorkflow automation

Theoretically Intelligent vs. Practically Effective: Why GPT-5.6 Sol Beats Fable for Product Work

Jul 10, 2026

An executive contrasts two AI models (Fable and Soul/GPT-5.6 Sol), arguing that Soul is superior for product work because it prioritizes practical effectiveness over theoretical intelligence. The speaker values the ability to ship products to customers and understand end-user goals over theoretical sophistication.

OpinionDiscussionAI model comparison (Fable vs Soul/GPT-5.6 Sol)Theoretical intelligence vs practical effectivenessProduct development and shipping

Build a harness when the same workflow needs the same setup and the same outcomes, every time

Jul 10, 2026

Building a harness for repetitive workflows allows you to be more prescriptive about job execution, resulting in greater efficiency, consistency, and better outcomes. Rather than explaining requirements to an AI agent each time, a harness lets you use a simpler interface like pasting a link while the agent already understands the intended task.

TechnicalOpinionWorkflow harnessesPrescriptive job executionAI agent automation

GPT 5.6-Sol vs. Claude Fable: Why OpenAI’s new model crushes my benchmark

Jul 9, 2026

Claire Vo compares OpenAI's new GPT 5.6 models (Soul, Terra, Luna) against Claude's Fable using her custom "How I AI" benchmark, finding that GPT 5.6 Soul excels at practical product work, prototyping, and natural communication, while Fable is theoretically intelligent but pedantic and difficult to collaborate with.

OpinionResearchGPT 5.6 Model Variants (Soul, Terra, Luna)Custom AI Benchmarking MethodologyModel Comparison: OpenAI vs. Anthropic

Context offloading is an underrated AI use case

Jul 9, 2026

The speaker highlights context offloading as an underrated AI use case, where AI serves as a safety net for routine cognitive tasks like email management and personal finances. Rather than adding new capabilities, AI reduces anxiety about missing important information or making mistakes by handling monitoring tasks, thereby freeing up mental bandwidth.

OpinionDiscussionContext offloadingAI as a safety netReducing cognitive anxiety

WTF is a harness?`

Jul 9, 2026

A harness is code that wraps around an AI agent to make it more effective. The speaker demonstrates how a harness can automatically investigate production issues by analyzing Sentry warnings, identifying root causes, determining impact, and recommending whether to create tickets or apply fixes.

TechnicalInsightfulAI harnesses and their purposeAutomated issue investigationProduction monitoring and Sentry integration

Stop prompting your AI agents. Start managing them.

Jul 8, 2026

The speaker discusses the shift from traditional agent prompting to agent management, highlighting the limitations of local Kanban board approaches and advocating for cloud-based VPS solutions with multiple communication channels to effectively manage autonomous AI agents across projects.

DiscussionTechnicalAgent orchestration and autonomous systemsShift from prompting to managementKanban boards for agent task management

How to build a custom AI harness with Claude SDK

Jul 8, 2026

A custom AI harness is code wrapped around an AI agent to make it more effective for specific workflows. The speaker demonstrates building a Sentry bug-debugging harness using Claude SDK with a terminal UI, showing how structured constraints, custom prompts, and specific tool adapters enable agents to handle complex tasks more efficiently than general-purpose AI tools.

TechnicalInsightfulAI harness definition and purposeCustom prompting and workflow encodingTool adapters and API integration

How I run autonomous coding agents from my phone with OpenAI Symphony + Linear

Jul 6, 2026

Alessio Finelli demonstrates how he uses OpenAI's Symphony framework combined with Linear and Codex to run autonomous coding agents from his phone, managing both software engineering tasks and a Pokémon card trading business through cloud-based VPS infrastructure instead of local machines.

TechnicalDiscussionAutonomous coding agents with OpenAI SymphonyCloud-based VPS infrastructure vs local runtimeLinear as state machine for task management

My taste and the automated benchmark disagreed almost completely

Jul 5, 2026

The speaker discusses discrepancies between their subjective evaluation of model performance and automated benchmark results, noting that different judges (Opus, 4A, 5.5) show varying levels of generosity and bias. They conclude that personal judgment matters significantly and plan to incorporate more subjective taste into evaluation metrics while retiring saturated benchmark tasks.

ResearchOpinionModel judging and evaluation methodologyJudge bias and self-evaluationDiscrepancy between metrics and subjective assessment

Task-by-task model recommendations

Jul 5, 2026

The speaker provides task-specific model recommendations across different use cases, suggesting GPT 5.5 for PRDs, Sonnet 4.6 for prototyping and casual interaction, and Opus 4.8 or Sonnet 5 for codebase work. Model selection varies based on complexity, with Opus 4.8 excelling at dense UI design and Sonnet suitable for simpler implementations.

TechnicalOpinionModel selection by task typeLLM benchmarking and performance evaluationCode and codebase handling

The How I AI Bench

Jul 3, 2026

The speaker introduces 'The How I AI Bench,' a new set of human and AI-graded benchmarks designed to evaluate language models on practical tasks like writing PRDs, solving bugs, and designing systems. They test Claude Sonnet 3.5 against these benchmarks and note that while it scores lower than some specialized benchmarks (69% on Agentic Coding SweetBench Pro, 82% on Terminal Bench 2.1), the difference may not be noticeable in real-world usage.

TechnicalOpinionAI model benchmarking methodologyThe How I AI Bench frameworkClaude Sonnet 3.5 performance evaluation

How a designer became a top engineer

Jul 2, 2026

Katie transitioned from designer to top-performing engineer, ranking in the 94th percentile for code throughput across the entire R&D organization. Her success stemmed from technical curiosity combined with supportive engineer mentors who reviewed her code and helped her improve her craft.

StoryInsightfulDesigner-to-engineer career transitionTechnical mentorship and code review cultureMeasuring engineering productivity through PR throughput

No meetings, no Jira, no text threads... and it shipped anyway.

Jul 1, 2026

A team successfully shipped a project in 10 weeks by eliminating traditional project management structures entirely—no meetings, Jira, documentation, or text communication. Instead, they relied solely on a 24/7 Zoom room where team members could work synchronously and asynchronously as needed.

StoryInsightfulElimination-based process designAsynchronous and synchronous collaborationMinimal documentation practices

Claude automates the busy-work so you can spend more quality time with your kids

Jun 11, 2026

The speaker discusses how Claude's Co-worker feature helps parents automate tedious online administrative tasks, freeing up time for more meaningful interactions with their children. By handling tasks like returns and help emails, AI removes low-value busywork rather than replacing genuine human experiences.

OpinionInsightfulAI-assisted task automationParenting and work-life balanceAdministrative busywork reduction

Use Claude as your personal shopping assistant

Jun 9, 2026

A parent describes using Claude as a household management and shopping assistant to find high-quality, naturally-made products from reputable brands. They created a project in Claude with specific brand criteria and used it to organize notes and vet brands. A key benefit highlighted was Claude surfacing that a previously reputable brand had declined in quality after a corporate takeover.

OpinionInsightfulUsing Claude as a personal shopping assistantEvaluating and vetting brands for quality and longevityOrganizing notes and lists with AI

She built a Claude shopping assistant to stop buying cheap junk

Jun 8, 2026

Nicole Ruiz demonstrates how she built a Claude project to automate high-quality shopping decisions for her family, using curated brand lists and purchasing criteria to filter out cheap, poorly-made products. She also shows how Claude Computer Use helps her draft return emails by pulling order details directly from her Gmail. The system is designed to reduce the mental overhead of conscious consumption so she can spend more time with her children.

InsightfulDiscussionClaude project for high-quality shoppingCurated brand lists and purchasing criteriaAI-assisted product vetting and brand history research

Creating a podcast hype video with Gemini Omni

Jun 7, 2026

The host demonstrates using Google Flow, a generative AI creative suite, to create a hype video for the 'How I AI' podcast. Using a fish-eye lens avatar of themselves, they generate a storyboard with cinematic shots including keyboard close-ups, office wide shots, and a humorous chair spin with a digital heads-up display overlay.

TechnicalFunnyGoogle Flow creative suiteAI-generated hype video productionStoryboard generation with generative AI

She shipped an app to the app store with zero coding knowledge

Jun 5, 2026

A person recounts shipping an app to the app store without learning to code, relying primarily on basic computer skills like copying, pasting, and file labeling. She candidly admits she still doesn't understand the underlying software infrastructure, including the deployment platform Railway, even after successfully completing the project.

FunnyStoryNo-code app developmentDeploying an app without technical knowledgeUsing Replit and Railway as development/hosting platforms

Google Omni made this hype video in less than 15 minutes

Jun 4, 2026

A creator demonstrates using Google Omni to generate a hype/promo video for their podcast 'How I AI' in under 15 minutes. The resulting video includes music, narration, and a polished promotional script. The creator expresses genuine amazement at the output quality.

FunnyTechnicalAI-generated video creationGoogle Omni capabilitiesPodcast promotion
Page 1 of 3Next

Get AI summaries like this delivered to your inbox daily