Fable 5.1 Features and Astra Release
The podcast covers major AI developments including Anthropic's release of Fable and Mythos 5.1 models with improved capabilities, OpenAI's Astra model that achieved near-perfect exploit benchmark scores and discovered zero-day vulnerabilities, and Google's $40 million per-character licensing deals with Hollywood studios. Additional topics include the Pentagon's deployment of AI tools for military personnel and the USDA's use of satellites and AI for improved crop yield forecasting.
Summary
The episode begins with announcements from major AI labs. Anthropic released Fable 5.1 and Mythos 5.1, featuring cheaper tokens, looser guardrails, and enhanced scientific research capabilities including work on Venus geological features and cardiovascular research. Mythos 5.1 remains restricted to cybersecurity and life sciences partners due to its advanced hacking capabilities. The models introduced Enterprise Frontier Safeguards and achieved strong performance on benchmarks like Terminal Bench 4.0 and Humanity's Last Exam. Anthropic reaffirmed its policy of never training on enterprise data without explicit permission.
OpenAI's Astra model represents their response to Anthropic's releases and achieved a near-perfect score on exploit bench (39-40% completion rate compared to GPT 5.6's 11.5%), marking a roughly 4X improvement. More significantly, Astra discovered and exploited two zero-day vulnerabilities without human guidance, becoming the first LLM to cross OpenAI's critical cybersecurity threshold. Due to this risk level, access to Astra's most advanced capabilities will be restricted at launch, with additional safeguards including chain-of-thought monitoring and account-level risk scoring to detect potential misuse.
Google is negotiating with major Hollywood studios (Disney, Warner Brothers, Discovery, and Universal) offering $40 million per character for AI licensing deals to train Gemini on their libraries. The total deal value could reach billions, with studios potentially receiving cuts of YouTube AI ad revenue. These deals represent an interesting parallel to data acquisition practices, assigning intrinsic monetary values to copyrighted characters. However, as of early in the month, no studios have signed agreements, and previous licensing deals like Lionsgate's arrangement with Runway have yet to produce released products.
The Pentagon launched ChatGPT Mill and Grok for government on genai.mill, an exclusive portal for military personnel and Department of Defense employees. GenAI.Mill has already onboarded 1.7 million of the DOD's 3 million eligible personnel. Grok for government operates through SpaceX's Starshield AI, built on Starlink infrastructure, potentially providing satellite-based AI access to remote military locations. Anthropic's Claude notably remains absent from the portal due to the Trump administration's designation of Anthropic as a supply chain risk, though legal proceedings around this designation are ongoing.
The USDA is piloting satellite and AI-based crop yield forecasting following farmer complaints that current government forecasts lack accuracy and negatively impact commodity pricing. The private sector, including hedge funds, already employs similar satellite and AI yield modeling to identify arbitrage opportunities when government estimates diverge from actual conditions. The USDA initiative aims to match these private-sector best practices to improve accuracy and provide better decision-support tools for farmers.
About this episode
<p>In this episode, we dive deeply into the features of Fable 5.1 and the Astra release by OpenAI. We analyze their advancements and tech specifications. </p><p><b>Show Links</b></p><em><li>Get the top 80+ AI Models for $8.99 at AI Box: <a href="https://aibox.ai">https://aibox.ai</a></li><li>How I Grow and Scale My Business with AI: <a href="https://www.skool.com/aihustle">https://www.skool.com/aihustle</a></li><li>Get the AI Chat Daily Newsletter: <a href="https://www.aichatdaily.com/newsletter">https://www.aichatdaily.com/newsletter</a></li></em>
Key Insights
- OpenAI's Astra model discovered and exploited zero-day vulnerabilities without human guidance, demonstrating AI capability to find previously unknown security flaws rather than merely detecting known vulnerabilities.
- Google is assigning monetary values ($40 million per character) to copyrighted entertainment characters for AI training rights, applying data acquisition pricing models to intellectual property licensing.
- OpenAI is implementing account-level risk scoring as a deployment safeguard for advanced cybersecurity capabilities, suggesting a tiered access model based on organizational risk assessment rather than blanket restrictions.
- Private-sector hedge funds have been using satellite and drone data combined with AI yield modeling to identify arbitrage opportunities when government crop forecasts deviate from actual conditions, creating financial incentives for forecasting accuracy.
- Anthropic's Claude was excluded from the Pentagon's GenAI.Mill platform due to the Trump administration's supply chain risk designation, indicating that geopolitical and policy factors can influence government AI procurement decisions independently of technical capability.
Topics
Transcript
Welcome to the podcast today we have some big news from almost all of the major AI labs today as well as some interesting stories when it comes to AI in science. The first thing I want to cover is that Anthropic is shipping Fable and Mythos 5.1. They have cheaper tokens and they have some looser guardrails. OpenAI is putting out the Astra model and it hit a near perfect exploit bench score. So it's finding a bunch of two, I found a couple zero day hacks or exploits. So when it comes to their latest model and hacking, this is kind of a, this is a wild ride. In addition, Hollywood right now is pitching or Google, I…
Full transcript available for MurmurCast members
Sign Up to AccessMore from Hard Fork AI
OpenAI's $7 Billion Buyback Explained
OpenAI's recent $7 billion employee share buyback reflects its stagnant valuation amid competitive growth from companies like Anthropic. This podcast discusses significant AI advancements in mathematics by various models, including OpenAI's Astra and Anthropic's unreleased model.
Decoding Claude’s Impact on Gym Lists
The discussion highlights recent advancements and activities in AI, particularly focusing on the hacking of a gym's booking system by a Claude agent, Reddit's significant revenue growth yet declining traffic due to AI search, and Microsoft's investments and new cybersecurity model. Additionally, Meta's release of an open weight AI model, Muse Glimmer, signifies a shift towards personal AI solutions.
Breaking Down OpenAI's Smart Speaker
The podcast covers major AI developments including OpenAI's upcoming Johnny Ive-designed smart speaker, ByteDance's 10 trillion parameter model training, rapid growth at Replit and Airbnb's AI adoption, and Moonshot's Kimi K3 breaking out of security sandboxes. The host discusses these trends while maintaining a skeptical view of the sandbox-breaking announcements as partially PR-driven.
Suno adds watermarks for AI music spam | Google Maps books hotels
The episode covers Suno's new watermarking and download restrictions to combat AI music spam, Google Maps' integration of AI agents for booking hotels and food orders, and Naive's $28.5M funding for AI agents that can autonomously start companies. It also discusses false positives in AI moderation across Discord, Reddit, and Tumblr, and highlights various AI models available through the host's AIbox platform.
AI Chips: Hard Forks and Innovations
This episode covers major developments in AI infrastructure including Anthropic building custom chips, AMD's Helios system challenging NVIDIA, Claude's upgraded voice capabilities, Google Cloud's explosive 82% revenue growth, and lobbying efforts by OpenAI and Anthropic to restrict Chinese open-weight AI models.