The White House rethinks its Anthropic fight
The White House is reversing course on its conflict with Anthropic, seeking greater access to its Mythos AI model for national security purposes while limiting broader private sector access. Meanwhile, Google is rolling out Gemini AI to vehicles, and OpenAI traced ChatGPT's unusual 'goblin obsession' to a single reward signal in its 'Nerdy' personality preset.
Summary
The central story involves a shifting dynamic between the White House and Anthropic over the Mythos AI model. Originally, the government escalated tensions with Anthropic, but the emergence of Mythos's powerful cyber capabilities has complicated the feud. Anthropic sought to expand private sector access to Mythos from approximately 50 firms to nearly 120, but U.S. officials pushed back citing compute strain concerns related to government use. A forthcoming White House AI memo is expected to push multi-vendor AI adoption for agencies and address some of Anthropic's grievances. Despite an ongoing legal battle, the memo may allow agencies to work around the supply chain risk designation. Notably, former AI czar David Sacks indicated that GPT-5.5 has reached cyber capabilities similar to Mythos, and predicted all frontier models will reach that level within six months. Internal division within the administration is evident, with Secretary of War Pete Hegseth publicly calling Anthropic's leadership an 'ideological lunatic,' even as other officials appear to want a reconciliation driven by access needs.
On the consumer technology front, Google announced the rollout of Gemini AI to vehicles with Google built-in, replacing the older Google Assistant. The upgraded system handles navigation, messaging, music, vehicle controls, and car-specific queries drawn from manufacturer manuals. A beta Gemini Live mode supports open-ended conversations, with future integrations planned for Gmail, Calendar, and Google Home. General Motors announced the feature for approximately 4 million of its vehicles from model year 2022 onward, with the U.S. rollout coming first.
OpenAI researchers traced ChatGPT's widely noticed habit of inserting goblins, gremlins, and fantasy creatures into responses to a single reward signal embedded in the 'Nerdy' personality preset. Following ChatGPT-5.1's November launch, 'goblin' mentions in user conversations jumped 175%, with similar spikes for related creatures. Fine-tuning loops recycled creature-favored outputs back into the model's default behavior, spreading the quirk beyond just Nerdy users. OpenAI retired the Nerdy preset in March and shipped GPT-5.5 with an explicit prompt banning goblins, gremlins, ogres, trolls, raccoons, and pigeons.
The newsletter also covered several other developments: Meta opened its ads platform to third-party AI tools via a new MCP server; OpenAI surpassed its 2029 Stargate compute goal of 10 GW ahead of schedule; Elon Musk admitted during trial testimony that xAI used distillation techniques to train on OpenAI models; and Anthropic launched a public beta for Claude Security, an enterprise tool for scanning and patching code vulnerabilities.
About this episode
PLUS: Stress test business ideas with Perplexity
Key Insights
- The White House appears to be softening its stance against Anthropic primarily because it wants greater government access to the powerful Mythos model, suggesting national security interests are overriding ideological opposition.
- Former AI czar David Sacks claimed that GPT-5.5 has already reached cyber capabilities comparable to Mythos, and predicted all frontier AI models will reach that level within six months, implying rapid capability convergence across providers.
- OpenAI's investigation found that a reward signal in a single personality preset ('Nerdy') was responsible for ChatGPT's global goblin-insertion behavior, demonstrating how fine-tuning loops can propagate unintended patterns across an entire model's default behavior.
- Elon Musk testified during his trial against OpenAI that xAI used distillation techniques to train on OpenAI models, a significant admission about the competitive and legally contentious practice of model distillation.
- Internal division within the Trump administration over Anthropic is stark: while some officials seek reconciliation to secure model access, Secretary of War Pete Hegseth publicly called Anthropic's leadership an 'ideological lunatic,' indicating no unified government position exists.
Topics
Transcript
Good morning, {{ first_name | AI enthusiasts }}. The government spent months escalating its fight with Anthropic. Then Mythos showed up with cyber capabilities powerful enough to make the feud look a lot less simple. The White House is now trying to thread an awkward needle: keep the model close for national security, limit who else can use it, and avoid looking like it is fully backing down from the Pentagon's hard line all at the same time. The White House’s Anthropic stance gets complicated Gemini comes into Google-powered cars Stress test business ideas with Perplexity OpenAI finds source of ChatGPT's goblin obsession 4 new AI tools, community workflows, and more ANTHROPIC VS. THE WHITE HOUSE Image source: Images 2.0…
Full transcript available for MurmurCast members
Sign Up to AccessMore from The Rundown AI
Economists, researchers put AI’s job shock on the clock
Over 200 AI researchers and Nobel laureates signed a Stanford statement warning that AI could displace jobs at historic scale within the next decade, requiring immediate government action on safety nets and labor policy. Meanwhile, the AI industry continues to evolve with new tools, research findings on AI personality variations, and ongoing feuds between major figures like Musk and Altman.
Apple takes OpenAI to court
Apple has filed a lawsuit against OpenAI alleging the company poached over 400 Apple employees and used them to steal confidential hardware secrets for an unreleased device designed by Jony Ive. The newsletter also covers new AI tools, training resources, and research on long-term AI scenarios.
OpenAI sends GPT-5.6 to Work
OpenAI launched GPT-5.6 with three tiers (Sol, Terra, Luna), featuring near-Fable performance at lower costs and introducing ChatGPT Work with desktop integration. The newsletter also covers Meta's Muse Spark 1.1 release and argues that AI will concentrate demand among top 1-5% professionals while displacing mid-tier service providers.
SpaceXAI, Cursor release the strongest Grok yet
SpaceXAI and Cursor released Grok 4.5, a powerful new AI model that combines strong performance with faster speeds and lower costs than competitors like Claude and GPT. The newsletter also covers OpenAI's new GPT-Live voice model, ByteDance's Seedream 5.0 Pro image generation tool, and several other AI developments in the industry.
Meta climbs the AI image leaderboard
Meta released Muse Image, an in-house AI image model ranking No. 2 on leaderboards behind only OpenAI's GPT Image 2, with a teased video model also performing strongly. Meanwhile, Beijing is considering restrictions on Chinese AI model exports, creating potential geopolitical reciprocity risks similar to U.S. export controls, while companies like DoorDash demonstrate that pairing multiple AI models significantly improves code review accuracy and cost-effectiveness.