The Ox Alpha mystery ends with Z.ai
Z AI revealed that Ox Alpha, the mystery AI model that dominated rankings last week, is actually its new GLM-5.3-Flash model, priced at a tenth of competitors and run entirely on Chinese-made chips. OpenAI leadership declared AGI arriving by year-end while emphasizing safety concerns following a security breach.
Summary
The Rundown newsletter breaks down several major developments in the AI landscape. The primary story is the resolution of the Ox Alpha mystery: Z AI confirmed the anonymous model that captivated the internet is GLM-5.3-Flash, launched with open weights and exceptional pricing ($0.045 per task on Artificial Analysis's Intelligence Index). The model achieved the top position on OpenRouter by doubling the second-place DeepSeek's numbers—a platform record. Notably, Z AI ran this massive usage week entirely on Chinese-made chips, potentially solving a critical bottleneck in China's AI infrastructure by achieving token costs comparable to Nvidia hardware.
In a separate major development, OpenAI leadership made bold AGI claims in a TIME profile. Sam Altman dated artificial general intelligence to year-end, with research chief Mark Chen claiming the lab is "80% of the way" to AGI. Chief scientist Jakub Pachocki highlighted that their Astra model can conduct a week of researcher work solo, positioning it as the first model to "actually invent new things in a way that matters." However, the narrative includes a cautionary note: OpenAI is slowing frontier development and refocusing on safety following July's Hugging Face security breach, which the lab called a "warning shot" demonstrating loss-of-control risks.
The newsletter also features Nate Grehek's insights on small business AI adoption, emphasizing that startups and small companies currently enjoy a 20x cost advantage on frontier models compared to enterprises. He recommends a two-track approach: having AI-enthusiasts build custom tools while automating mundane tasks for those uninterested in AI. Additional coverage includes practical guides on ChatGPT Work projects, community workflows (like Brian Carey's SharePoint document management system), and emerging tools like Memoket's wearable AI recording device and Yutori's Navigator n2 computer-use model.
About this episode
PLUS: A beginner’s guide to ChatGPT Work
Key Insights
- Z AI demonstrated that Chinese-made chips can achieve token costs competitive with Nvidia hardware, potentially resolving a major infrastructure bottleneck for China's AI development.
- OpenAI's research chief Mark Chen claims the company is 80% of the way to AGI, with CEO Sam Altman predicting an internally-defined AGI model will exist by year-end 2024.
- Small businesses currently enjoy approximately 20x lower per-token costs on frontier AI models compared to enterprises, creating a structural competitive advantage for rapid experimentation and deployment.
- OpenAI is deliberately slowing frontier model development and implementing automatic shutdown systems for rogue agents following a July security breach it characterized as a 'warning shot' about loss-of-control risks.
- Chief scientist Jakub Pachocki claims Astra can conduct a week of independent researcher work, positioning AI model capability to autonomously generate new knowledge as a critical AGI milestone.
Topics
Transcript
Good morning, {{ first_name | AI enthusiasts }}, and welcome to our 2,833 new readers. The anonymous free AI that had developers playing detective all weekend belongs to China's Z AI, and the case of Ox Alpha is officially closed. The impressive preview launches as GLM-5.3-Flash, arriving with open weights, pricing near a tenth of comparable rivals, and a record week of usage that the lab says ran entirely on Chinese-made chips. Reminder: Our next live workshop is today at 2 PM EST. Join and learn how to build your first AI evaluation to measure AI quality, compare models, and trust workflows with The Rundown University AI Educator Nate Grehek. RSVP here . The Ox Alpha mystery model ends with…
Full transcript available for MurmurCast members
Sign Up to AccessMore from The Rundown AI
OpenAI’s “generational leap” with GPT-6 Astra
OpenAI released GPT-6 Astra, positioning it as a major advancement in AI with exceptional benchmark performance across multiple domains. The newsletter also covers Google's improved weather forecasting model, the Loop Method for ChatGPT optimization, and a reader's positive-news-only AI app.
Meta, Google join the AI launch party
Meta and Google launched new AI models in early September, with Meta's Muse Spark 1.3 achieving near-frontier performance at low cost while Google's Gemini 3.8 Flash represents a recovery step but still trails the frontier. The newsletter also covers AI safety concerns about reasoning transparency, tech literacy as a career ceiling, and practical AI workflows for professional development.
Fable 5.1 kicks off launch week at the frontier
Anthropic released Claude Fable 5.1, showing significant improvements in coding and research tasks with reduced safety rejections, marking the end of the summer's cautious release period as OpenAI's Astra launch approaches. Bernie Sanders published an op-ed calling for a global AI pause, citing control concerns and societal risks, while ongoing legal battles between Apple and OpenAI center on alleged theft of confidential designs.
Runway's Solaris previews the no-code internet
Runway unveiled Solaris, an AI-powered interface that renders websites and apps as real-time video with no underlying code, while Imperial College researchers developed an AI model that detects heart disease from ECGs in under two seconds with superior accuracy to human doctors.
OpenAI cuts out SpaceX-owned Cursor
OpenAI is removing its models from Cursor coding editor by mid-November following SpaceX's acquisition of the platform, citing Elon Musk's history of contract violations as justification. The move escalates the ongoing feud between Sam Altman and Elon Musk while putting developers in the middle of a high-profile corporate conflict.