An Insane Week for AI Agents — GLM 5.2, Codex Skills, Claude, and Cursor
TL;DR · What you'll learn
- 1 It's been an insane week for AI agents — latest Claude Fable 5 news, a new Codex feature that turns screen recordings into skills, the world's best open-source model, and the SpaceX × Cursor acquisition all hit at once.
- 2 China's Z.ai released GLM 5.2 — an open-source model that matches GPT 5.5 and Opus 4.8 at roughly 1/5 to 1/6 the cost of frontier models. Hooks into Cursor in seconds.
- 3 On benchmarks GLM 5.2 holds its own with Opus and GPT 5.5. The author's read of the field: Fable leads, GPT 5.5 next, Opus trails slightly — and now an open-source contender has pulled even with that pack.
- 4 Codex's new 'record and replay' is the standout feature. Record yourself doing something on screen, hand it to Codex, and turn it into a skill. The demo records uploading a video as a Typefully draft and replays it with Codex's computer use.
- 5 During the demo, Typefully blocked the upload because the file was over 512MB, but the click-through and upload flow itself worked. Recording → skill → autonomous execution is now a real, runnable loop.
- 6 More Chinese open-source models are on the way, and Gemini is expected to announce both a new model and direction on its 'super app.' Strategy alignment across major labs is the next thing to watch.
- 7 The author calls out Google directly: 'too many products, I can't tell anyone which one to use.' Antigravity, AI Studio, Jules, Gemini Desktop — Google needs to pick *one* super app to be competitive again.
Read as slides
3 slides total
GLM 5.2 marks an inflection point for open-source AI
The biggest story of the week was Z.ai's GLM 5.2. It matches GPT 5.5 and Opus 4.8 in capability while costing roughly 1/5 to 1/6 of frontier models. The whole thing is open-source, and integrating it into Cursor takes seconds.
The author maps the current frontier as Fable on top, GPT 5.5 close behind, and Opus trailing slightly. The real shock is that GLM 5.2 has now pulled level with that group. The single-frontier-leader narrative is collapsing in real time — both benchmarks and live use show the open-source side genuinely catching up.
Codex's 'record and replay' — turning screen actions into skills
On the Codex side, 'record and replay' is a serious step forward. Codex's computer use feature lets you record a screen workflow and register it as a skill. After that, Codex can call the skill and execute the workflow autonomously.
The live demo recorded 'upload my latest video to Typefully as a draft' as a skill, then asked Codex to run it via computer use. Recordings up to 30 minutes are supported, so even long workflows can be captured. Typefully's 512MB upload limit blocked the actual file in the demo, but the click-through, navigation, and upload flow on Comet executed correctly. The distance between 'record once' and 'automate from now on' just got dramatically shorter.
The super-app race and a pointed message for Google
Looking forward, more Chinese open-source models are coming, and Gemini is expected to announce both a new model and a direction for its super app. Strategic alignment across the big labs is the thing to watch next. The author argues we've already been in the 'super app era' for 100 days.
The sharpest comment goes to Google. As a content creator, the author can't recommend Antigravity vs AI Studio vs Jules vs Gemini Desktop because there's no clear answer. Google has too many products, none of which feels like the super app. Picking *one* and committing is the prerequisite for Google getting back into real competition — that's the takeaway.
Editor's Take
The most operationally significant item in the roundup is Codex's 'record and replay.' If you can record screen actions, turn them into a skill, and auto-execute, the barrier to automation drops from writing code to 'demonstrating the steps once.' Combined with GLM 5.2 nearing the frontier at a fifth to a sixth of the cost, the democratization of agent building is advancing on two axes at once — ease and affordability.
Source
Riley Brown
AI Agents Just Changed Forever: GLM 5.2, Codex Skills, Claude & Cursor
This article auto-summarizes the YouTube video's transcript with Claude. Please refer to the original video for nuance and exact wording.
Watch on YouTube →Related
3 articles
AI News & Strategy Daily | Nate B Jones
Claude Fable 5 Bossed 20 Cheap AI Agents to Build a Whole Website for $8 -- and Caught Its Own Cheating Along the Way
A swarm of roughly 20 cheap AI agents, orchestrated by Fable, rebuilt the creator's wife's website in one hour -- better than six days of hands-on work with Codex the month before.
AIで創作効率化とマネタイズby嗚呼蛙
Seven Codex Features Worth Knowing, Even on the Free Tier: From Browser Control to Phone Access
With Fable 5's return generating buzz for Claude Code, this video revisits seven useful features of Codex, OpenAI's agent, usable even without paying.
Vaibhav Sisinty
18 AI Updates This Week: A Free Tool That Merges ChatGPT and Claude, and the Truth Behind Meta's Claude Code Ban
A tool called Hermes added a 'mixture of agents' feature that runs two AI models at once and merges their best answers, scoring 8% above Opus alone and 11% above GPT 5.5 alone on its own benchmark.