Claude Daily Claude Daily
RSS
Riley Brown 22 min video 3 slides

An Insane Week for AI Agents — GLM 5.2, Codex Skills, Claude, and Cursor

AI Agents Just Changed Forever: GLM 5.2, Codex Skills, Claude & Cursor
25,900 views 7 highlights

TL;DR · What you'll learn

  • 1 It's been an insane week for AI agents — latest Claude Fable 5 news, a new Codex feature that turns screen recordings into skills, the world's best open-source model, and the SpaceX × Cursor acquisition all hit at once.
  • 2 China's Z.ai released GLM 5.2 — an open-source model that matches GPT 5.5 and Opus 4.8 at roughly 1/5 to 1/6 the cost of frontier models. Hooks into Cursor in seconds.
  • 3 On benchmarks GLM 5.2 holds its own with Opus and GPT 5.5. The author's read of the field: Fable leads, GPT 5.5 next, Opus trails slightly — and now an open-source contender has pulled even with that pack.
  • 4 Codex's new 'record and replay' is the standout feature. Record yourself doing something on screen, hand it to Codex, and turn it into a skill. The demo records uploading a video as a Typefully draft and replays it with Codex's computer use.
  • 5 During the demo, Typefully blocked the upload because the file was over 512MB, but the click-through and upload flow itself worked. Recording → skill → autonomous execution is now a real, runnable loop.
  • 6 More Chinese open-source models are on the way, and Gemini is expected to announce both a new model and direction on its 'super app.' Strategy alignment across major labs is the next thing to watch.
  • 7 The author calls out Google directly: 'too many products, I can't tell anyone which one to use.' Antigravity, AI Studio, Jules, Gemini Desktop — Google needs to pick *one* super app to be competitive again.

Read as slides

3 slides total

01 Slide 1 / 3
Watch at 00:40

GLM 5.2 marks an inflection point for open-source AI

The biggest story of the week was Z.ai's GLM 5.2. It matches GPT 5.5 and Opus 4.8 in capability while costing roughly 1/5 to 1/6 of frontier models. The whole thing is open-source, and integrating it into Cursor takes seconds.

The author maps the current frontier as Fable on top, GPT 5.5 close behind, and Opus trailing slightly. The real shock is that GLM 5.2 has now pulled level with that group. The single-frontier-leader narrative is collapsing in real time — both benchmarks and live use show the open-source side genuinely catching up.

Claude Daily 01 / 03
02 Slide 2 / 3
Watch at 10:17

Codex's 'record and replay' — turning screen actions into skills

On the Codex side, 'record and replay' is a serious step forward. Codex's computer use feature lets you record a screen workflow and register it as a skill. After that, Codex can call the skill and execute the workflow autonomously.

The live demo recorded 'upload my latest video to Typefully as a draft' as a skill, then asked Codex to run it via computer use. Recordings up to 30 minutes are supported, so even long workflows can be captured. Typefully's 512MB upload limit blocked the actual file in the demo, but the click-through, navigation, and upload flow on Comet executed correctly. The distance between 'record once' and 'automate from now on' just got dramatically shorter.

Claude Daily 02 / 03
03 Slide 3 / 3
Watch at 20:58

The super-app race and a pointed message for Google

Looking forward, more Chinese open-source models are coming, and Gemini is expected to announce both a new model and a direction for its super app. Strategic alignment across the big labs is the thing to watch next. The author argues we've already been in the 'super app era' for 100 days.

The sharpest comment goes to Google. As a content creator, the author can't recommend Antigravity vs AI Studio vs Jules vs Gemini Desktop because there's no clear answer. Google has too many products, none of which feels like the super app. Picking *one* and committing is the prerequisite for Google getting back into real competition — that's the takeaway.

Claude Daily 03 / 03

Editor's Take

The most operationally significant item in the roundup is Codex's 'record and replay.' If you can record screen actions, turn them into a skill, and auto-execute, the barrier to automation drops from writing code to 'demonstrating the steps once.' Combined with GLM 5.2 nearing the frontier at a fifth to a sixth of the cost, the democratization of agent building is advancing on two axes at once — ease and affordability.

Source

AI Agents Just Changed Forever: GLM 5.2, Codex Skills, Claude & Cursor

Riley Brown

AI Agents Just Changed Forever: GLM 5.2, Codex Skills, Claude & Cursor

Published 6/21/2026 22 min 25,900 views

This article auto-summarizes the YouTube video's transcript with Claude. Please refer to the original video for nuance and exact wording.

Watch on YouTube

Related

3 articles