Welcome back. OpenAI clearly wants to dominate personal AI agents after moves by SpaceXAI and Meta. In a viral interview moment, OpenAI's own CFO accidentally named Meta's Muse instead of their own brand-new agent, Dots. Meanwhile, Anthropic highlights the cyber risks behind open-weight models.
Also: A therapeutic game where you can nuke any webpage, a quick fix to stop Codex from thrashing your SSD limit, and an essential checklist for anyone shipping iOS or Android apps.
Today’s Brief
Recap of OpenAI’s DevDay
Cursor lets you visualize your codebase
A Google dev’s agentic coding workflow (tutorial)
Anthropic’s guide to choosing effort levels in Claude Code (cookbook)

TODAY IN PROGRAMMING
OpenAI has a treat for all developers: The ChatGPT maker just wrapped its biggest event yet, offering 20 major announcements for developers. The main highlights of the events were always-on agent Dots that takes on GrokBot and Muse along with the Decisions API that is meant to work like the new Jev model. There’s a lot to catch up on, and we broke down the biggest launches in the Insight section below.
Anthropic highlights how open-weight models could be a cyber threat: The AI lab tested Chinese GLM-5.3 and found that it can build working cyber-attacks almost as well as Claude Mythos Preview. Keep in mind, Mythos is the frontier model that Anthropic locked down over safety risks as compared to the open-source Chinese model that anyone can use right now. In testing, it dug up zero-day browser bugs and pulled local files straight off a target machine. If you ship code, it’s time to check your defenses, as attackers could already have this in their toolkit. See the findings.
ElevenLabs cuts speech latency to 150ms: The voice AI lab shipped Eleven v4 and v4 Turbo, its fastest speech models yet. Turbo hits a median 150 ms time-to-first-speech, making the awkward pause virtually vanish for callers. It starts streaming audio back before your LLM even finishes generating the sentence, built for real-time agent loops. You get support for 90+ languages alongside Professional Voice Clones, accessible via REST, WebSockets, or the Python and TypeScript SDKs. Find more details on the API here.

PRESENTED BY CONSCIUM
What happens when your AI agent goes off script?
Your AI agent passes its tests. Then a real user does something you didn't expect.
Simulate real users, edge cases and messy inputs.
See exactly where your agent breaks.
Get recommended fixes in minutes.
Connect your agent and see how it holds up under pressure.

INSIGHT
Everything OpenAI announced at DevDay 2026
OpenAI dropped a massive stack of updates at DevDay, and we spent the hours after tearing through the livestream, tracking dev reactions on X, and reading/rereading the docs. Here are the big releases actually worth your build time:
GPT-6.1 Sol: Easily the biggest announcement of the event. OpenAI is delivering near-Astra intelligence at one-fifth the standard token price ($2 in / $10 out per million), with cached input at just $0.10 per million which is half of GPT-6 Sol's cached rate.
Codex Cloud: Your agent can finally keep coding around the clock in the background. Reusable cloud environments keep your repositories, dependencies, and settings ready to go. When you need to check progress or change direction, you can manage the entire workflow right from your phone.
Codex CLI, Code Review, Security Cloud: Managing your agents gets a lot easier with new voice controls and a dedicated /agents dashboard. Plus, cloud reviews and security scans can now track down bugs and spin up fixes while you step away.
Ultrafast: OpenAI’s new $500/month Pro 500 plan includes Ultrafast which lets Codex and Work generate up to 300 tokens per second, or 8x faster. So you spend way less time watching your agent type.
Decisions API: Well, Jev is already facing serious competition. OpenAI's Luna-powered API selects from options you define, leaning on text or image context to help you make calls fast. Think request routing or deciding your agent's next move. It is in limited preview right now.
Agents API, Bedrock Managed Agents: If you’re building an agent that needs to use other software, you can now add computer use, with orchestration and context handling managed for you. There’s an AWS option, too.
Banked reset: Think of it as a one-time reset for your Codex limits that you can stash away until you really need it. Simon Willison noted that OpenAI rolled this out globally during DevDay, so check your dashboard to see if you qualify and when it expires.

IN THE KNOW
What’s trending on socials and headlines
Break it Down: See how a query runs through SQLite's codebase in this 7-minute explainer built entirely with Opus 5.5. Every open-source repo needs one (1.2K bookmarks).
URL Wrecker: A viral game lets you point a stickman at any website and destroy it with flamethrowers, grenades, and pure chaos. Devs are loving it (5M views).
Silent Killer: Did you know Codex is constantly writing diagnostic logs and eating up your disk space? Run this single command to kill it (1.1M views).
Bot Squad: SpaceXAI’s lead designer and engineer behind Grok Bot revealed the 14 bots running their work and life. One books trips, another ships code before review.
Cursor Charts: Cursor now builds charts and diagrams right inside the chat. One new command does the whole thing, and it's already live (2.8K likes).

TOP & TRENDING RESOURCES
Top Tutorial
The ideal agentic coding workflow (by a Google engineer): This tutorial walks through the full setup, from giving agents the right memory and context to planning before they code, running multiple agents in parallel with Git worktrees, and reviewing AI-generated code based on risk. You’ll learn how to turn one-off prompts into repeatable loops that can research, build, test, fix, and review with less babysitting.
Top Tool
TapKit: Give your AI agent control of a real iPhone from your Mac. It can see the screen, tap, swipe, type, open apps, handle 2FA, use iCloud, and even make purchases through apps, turning the phone into another computer your agent can operate.
Trending Cookbook
Choosing the right effort level in Claude Code (by Anthropic): This guide shows when spending more compute actually helps. Use low for quick sketches and easy edits, medium for everyday engineering, high when testing and edge cases matter, and max for hard autonomous work.

AI CODING HACK
How to speed up Claude Code with six lines in CLAUDE.md
Claude Code pads sessions with features, tests, and refactors nobody asked for, then reports done without checking its work. A dev shared six CLAUDE.md lines that pair Sonnet 5.5 with Jev, TypeSafe's decision model, and posted a side-by-side timer showing a 4x speedup.
Step 1: Install the Jev integration. It requires a TypeSafe API key and Node 20+
Step 2: Paste the full prompt into your project's CLAUDE.md.
Step 3: Restart Claude Code.
Jev now picks each session's model and effort, and Claude stops before adding work you didn't request. Rules 5 and 6 improve sessions even without Jev connected.

IN CASE YOU MISSED IT
Our most-clicked story from yesterday
An ex-Meta engineer shared her method for becoming an AI engineer in 2026, no degree required, covering skills, projects, and proof.
Grow customers & revenue: Join companies like Google, IBM, and Datadog. Showcase your product to our 350K+ engineers and 150K+ followers on socials. Get in touch.
Whenever you're ready to dive deeper
We put together a few guides on coding agents, agentic engineering, and leadership frameworks to help you level up in your career. Browse all our guides.
What did you think of today's newsletter?
You can also reply directly to this email if you have suggestions, feedback, or questions.
Until next time — The Code team






