Using Claude Code: The Unreasonable Effectiveness of HTML This is now also on the [Claude Blog](https://claude.com/blog/using-claude-code-the-unreasonable-effectiveness-of-html). Markdown has become the dominant file format used by agents to communicate with us. It’s simple, p…
New Anthropic research: Natural Language Autoencoders. Models like Claude talk in words but think in numbers. The numbers—called activations—encode Claude’s thoughts, but not in a language we can read. Here, we train Claude to translate its activations into human-readable text…
Introducing GPT-Realtime-2 in the API: our most intelligent voice model yet, bringing GPT-5-class reasoning to voice agents. Voice agents are now real-time collaborators that can listen, reason, and solve complex problems as conversations unfold. Now available in the API along…
Codex now works directly in Chrome on macOS and Windows. It’s even better at working with apps and sites in Chrome, and now works in parallel across tabs in the background without taking over your browser. To get started, install the Chrome plugin in the Codex app. https://t.c…
New Anthropic research: Teaching Claude why. Last year we reported that, under certain experimental conditions, Claude 4 would blackmail users. Since then, we’ve completely eliminated this behavior. How?
Just gonna leave this here. https://t.co/EOI980j9e9 https://t.co/HxbCl2Izcz
React Doctor v2 is here Your agent writes bad React code, this catches it Works with Next.js, Vite, React Native. Fix your app in minutes npx react-doctor@latest https://t.co/SEFExJ2EWU
Last week we shipped 50+ Claude Code reliability fixes. This week it's 60+ more. Smoother long-running sessions, a more efficient agent loop, auth that works in more environments, and terminal fixes: 🧵
Introducing the Printing Press, a CLI-factory and a CLI-library. Built with @trevin. 🏭🖨📚 Most APIs suck for agents. Most MCPs suck for agents. Most official CLIs suck for agents. They waste tokens and time. @steipete started making his own because of this. 📚 A Library of a…
a little project i've been hacking on: https://t.co/zTuwWy44ly bugs expected. more topics soon.
The more I replace plans with prototypes, the better the outputs Who'd have thought that low fidelity prototypes were better than walls of spec Oh yeah, the entire industry for 20 years Stop going against decades of knowledge because someone in SF shipped it as a 'mode'
Multiple security vulnerabilities affecting React Server Components and Next.js have been disclosed. We strongly recommend updating your applications immediately. Cloudflare WAF managed rules already mitigate the disclosed denial-of-service vulnerabilities, and we are investiga…
An update regarding the future at @Cloudflare. I’ve shared my full message to the team and details on the support we're providing those departing here: https://t.co/8djT55aVSP
The new @digg alpha is coming soon. First up: AI news. 9M+ graph connections. 15+ AI judges. Real-time X ingestion. Sentiment analysis, clustering, and signal detection built to surface what actually matters. https://t.co/sYSAnVDxWE
Real-time World Models are the next AI frontier. Today, we're taking the first step towards this reality: our early preview lets you experience worlds generated in real-time, running on our global low-latency infrastructure. Try it now: https://t.co/h0XDYsHcGB https://t.co/DJP…
"Technical writing completely changed my life." - @trq212 In under 2 years, Thariq (@AnthropicAI) cracked the code on writing technical articles that consistently hit 1M+ views. In this 20-min workshop, he breaks down: → his exact writing workflow → the tactics behind articl…
small ship / passion project, more details soon https://t.co/Njxmsg9ISm 1. call responses via cli with all cloud tools 2. unix style structured outputs via cli 2. image gen/edit, transcription, tts 3. make projects and provision api keys more docs soon
I gave a viral talk recently, and @swyx asked me to put something together to explain how I did it - to help future AIE speakers and anyone who wants to learn. I am, oddly, extremely qualified to do this because I spent 6 years as a voice coach. So I've not only given countless…
Strong Opinions, Loosely Held on Agent + Harness Engineering: 1. You can outperform any default harness+model (including codex & claude code) on pretty much any Task by engineering the harness around it. Using the exact same model, curate prompts, tools, skills, hooks for that…
We partnered with @PrimeIntellect to build Fast Ask, a small RL-trained subagent that helps our Sheets agent find answers in spreadsheets. It scores +4% over Opus on exact match accuracy at Haiku latency. https://t.co/GJQvHJjABl
Ranked by when the bookmark was saved during this London date range. Interactions are likes, reposts, replies, and quotes combined. Metrics are X's latest public counts; older items can be stale once they leave the recent-bookmarks sync window.