A frontier without an ecosystem is not stable I’ve been thinking a lot about the future of the firm in an AI-driven economy. This transition is different than any previous platform shift. In the past, we used digital systems to enhance human capital. This is the first time we …
We're launching code storage and git hosting. Origin gives teams and agents a place to host, review, and collaborate on code. Available this fall. Join the waitlist. https://t.co/uamaIarJXY
New in Claude Design: it stays on brand with your design system across projects, lets you edit directly on the canvas, syncs with Claude Code, and connects to more of the tools you already use. https://t.co/MK8YvLP8zV
Introducing eve, an agent framework. 𝚊𝚐𝚎𝚗𝚝/ 𝚊𝚐𝚎𝚗𝚝.𝚝𝚜 𝚒𝚗𝚜𝚝𝚛𝚞𝚌𝚝𝚒𝚘𝚗𝚜.𝚖𝚍 𝚝𝚘𝚘𝚕𝚜/ 𝚜𝚔𝚒𝚕𝚕𝚜/ 𝚜𝚊𝚗𝚍𝚋𝚘𝚡/ 𝚜𝚌𝚑𝚎𝚍𝚞𝚕𝚎𝚜/ Like Next.js, for agents. https://t.co/ezUIGLkSqG
Reminder that you can use the Codex App, CLI and SDK with any open source model, not just with OpenAI models. https://t.co/spPifB4ck3
In light of what happened, I'm doubling down on skills like /improve. A frontier model got pulled. If it happened once, it's gonna happen again. Fable today. 4.9 tomorrow or maybe gpt 6 one day. So, treat intelligence as borrowed. Drain intelligence when it's available. Build…
It seems a mistake to call oneself a "non-technical founder." You're treating not knowing how to do something as a part of your identity. Surely it's better just to fix that.
Our latest economic research introduces a framework for tracking Claude Code as it scales. Who is using Claude Code, and what are they using it for? How is the value of tasks changing? And how much does domain expertise shape whether a session succeeds? https://t.co/IjjwQvrESo
Got a PayPal verification text and thought I been hacked, but it was just codex signing up for a web service it needed.
We're introducing GLM-5.2, our latest flagship model for long-horizon tasks. It marks a substantial leap in long-horizon task capability over its predecessor GLM-5.1 and, for the first time, delivers that capability on a solid 1M-token context. GLM-5.2's new capabilities include…
More of Codex is rolling out across Europe this week. We’re bringing Computer use, the Codex Chrome extension, personalized memory, and Chronicle to Codex users in the EEA, UK, and Switzerland. https://t.co/tsriEswcyY
My heuristic is that any diff an agent generates over ~1500 lines is too big and is indicative that the problem needs to be decomposed. This is my general pattern now for feature work: 1. Try to implement the whole feature, loosely guided. I call this the "draw the owl" prompt …
Introducing LifeSciBench, a benchmark for measuring and improving how well AI supports real-world life science research. Developed with 173 scientists from biotechnology and pharmaceutical research, LifeSciBench includes 750 expert-authored tasks across seven biological researc…
I basically never write my own /goal anymore. I ask Codex to write one for itself, and one for each agent it spawns. Like this 👇 https://t.co/8ykoPJNLmC
AI is making marketers lazy. So we made the website do the work instead. Today, we're launching @ployai: the all-in-one marketing platform that turns your website into your hardest working employee. And we're coming out of stealth today with a $27M seed led by @ycombinator and…
We built an internal AI system called Builderbot. It coordinates agents across our entire codebase. Engineers tag it in Slack, and it researches, plans, and ships. The story so far: - 200,000 operations per day. - 1,500 pull requests merged per week. - 15% of all production cod…
Introducing /visual-plan - a skill to generate rich, visual plans for Claude Code and Codex. Plan mode in Claude Code is incredible. But I always find my eyes glazing over when it gives me this huge markdown essay in my terminal. I found I can make much better visual plans wi…
frontier labs are absolutely scamming you on API pricing btw GLM-5.2 is $4.4 output at 744B@40B DeepSeek-V4-Pro is $0.87 output at 1.6T@49B (and they are both making money, without any fancy Blackwell chips) Sonnet 4.6 is $15 output Opus 4.8 is $25 output GPT-5.5 is $30 output…
Just Shipped: Flue 1.0 Beta Flue is the TypeScript framework for building the next generation of agents, designed around an open agent harness with zero LLM lock-in. It’s like Astro, for agents. Flue 1.0 has been redesigned around three core primitives: 🔁 Workflows — structu…
GLM 5.1 vs GLM 5.2 6 advanced HTML canvas challenges: 💧 Ink diffusing in water ⚔️ Energy-blade duel 📱 Slide to unlock 🅿️ 360° parking assist 🔥 Burning letter to ash 🏠 Build-a-house sequence Pure canvas, zero libraries. https://t.co/1FJWuusrmI
Ranked by when the bookmark was saved during this London date range. Interactions are likes, reposts, replies, and quotes combined. Metrics are X's latest public counts; older items can be stale once they leave the recent-bookmarks sync window.