We're publishing our most detailed threat intelligence report to date. It covers how people tried to misuse Claude—for cyberattacks, influence operations, surveillance, biology, and building weapons—and how we found and stopped them. We disrupted every operation in the report…
So Jacob Coxon, who dramatically resigned from Anthropic yesterday, worked there for a grand total of six weeks. He started with them in July. All of his socials appeared yest. It has all the signs of a highly coordinated op through doomer mega donors and the corporate media.
Super Smash Bros Melee in MR. Where the players fight on your own table as a platform. Even hang off the table! On the Meta Quest. https://t.co/gCGBuDCJPP
GPT-Live-1 is now available in the API. Bring ChatGPT’s natural back-and-forth to your app, with voice agents that listen while they speak and work with the models and harness you choose. https://t.co/gIl1gwsBDV
Now everyone can put data to work. We’re introducing a new Data agent in ChatGPT Work so you can turn your company’s data into answers, interactive dashboards, and action—just by asking. Just add the Data Plugin in ChatGPT Work, connect to the data sources and context you alre…
Introducing SWE-2, our closest model yet to the frontier. On leading evals, it scores on par with recent frontier models – at up to 70% lower cost. We scaled RL to multiple trillions of parameters, with a refined recipe that pushes the Pareto curve on both capabilities & c…
Go from idea to a working agent faster with the Agents API. Build and run cloud agents with the Codex harness, fully managed by OpenAI. We handle orchestration, long-running sessions, and context management. You focus on what makes your agent unique. Available in public beta.…
We just rolled out CursorBench 4.0! It includes new tasks for how well models follow instructions, work on challenging projects over time, and is more difficult than before (so all models score lower). https://t.co/cYVgRBEFWE
🌊 SYSTEM PROMPT LEAK 🌊 Got the full system prompts and tools for GPT-6 Astra! 🚀 This MASSIVE dump comes in at >330k characters for the prompts and >1.1M for the tools 🤯 Hope you enjoy! 🤗 https://t.co/qRGW8pFb1x Won't all fit in a tweet, of course, but here's the first …
Here you go, link in the next post https://t.co/HIyZvKM4l8
As promised, my 𝗖𝗵𝗮𝘁𝗚𝗣𝗧 𝗪𝗼𝗿𝗸 onboarding guide. A companion to the tips so far, with more to come! Use it to build a setup around your actual job, whether you’re starting fresh or improving how you already work https://t.co/S73KMKGUCT
I co-founded Opendoor ($10B+ rev). Today, I’m launching Summation. We raised $35M from Benchmark & Kleiner Perkins to build an AI analyst you can trust. Here’s where Claude and ChatGPT fall short: 🧵 https://t.co/ZmXGGU8R8h
The world around you is designed in CAD. How well can AI build it? Introducing CADArena, benchmarking AI CAD generation across the tools engineers use Agents can now build accurate geometry, but struggle to create feature trees engineers can maintain and edit
We launched https://t.co/0KJOrjIfNo today! 🎉 For three years, I’ve wanted a place for Ramp designers to share their work and personality, and for y'all to get to know us. So @lalizlabeth, @baothiento, Nicholas Ano, @PaulJun_, @viktorhofte, @pontusab, and I built one! Just V1…
introducing ═══════════ 𝚌𝚏𝚘.𝚊𝚒 ═════════════ tl;dr 1/ runway is now https://t.co/Uqg6xlaJIU 2/ you can hire @arithecfo and he can start TODAY 3/ you can give @arithecfo a test drive by replying to this thread with any ridiculous thing you want him to model for you, and …
Introducing Gumball Model agnostic, proactive, self improving. All in your company’s private cloud https://t.co/1ndSTbsdBh
this css-only annotation library is so cool and elegant! built by Maxim Syabro. check out the demo site here: https://t.co/wCGBkZwDtv https://t.co/yxCbSxyoIX
Why the world's best AI startups write bad prompts (& how to fix this) ## TLDR 1. Most prompts are bad because prompt evolution tends to be accretive: we only add, never remove, over time. This leads to spaghetti prompts, with contradictions and ambiguity. This has real busine…
got my agent to make a wiki about me and told it to do its own research the result is a bit creepy (accurate) but super cool https://t.co/1RJ9zARLXZ
the average on-platform customer on sfcompute saves about 25% it's the difference between $4.5/hr and $3.3/hr we wrote about why and how you can scale faster with reduced risk https://t.co/aPfmMstxk9
Ranked by when the bookmark was saved during this London date range. Saves means the post's public bookmark count on X, not saves on this site. Metrics are X's latest public counts; older items can be stale once they leave the recent-bookmarks sync window.