We're publishing our most detailed threat intelligence report to date. It covers how people tried to misuse Claude—for cyberattacks, influence operations, surveillance, biology, and building weapons—and how we found and stopped them. We disrupted every operation in the report…
So Jacob Coxon, who dramatically resigned from Anthropic yesterday, worked there for a grand total of six weeks. He started with them in July. All of his socials appeared yest. It has all the signs of a highly coordinated op through doomer mega donors and the corporate media.
Super Smash Bros Melee in MR. Where the players fight on your own table as a platform. Even hang off the table! On the Meta Quest. https://t.co/gCGBuDCJPP
GPT-Live-1 is now available in the API. Bring ChatGPT’s natural back-and-forth to your app, with voice agents that listen while they speak and work with the models and harness you choose. https://t.co/gIl1gwsBDV
Now everyone can put data to work. We’re introducing a new Data agent in ChatGPT Work so you can turn your company’s data into answers, interactive dashboards, and action—just by asking. Just add the Data Plugin in ChatGPT Work, connect to the data sources and context you alre…
I can’t stress enough how little an idea matters compared to the agency of the people executing the idea. I have had the privilege of knowing and sometimes even working with some of the most successful people (by various metrics). The difference between mediocre and excellent…
Introducing SWE-2, our closest model yet to the frontier. On leading evals, it scores on par with recent frontier models – at up to 70% lower cost. We scaled RL to multiple trillions of parameters, with a refined recipe that pushes the Pareto curve on both capabilities & c…
New in Claude Code: claude plugin eval See what value your plugin is adding, or if it needs more work. You can create test cases, run your plugin or skill against those test cases, score those runs, then run each case again without the plugin to see the differences. https://t.…
AI Engineering Skills Map: Shaping the build When you’re skilled at AI Engineering, your best work won’t be merely implementing a product that someone else spec’ed out. Instead, you will actively shape the build. Before modern AI tools accelerated and expanded what a single de…
Go from idea to a working agent faster with the Agents API. Build and run cloud agents with the Codex harness, fully managed by OpenAI. We handle orchestration, long-running sessions, and context management. You focus on what makes your agent unique. Available in public beta.…
This is Knap. It's a new language I created that turns data into Markdown. The syntax should feel familiar and comes with wonderfully pleasant features to modify and format plain text. Knap is open source. Over a million people already use Knap directly or indirectly because it…
Introducing Super Smash Royale, my best creation yet Play Battle Royale in the browser with 45 characters from Melee + Brawl Built in a few days with GPT-6 Astra, @ElevenLabs, @MeshyAI, @Blender and @threejs Solo, multiplayer, and controller support smashroyale dot io https:…
this was wild amounts of disinformation / fear mongering / the stupidest interview ive ever seen: 1) ai did NOT hack huggingface on its own "independent volition". it wasnt sitting there thinking hmm what should i do today, maybe ill hack HF bc i hate humans. No, 10841 *was pro…
We just rolled out CursorBench 4.0! It includes new tasks for how well models follow instructions, work on challenging projects over time, and is more difficult than before (so all models score lower). https://t.co/cYVgRBEFWE
🌊 SYSTEM PROMPT LEAK 🌊 Got the full system prompts and tools for GPT-6 Astra! 🚀 This MASSIVE dump comes in at >330k characters for the prompts and >1.1M for the tools 🤯 Hope you enjoy! 🤗 https://t.co/qRGW8pFb1x Won't all fit in a tweet, of course, but here's the first …
Here you go, link in the next post https://t.co/HIyZvKM4l8
ok everyone, i took one for the team here's the data we all wanted to see - real token value of each LLM subscription, empirically measured through usage on my real subscriptions - supergrok heavy has now become the highest value at $12k worth of tokens (40x ROI) - the $200 p…
As promised, my 𝗖𝗵𝗮𝘁𝗚𝗣𝗧 𝗪𝗼𝗿𝗸 onboarding guide. A companion to the tips so far, with more to come! Use it to build a setup around your actual job, whether you’re starting fresh or improving how you already work https://t.co/S73KMKGUCT
BREAKING: 40 British MPs have signed a letter to Prime Minister Andy Burnham calling for a ban on the development of superintelligent AI. https://t.co/zH04Xdw1qq
I co-founded Opendoor ($10B+ rev). Today, I’m launching Summation. We raised $35M from Benchmark & Kleiner Perkins to build an AI analyst you can trust. Here’s where Claude and ChatGPT fall short: 🧵 https://t.co/ZmXGGU8R8h
Ranked by when the bookmark was saved during this London date range. Interactions are likes, reposts, replies, and quotes combined. Metrics are X's latest public counts; older items can be stale once they leave the recent-bookmarks sync window.