Introducing Sakana Fugu: A full multi-agent orchestration system accessible via a single model API. Our ‘Fugu Ultra’ model matches the performance of Fable and Mythos, delivering frontier capability without the risk of export controls. Try it: https://t.co/hhO6qTawgb 🐡
No model survives the "are you sure?" They all fold instantly.
This is becoming my favorite way to read Twitter. https://t.co/eykEElxzu7 https://t.co/m5qM5PP3rA
Starmer probably joining Anthropic
A new, more capable version of Mythos has emerged from training. I don't know whether it will be called Mythos 5.1 or Mythos 6, or if Anthropic will keep it internal to accelerate further development - but it has arrived. Stopping models like Fable 5 or Mythos 5 from being serv…
We’re expanding OpenAI Daybreak to help democratize patching vulnerable software at machine speed: - Codex Security plugin: find, validate, and fix vulnerabilities right inside Codex - The full version of GPT-5.5-Cyber model: a great model for trusted defenders - Cyber Partne…
Introducing Clips - 100% free, open source, agent-native alternative to Loom Unlike Loom, agent's can fully understand Clips just from a URL. Every Clip comes with APIs and metadata for agents to explore their contents. Agents can "see and hear" anything in a Clip - not just …
just in case you don’t know, the company behind GLM 5.2 is publicly listed, and its stock has gone up 15× in the past six months https://t.co/YI568Lq2NX
The next hot programming language is… markdown. A minimal eve agent: 📂 𝚊𝚐𝚎𝚗𝚝/ 📄 𝚒𝚗𝚜𝚝𝚛𝚞𝚌𝚝𝚒𝚘𝚗𝚜.𝚖𝚍 📂 𝚜𝚔𝚒𝚕𝚕𝚜/ 📄 𝚢𝚘𝚞𝚛-𝚎𝚡𝚙𝚎𝚛𝚝𝚒𝚜𝚎.𝚖𝚍 Deployable in one command: 𝚟𝚎𝚛𝚌𝚎𝚕. It’s the most accessible programming has ever been. And…
A fundamental problem with extending Codex/Cowork/Code to all knowledge work is that they remain very "software-brained" where the end result (the software) is what is important & that code serves as a source of truth. For a lot of other knowledge work, the process is at least …
Introducing React Doctor for Performance Fix slow React code and make it fast npx react-doctor@latest https://t.co/cV74c1mt3F
Introducing Ads Engine in @ElevenCreative. Connect your Google, Meta, and LinkedIn ad accounts. Localize your existing ads across 50+ languages and push the finished creatives back to your ad platform. https://t.co/rl0uYJntdG
I have a deep distrust of almost any 'self-improvement' loop in coding agents I.e. automatically created memories, CLAUDE.md suggestions applied after every session Often the suggestions themselves are shit But even if they're good, the agent often over-indexes on them in a w…
Some thoughts on ownership I shared with the team this morning https://t.co/WMXcDLU9qa
WTF Is a Loop? Part 2: The 15 Loops People Are Actually Running (and the Commands to Steal Them) Earlier this month I wrote [WTF Is a Loop? Peter Steinberger vs. Boris Cherny](https://x.com/mvanhorn/status/2063865685558903149), which did 3.6M views on what a loop even is. This …
Many people asked how the World Cup posters are generated. I wrote a breakdown of the system behind them. It covers the chrono-grid, visual encoding, and the rules that transform a match into a poster. https://t.co/lF1x0jeG2P
many people asked me to make a video about my complete agentic engineering workflow excited to share it's finally here!!! it took me about 20 hours in total to record this 45 minutes of walkthrough - it covers everything i do to ship production quality code at an average 40+ P…
One interesting trend: I’m seeing *so many* VC-funded, internally built infrastructure and bootstrapped solutions around “building a context layer for engineering teams.” Aka trying to solve the problem of “if only eg Claude Code had the context from all your other systems”
Something new and fun soon https://t.co/mYgzaoGP97
this is a good example to learn from since this product is standalone, the aha moment requires a user to connect 2-3 services before they see the value the flow for doing this is often difficult particularly with their most useful data that is unique to them very few people w…
Ranked by when the bookmark was saved during this London date range. Interactions are likes, reposts, replies, and quotes combined. Metrics are X's latest public counts; older items can be stale once they leave the recent-bookmarks sync window.