Introducing Sakana Fugu: A full multi-agent orchestration system accessible via a single model API. Our ‘Fugu Ultra’ model matches the performance of Fable and Mythos, delivering frontier capability without the risk of export controls. Try it: https://t.co/hhO6qTawgb 🐡
No model survives the "are you sure?" They all fold instantly.
This is becoming my favorite way to read Twitter. https://t.co/eykEElxzu7 https://t.co/m5qM5PP3rA
Starmer probably joining Anthropic
A new, more capable version of Mythos has emerged from training. I don't know whether it will be called Mythos 5.1 or Mythos 6, or if Anthropic will keep it internal to accelerate further development - but it has arrived. Stopping models like Fable 5 or Mythos 5 from being serv…
We’re expanding OpenAI Daybreak to help democratize patching vulnerable software at machine speed: - Codex Security plugin: find, validate, and fix vulnerabilities right inside Codex - The full version of GPT-5.5-Cyber model: a great model for trusted defenders - Cyber Partne…
Introducing Clips - 100% free, open source, agent-native alternative to Loom Unlike Loom, agent's can fully understand Clips just from a URL. Every Clip comes with APIs and metadata for agents to explore their contents. Agents can "see and hear" anything in a Clip - not just …
just in case you don’t know, the company behind GLM 5.2 is publicly listed, and its stock has gone up 15× in the past six months https://t.co/YI568Lq2NX
A fundamental problem with extending Codex/Cowork/Code to all knowledge work is that they remain very "software-brained" where the end result (the software) is what is important & that code serves as a source of truth. For a lot of other knowledge work, the process is at least …
Introducing React Doctor for Performance Fix slow React code and make it fast npx react-doctor@latest https://t.co/cV74c1mt3F
Introducing Ads Engine in @ElevenCreative. Connect your Google, Meta, and LinkedIn ad accounts. Localize your existing ads across 50+ languages and push the finished creatives back to your ad platform. https://t.co/rl0uYJntdG
I have a deep distrust of almost any 'self-improvement' loop in coding agents I.e. automatically created memories, CLAUDE.md suggestions applied after every session Often the suggestions themselves are shit But even if they're good, the agent often over-indexes on them in a w…
Some thoughts on ownership I shared with the team this morning https://t.co/WMXcDLU9qa
WTF Is a Loop? Part 2: The 15 Loops People Are Actually Running (and the Commands to Steal Them) Earlier this month I wrote [WTF Is a Loop? Peter Steinberger vs. Boris Cherny](https://x.com/mvanhorn/status/2063865685558903149), which did 3.6M views on what a loop even is. This …
Many people asked how the World Cup posters are generated. I wrote a breakdown of the system behind them. It covers the chrono-grid, visual encoding, and the rules that transform a match into a poster. https://t.co/lF1x0jeG2P
One interesting trend: I’m seeing *so many* VC-funded, internally built infrastructure and bootstrapped solutions around “building a context layer for engineering teams.” Aka trying to solve the problem of “if only eg Claude Code had the context from all your other systems”
this is a good example to learn from since this product is standalone, the aha moment requires a user to connect 2-3 services before they see the value the flow for doing this is often difficult particularly with their most useful data that is unique to them very few people w…
Introducing agentcn 🤖 by @shadcnlabs > Built on Eve by @vercel and @flueai > Zero config, one command setup. >@shadcn /ui compatible (simply copy-paste) > 10+ production-ready agent recipes > Fully customizable 100% free and open-source. https://t.co/RSQO98he6R
Every year around my birthday I publish an updated "reflections"-ish piece. This one is two and a half months late, but my thoughts on optionality, depreciating utility of money, tech sales, intuition, and more. https://t.co/1CxV6UN9sB
Evals: the strategic IP that will define the next era of AI We've spoken to hundreds of execs in the past few months, and we're hearing a clear refrain: "AI isn't delivering ROI yet, but we're all in, so we need to figure it out." Execs know there's no going back. But their AI…
Ranked by when the bookmark was saved during this London date range. Saves means the post's public bookmark count on X, not saves on this site. Metrics are X's latest public counts; older items can be stale once they leave the recent-bookmarks sync window.