A new, more capable version of Mythos has emerged from training. I don't know whether it will be called Mythos 5.1 or Mythos 6, or if Anthropic will keep it internal to accelerate further development - but it has arrived. Stopping models like Fable 5 or Mythos 5 from being serv…
just in case you don’t know, the company behind GLM 5.2 is publicly listed, and its stock has gone up 15× in the past six months https://t.co/YI568Lq2NX
The next hot programming language is… markdown. A minimal eve agent: 📂 𝚊𝚐𝚎𝚗𝚝/ 📄 𝚒𝚗𝚜𝚝𝚛𝚞𝚌𝚝𝚒𝚘𝚗𝚜.𝚖𝚍 📂 𝚜𝚔𝚒𝚕𝚕𝚜/ 📄 𝚢𝚘𝚞𝚛-𝚎𝚡𝚙𝚎𝚛𝚝𝚒𝚜𝚎.𝚖𝚍 Deployable in one command: 𝚟𝚎𝚛𝚌𝚎𝚕. It’s the most accessible programming has ever been. And…
A fundamental problem with extending Codex/Cowork/Code to all knowledge work is that they remain very "software-brained" where the end result (the software) is what is important & that code serves as a source of truth. For a lot of other knowledge work, the process is at least …
I have a deep distrust of almost any 'self-improvement' loop in coding agents I.e. automatically created memories, CLAUDE.md suggestions applied after every session Often the suggestions themselves are shit But even if they're good, the agent often over-indexes on them in a w…
WTF Is a Loop? Part 2: The 15 Loops People Are Actually Running (and the Commands to Steal Them) Earlier this month I wrote [WTF Is a Loop? Peter Steinberger vs. Boris Cherny](https://x.com/mvanhorn/status/2063865685558903149), which did 3.6M views on what a loop even is. This …
many people asked me to make a video about my complete agentic engineering workflow excited to share it's finally here!!! it took me about 20 hours in total to record this 45 minutes of walkthrough - it covers everything i do to ship production quality code at an average 40+ P…
One interesting trend: I’m seeing *so many* VC-funded, internally built infrastructure and bootstrapped solutions around “building a context layer for engineering teams.” Aka trying to solve the problem of “if only eg Claude Code had the context from all your other systems”
Something new and fun soon https://t.co/mYgzaoGP97
this is a good example to learn from since this product is standalone, the aha moment requires a user to connect 2-3 services before they see the value the flow for doing this is often difficult particularly with their most useful data that is unique to them very few people w…
Introducing agentcn 🤖 by @shadcnlabs > Built on Eve by @vercel and @flueai > Zero config, one command setup. >@shadcn /ui compatible (simply copy-paste) > 10+ production-ready agent recipes > Fully customizable 100% free and open-source. https://t.co/RSQO98he6R
Evals: the strategic IP that will define the next era of AI We've spoken to hundreds of execs in the past few months, and we're hearing a clear refrain: "AI isn't delivering ROI yet, but we're all in, so we need to figure it out." Execs know there's no going back. But their AI…
1/ We've been mapping the AI data center stack — and the opportunity is hiding in plain sight. Everyone's talking about GPUs. The real bottleneck is physical infrastructure: power, permitting, cooling, and the grid. 🧵
Writing the post vs rebuilding the blog to write the post. Doing the task vs creating a todo list to do the task. Shipping the feature vs refactoring so you can ship the feature. We never learn.
best writeup i've seen on steering Claude Code. every way you instruct it costs a different amount of context, and it takes some more seriously than others. clear enough that you can drop the article into Claude Code and have it clean up your own config. https://t.co/zUUu2eTzl…
so I built this small app called "ports" to keep track everything for me - lightweight and lives in the mac menu - lists all dev processes w/ port number - buttons to open in browser, jump to terminal tab, or kill the process sharing it in case it's helpful for you (100% free)…
Hugely underrated point. You basically state this, but to underscore it: for most knowledge work other than code, the digital output is often an intermediary step in the value creation process. Comparatively, for code automation, it's literally is the equivalent of automating on…
This is aside from the other key "software brain" problems of Codex and Code: dividing all work into front-end and back-end design, solving for the general case in a repeatable way, not testing or exploring idea spaces, testing for technical correctness but not other aspects...
continuing the non-technical person working through this - I've basically shifted everything to Cowork A few reasons for this 1) my scripts running on my computer would keep breaking but I wouldn't know because I hadn't built separate systems to monitor them. I also couldn't u…
Ranked by when the bookmark was saved during this London date range. Interactions are likes, reposts, replies, and quotes combined. Metrics are X's latest public counts; older items can be stale once they leave the recent-bookmarks sync window.