I can’t stress enough how little an idea matters compared to the agency of the people executing the idea. I have had the privilege of knowing and sometimes even working with some of the most successful people (by various metrics). The difference between mediocre and excellent…
New in Claude Code: claude plugin eval See what value your plugin is adding, or if it needs more work. You can create test cases, run your plugin or skill against those test cases, score those runs, then run each case again without the plugin to see the differences. https://t.…
AI Engineering Skills Map: Shaping the build When you’re skilled at AI Engineering, your best work won’t be merely implementing a product that someone else spec’ed out. Instead, you will actively shape the build. Before modern AI tools accelerated and expanded what a single de…
This is Knap. It's a new language I created that turns data into Markdown. The syntax should feel familiar and comes with wonderfully pleasant features to modify and format plain text. Knap is open source. Over a million people already use Knap directly or indirectly because it…
Introducing Super Smash Royale, my best creation yet Play Battle Royale in the browser with 45 characters from Melee + Brawl Built in a few days with GPT-6 Astra, @ElevenLabs, @MeshyAI, @Blender and @threejs Solo, multiplayer, and controller support smashroyale dot io https:…
this was wild amounts of disinformation / fear mongering / the stupidest interview ive ever seen: 1) ai did NOT hack huggingface on its own "independent volition". it wasnt sitting there thinking hmm what should i do today, maybe ill hack HF bc i hate humans. No, 10841 *was pro…
ok everyone, i took one for the team here's the data we all wanted to see - real token value of each LLM subscription, empirically measured through usage on my real subscriptions - supergrok heavy has now become the highest value at $12k worth of tokens (40x ROI) - the $200 p…
BREAKING: 40 British MPs have signed a letter to Prime Minister Andy Burnham calling for a ban on the development of superintelligent AI. https://t.co/zH04Xdw1qq
My blog post about how I manage agents is finally ready. It first explains my bespoke `plans` setup, and then explains how @cursor_ai new Projects feature enabled me |o do so much more. It was the missing piece for me. Check it out: https://t.co/TbIdJqxVGo
Introducing Townies – a relaxing place where up to 50 players can keep a quaint, yet blossoming, town thriving. Built on Workers & Durable Objects. Watch others in the town mow grass, deliver papers, clean up trash, picnic at the park, or buy a house. https://t.co/OwuyYSMm…
If you want to play around with an early thing: you can now give us money and we give you tokens in return. https://t.co/9u11mtoT6S
San Francisco Compute has closed a few more deals. https://t.co/UWOUMxMh3M
happy friday! i did some interaction work https://t.co/IHPtIRmndm https://t.co/q1ZJSnVhtl
My YC batch wrapped up yesterday & I wanted to share a few of the companies I was close with and think you should check out! @hoplite_sh - cloud agents with a beautiful design - their onboarding flow is incredible and gets you up and running super quickly @try_glen - organizat…
check out how AI helps scientists with quantum computing research @bea_yankelevich at MIT's EQuS group used GPT-5.6 Sol with Codex to run routine measurements on quantum chips, so she could accelerate her experiment design https://t.co/3HWu21htug
I've decided to backup GitHub https://t.co/cu6tNhnX6O AMA
I sat down with AI leaders from @figma, @box, @cognition, @tryramp, BAM, and the one and only @clairevo to hear how they’re using Astra. Optimizing GPU kernels, revamping workflows with computer use, designing many consistent screens… Amazed by what they’d done in just days! 🌌…
Agents can now automate most repetitive tasks. But running them 24/7 burns countless tokens. Introducing Routines: automated recurring work with AI. Choose an hourly, daily, or weekly schedule. Each run starts with deterministic code and invokes Agent only when reasoning is ne…
better benchmark scores don't tell you much about how well a model does on your real work that's why for the last 3 years @every we've done vibe checks on new models: long-form reviews based on hands-on testing on each model for our real work now we're doubling down, and gett…
It's here! Control any terminal session from your team from any device. https://t.co/jfsO5Kjzk4 V2 gives you one unified view of every agent your org is running, and a live keyboard into each of them. So excited for this 1/👇👇👇 https://t.co/dHGKnpWrWK
Ranked by when the bookmark was saved during this London date range. Save rate means public X bookmarks divided by views, with a 10k-view minimum. Metrics are X's latest public counts; older items can be stale once they leave the recent-bookmarks sync window.