bookmarks
8 bookmarks

Why the world's best AI startups write bad prompts (& how to fix this) ## TLDR 1. Most prompts are bad because prompt evolution tends to be accretive: we only add, never remove, over time. This leads to spaghetti prompts, with contradictions and ambiguity. This has real busines…↗

New content strategy: be the Robinhood of AI Take concepts/products/best practices from Token Street (SF/X/engineering bubble) & translate for Main Street (non-technical knowledge workers/execs). WTF is loop engineering vs. graph engineering? WTF are evals & how to set up for…↗

You just hired a million bad employees. AI was supposed to replace human labor. It did the opposite. For the first time in history, humans are cheaper than software. And AI is creating more jobs than it eliminates. --- Technology has always solved one problem by creating a…↗

·@gsivulka·Jul 15, 2026·media·agents·ai-workforces·claude-code·evals

Pi's Edit Tool Thread on pi's edit tool since that came up in discussions at @aiDotEngineer, in part in relation to some people seeing edit failures even on SOTA anthropic models. pi's edit tool is intentionally boring and it is not modeled after one provider's preferred edit o…↗

Evals: the strategic IP that will define the next era of AI We've spoken to hundreds of execs in the past few months, and we're hearing a clear refrain: "AI isn't delivering ROI yet, but we're all in, so we need to figure it out." Execs know there's no going back. But their AI…↗

we built the first sane way to debug your agent locally. you can see your traces. codex/claude code can too. this lets them write evals and test your agents automatically. best part: it's completely free and open source. install with 1 line. (github below)↗