After co-inventing ChatGPT, I kept asking myself: why have superhuman chat models not led to AGI? I’ve spent the last 2 years in stealth building a new way to train models (RLCD), and a new type of frontier AI model that we are releasing today: Jev • 20-200x faster • 40-400x c…
I have conducted an audit of Anthropic's finances. What I have found is so shocking that I am calling for a Congressional investigation. Anthropic is not just seeking regulatory capture. It has built a regulatory capture machine that cannot be turned off. Structural financia…
There are two ways AI progress could go very badly and that we must avoid. First, we could lose control of the future to AI. This is unacceptable; we are unapologetically on Team Humanity, and AI must always serve people. To ensure that, we need ways to ensure that alignment an…
An alternative hypothesis: - model performance is plateauing - compute is getting much more expensive - AI data centers are massively unpopular - slowing down AI development is a way to explain slowing progress, reduce spending, and try to regain some goodwill. - This is an eff…
Holy crap POTUS just phoned in @JensenHuang live on stage at All In Summit We will not lose the ai race! And whatever Dario said this weekend won’t stop our progress This made my morning! $NVDA https://t.co/Sr4wV9kMfz
I must be among an extremely small group of people (n=1?) that have both 1) trained a frontier LLM and 2) designed and synthesized custom viruses in a lab with my own two hands. And I think that the takes on AI killing us all by creating dangerous viruses is total bogus.
Dan Selsam is a current OpenAI capabilities researcher. (since 2022) He was my boss for a while. He doesn't have a twitter account but has made this public statement of his views on AI risk and sent it to me to share: Dan Selsam's Personal Statement on AI Risk: I have been wor…
open the frontier the frontier is the edge of what we know. no company owns what comes next. i want more people to be able to advance it. i favor open releases that people can examine, use, and improve together without waiting. i want more companies to choose openness. i'm not…
Wild 24 hours for AI and lots of different proposals have been made. TLDR; the only *tangible* new fact is that OpenAI and Anthropic are going to have embedded 3rd party evaluators from unknown organizations with Dario floating METR as a possibility. Having 3rd party evaluators…
Here's the link to it :з https://t.co/P3b6RulpvL
Morning Bathrobe Rant: Rethinking Harnesses. https://t.co/e3Lbj777mw
Introducing shadcn/lint. An agent-first linter for Tailwind design systems. You define what’s allowed. When an agent breaks a rule, the error explains what’s wrong and how to fix it using your components, variants and theme. There’s a lot you can do with this. Let me show you …
At Anthropic, Claude now writes 80% of our code. Engineers ship 8x more code per quarter. Side effect: Tests grew 10x. CI jobs up 25x in 6 months. Here's what helped us scale: https://t.co/WOL2r60vAE https://t.co/7vf60noTQt
Here's what OpenAI's agentic software factory looks like, today. Details: https://t.co/UUZufLxG2r (thanks to all the OpenAI folks who explained how it works! And Perf Factory looks especially interesting to me) https://t.co/NEyldtr1Dq
After months of writing, 'How to Unclench' is finally live! It's an interactive essay packed with stories, science & guided practices to help you unclench. → https://t.co/UTYPJkkrKH https://t.co/U9ibMPcjtU
Say hello (literally) to Gemini 3.8 Live and 3.8 Live Extended Thinking, our new SOTA live audio models, available with frontier price + performance. 3.8 Live supports 97 languages (can seamlessly switch), async tool calls, and more! https://t.co/z29J9kelr3
for the skeptics in government and elsewhere: “pacing the frontier” will compress the margins of the frontier labs. it is a heavy cost imposed asymmetrically on model developers with the strongest AIs in America. by its nature, it would be a terrible regulatory capture tactic
Introducing Bolt Forge. Free until Oct 14th: - Up to 50x more usage - The new frontier: GLM, DeepSeek, Kimi - Zero usage charges Live now in your model picker on https://t.co/UH6gFfHvbp And one more thing... 👇 https://t.co/GnrbkRdzsg
Making Startups Powerful: https://t.co/C5yIvFWXoC
On macOS and wanna see something interesting? Ask your agent: "look at ~/Library/Application Support/Knowledge/knowledgeC.db and tell me some interesting facts" I had no idea this stuff was all getting logged and it can infer a lot about your activity.
Ranked by when the bookmark was saved during this London date range. Save rate means public X bookmarks divided by views, with a 10k-view minimum. Metrics are X's latest public counts; older items can be stale once they leave the recent-bookmarks sync window.