📢Meet Qwen3.8-Max — our most capable model to date. Next week, the open weights of Qwen3.8-Max will be released, and Qwen3.8-27B is also going open-weights to meet you all!🎉 Qwen3.8-Max, a new bar for coding and cowork at 2.4T parameters: - Autonomous coding: 10+ days of s…
An internal version of our next major model produced 10 new results on long-standing open problems in mathematics and theoretical computer science, using roughly $2,000 worth of tokens at GPT-5.6 Sol API rates. https://t.co/4cgowmPOpY
After doing ~60B tokens, this is my full AGENTS.md https://t.co/VVXIwtSAcO
Here is my AI investing guide. Sitting here August 2026, my current best thoughts are as follows: 1. LPS (Land Power Shell) is still the most obvious and fastest path to cash on cash returns. Lots of value can be assembled and traded quickly at this layer. And as data centers …
finally a planner that thinks in time, not lists.. your day curls around the clock face https://t.co/6XOpYCt2Ir
We've decided to open-source the CRM we built for ourselves at Comp AI. It's agentic-first, which we mean literally: durable research agents. MIT license. Built using next, eve, and context. https://t.co/JLwQCXinW1
Codex tip: tell Sol to configure 𝗟𝘂𝗻𝗮 𝗠𝗮𝘅 𝗮𝘀 𝗮 𝘀𝘂𝗯𝗮𝗴𝗲𝗻𝘁. 𝗽𝗿𝗼𝗺𝗽𝘁: create a custom agent named luna_worker at ~/.codex/agents/luna-worker.toml. use these settings: model = "gpt-5.6-luna" model_reasoning_effort = "max" give it a description and instruct…
I made this Three.js bookshelf inspired by the Stripe Press site. It scrolls horizontally, and you can open each book and flip through the pages. It took quite a few prompts to get the textures and small details right. I’m open-sourcing the code and prompt. Demo: https://t.co…
Today, we're introducing @Intelligence_ai. In 6 months, as a team of 10, we scaled from $5M to $60M ARR and 5.5M users across 190+ countries. We raised a $7.9M seed, led by @IndexVentures with participation from @conviction, @A_StarVC, and @combinator to build DesignArena, a u…
We built an agent that powers our company's internal operations called @𝚟. Every day-to-day job at Vercel now involves @𝚟. It's growing exponentially both in daily interactions and token use. One can extrapolate to the day AI agents run entire companies from here. It's an ex…
I just released 70K+ high-quality hand-drawn icons, completely free to use. Use them however you want. No attribution required. Check them out 👇 https://t.co/KQ3VBtkKj8
Made a quick survey about the economics of AI: https://t.co/G9Xl7RW9Si.
One of our big findings in our study at Procter and Gamble was that AI blurred the lines between jobs. Now OpenAI has a similar finding. Organizational boundaries are becoming porous, the walls thinning. Companies are going to need to think about division of labor in a new way.…
Rick Rubin's House on the Mountain test: Create according to your own taste, not for applause, critics, algorithms, or market demand. "Imagine going to live on a mountaintop by yourself, forever. You build a home that no one will ever visit. Still, you invest the time and effo…
AI forecasting is now approximately superhuman. Today, FutureSearch is exiting our public beta and launching to everyone. FutureSearch is the original AI forecasting company, started in August 2023. We’re currently #1 of 194 in the most competitive AI forecasting tournament, an…
After doing ~600B tokens, this is my full AGENTS.md https://t.co/5vuObTXmN6
Introducing Supabase Evals. Our benchmark for how well AI coding agents build with Supabase. We run agents like Claude Code, Codex, and Open Code against real tasks and score what they do. https://t.co/N5oKIcQwh9
Introducing Ori Eval: the easiest way to write your first eval. There's no definitive best model, only the best model for each task. Ori Eval leverages OpenRouter's APIs for each task in your codebase, and then evaluates the results. curl -fsSL https://t.co/ABRt1wxtZ4 https://…
Introducing Modern Claudefare Built fully with Opus 5 on High Mode over a few days using principles from @mattshumer_’s Gauntlet Loop - 84,100 lines of code Includes remakes of 4 beloved maps: 0:00 - Rust 0:40 - Highrise 0:57 - Nuketown 1:26 - Terminal Solo & multiplayer (w/ …
An internal version of Astra, our next major model, found new results across 10 long-standing open problems in math and theoretical computer science. The total token cost to find all 10 solutions? Roughly $2,000 at Sol API rates. Astra then formalized each argument in Lean. h…
Ranked by when the bookmark was saved during this London date range. Interactions are likes, reposts, replies, and quotes combined. Metrics are X's latest public counts; older items can be stale once they leave the recent-bookmarks sync window.