This is GPT-6 Astra. Anything you can do on a computer, Astra can do for you. Fast. https://t.co/gDd0IsewJw
https://t.co/U1TxCPDwwY
Claude can now use your computer in the background in Claude Cowork and Claude Code. Give it something to do on your desktop and Claude clicks, types, and opens apps just like you would, while you work on something else. https://t.co/AOiup03pQK
I had early access to GPT-6. This model will completely shatter your understanding of what's possible, and it will start a new era of creativity. It's hard to explain how big of a jump this is, so I'll share my tests. 1/6 Made this 3D model and animation in code from an image.…
GPT-6 Astra represents a step-function change in model capability for interactive reasoning problems. It scores 66% on ARC-AGI-3 using our standard harness, and nearly 100% with a continuous conversation harness and custom compaction, at a cost of roughly $360 per game. In fact…
gpt-6-astra gpt-6-astra-aeon ^ new slugs added to a statsig feature flag that appeared ~30 mins ago in Codex Desktop "aeon" means "an immeasurably or indefinitely long period of time, an age, or eternity"
Playgrnd® beta is live. No account. No paywall. 100% free. 32 tiny tools for making weird, beautiful things, with new ones dropping every day. Animate in one click. Dither everything. Export SVGs. Go make something weird. 🛝 https://t.co/Vs4BkFweHm https://t.co/QSPhHPtnL9
New OpenAI repo with a Lean formalization by GPT-6-Astra proves that there are infinitely many pairs of consecutive primes whose distance is at most 186 https://t.co/kegFNJKQrq https://t.co/bc8k6H1OdX
GPT-6-Astra Benchmarks ARC-AGI-3 - 98.6% FrontierMath Tier 4 v2 - 97.6% DeepSWE v1.1 - 74.1% ExploitBench - 100%
Your input needed: would you use this? This is an early look at how we're thinking about making Claude Code way more extensible. It's a little crazy, and very exciting. More details here: https://t.co/X2SbdZYxvi
Today we're announcing Base Labs, a dedicated research organization focused on advancing open-source AI. We believe in a healthy, open frontier model ecosystem. To enable this, we are working on: - Blue-sky research on continual learning, the science of RL, and how models lear…
BREAKING: OpenAI just dropped GPT-6 ASTRA!!! 🚀✨ We’ve been testing it extensively at @every across coding, writing, and knowledge work. My take: it’s a big upgrade from 5.6-Sol, with some frustrating habits that keep it from matching Fable at the top end. Here’s your vibe ch…
I have tested GPT-6 Astra this week. It is a show horse, not a workhorse. Super fun to use and showy, does crazy good stuff, and design-wise it's the most interesting model I've run. But I don't know if I'd use it for a real day to day engineering/product work. It's not as trust…
Today we are releasing our speculative decoding implementation in our inference engine uzu. Initially for Qwen3.6 27B, with support for Qwen3.8 27B and Muse Glimmer coming soon. On Apple M5-series chips, we outperform MTPLX (MLX + speculative decoding) by almost 2x, and llama.…
[Channels Mariah]: It's tiiiiiiiiiime! Today @every is releasing Compound Writing, so now you - yes, you - can install my tendency to overthink a paragraph. I started this past winter by pulling down the Compound Engineering plugin and telling Claude (4.5 at the time, what a …
GPT 6 Astra is here. We ran the numbers on AutomationBench: It's the highest score we've ever recorded. Clean sweep across every domain. Scores 41.4% at Max effort. For context, no model had cleared 40% before today (GPT-5.6-Sol scored 28.8%) 𝗕𝗲𝘀𝘁 𝗳𝗶𝘁 𝗳𝗼𝗿: reconcili…
New plugin alert! The Kernel Browser plugin lets your agents create and control remote browsers and while showing you an in-app preview of them at work! https://t.co/dgwZCrnPpX
New episode with Matan Grinberg (@matanSF), CEO & co-founder of @FactoryAI !! We get into so many good topics: Studying physics for a decade, how to measure ROI of AI spend, who to hire in a world where we don't code, policy, the risk of Chinese open models vs. closed-model mon…
PSA if you're using Claude Code within @get_bb_app: BB uses a local proxy, so Claude Desktop's “load tools when needed” setting doesn’t carry over to BB-launched sessions and Claude loads every MCP tool schema upfront. For me, that meant 167k tokens already filling the contex…
Ranked by when the bookmark was saved during this London date range. Saves means the post's public bookmark count on X, not saves on this site. Metrics are X's latest public counts; older items can be stale once they leave the recent-bookmarks sync window.