Claude Cowork and chat are merging into one Claude. Ask a quick question or hand over a report, and Claude takes it from there, even after you close your laptop. If something's unclear, Claude asks—you keep the final say. Rolling out to Pro and Max over the next few weeks. htt…
Projects now run from one conversation, starting in Claude Code. You describe what needs doing, and Claude directs parallel threads that keep working after you close your laptop. In beta today for select Pro and Max users in cloud sessions; coming to all Claude users soon. http…
Today we're rolling out Projects in Claude Code on desktop and web. A project is one conversation with Claude. It splits the work into threads itself, runs them as parallel cloud sessions, passes context between them, and keeps going when you leave. In beta for select users. h…
found the perfect use case for @typesafeai Jev: instant compaction in 2026, why is compaction still a summarization prompt? Jev can make it instant by scoring every tool call and dropping what’s irrelevant https://t.co/h6NzKzNpCg
Nearly half a year of silence. We spent it studying one problem: how far RL can scale. MiMo-V2.6 is in the middle of its RL run right now. Three things we scaled: compute (~2B tokens per step, 1568 prompts × 16 rollouts, fully async), environments and harnesses (multi-task agen…
Has anyone yet tried to get Astra and Fable to agree on the perfect styleguide? And then host it somewhere with a web mcp.
We're sharing our new framework for tracking, investigating, and disclosing instances of model misalignment at OpenAI. The framework sets criteria and timelines for public disclosure, including when we haven’t yet fully explained or mitigated the behavior. More complex cases ma…
I made a website with 25 mini rooms, each with Claude keeping people company. Created with Claude Opus 5. https://t.co/VMRs5waWO5
🥷 New stealth model: Union Alpha (@unionalphaai) A multimodal model for research, coding, and agentic workflows. - Free to use - 256K context - Tool calling - Frontier-level general-purpose performance Try it now and share your feedback: https://t.co/SZbdoGOkdb
We open sourced BrowserSkill, a bridge between your agent and your actual browser. most tools give the agent a blank browser. We let it borrow a tab from yours, then hand it back. > login state is already there, it just works where you're signed in > captchas and confirmation …
SITUATION DETECTED: GPT-6 Astra cracked a previously unsolved 1941 German Army Enigma message in about 10 hours, searching archives, writing attack code, and testing keys autonomously.
introducing Craft a collection of design engineering concepts oss, free, and wip https://t.co/MTJJqYFvqd https://t.co/GAyyKqHZtT
We built a time machine for the web. Introducing Exa Snapshot: an index of 400 billion historical snapshots of webpages that lets you search as if it's the past. Snapshot is already being used for backtesting prediction models, RL at labs, exploring the pre-AI web, and more. h…
All the grifters are completely wrong about Jev’s architecture so I decided I’d release an open-weight version. BUT training takes time, so while we all wait I decided I’d drop the sauce. https://t.co/TK9FIiUxAJ
Here's the DeepSWE result for Union Alpha, a stealth model we just launched. Try it now! Works in every harness. $ ori [code | your-fav-harness] --model=stealth/union-alpha https://t.co/RifNfhbHvv
I reverse-engineered a jev-like architecture given its type. You can find the repo here to train your own jevlikes: https://t.co/UVPfP6OoqZ https://t.co/ebfpaPeb1U
I made a list of great startups to join. It's called the Breakout List. The list has 92 companies. These are the 20 with 25 or fewer employees: - Hone (@moritz_stephan, @CarloWillem, @oqbrady) - Normal (@ansonyuu, @hudzah) - Standard Intelligence (@G413N, @devanshpandey) - Taci…
Today, we are announcing our Series A and a new product: @raindrop_ai Simulations. We've now raised $50m from @CRV and @lightspeedvp to protect the world from agent failures, big and small. https://t.co/LMLgYcf4F2
rebuilt the X algorithm with Jev - uses real weights - simulates virality of your post - has a global feed (you see everyone) it's insanely accurate https://t.co/a7hHlj6cYK
run fast-jev-compaction: https://t.co/htDJNJnLaa
Ranked by when the bookmark was saved during this London date range. Save rate means public X bookmarks divided by views, with a 10k-view minimum. Metrics are X's latest public counts; older items can be stale once they leave the recent-bookmarks sync window.