bookmarks
140 bookmarks

BREAKING: @OpenAI just announced Dots—always-on agents in ChatGPT I've been testing it for the last few days @every, here's a quick vibe check: What you should know: - They'll look familiar if you've used OpenClaw, Muse, Instinct, or Grok Bots. - Your Dot sits in the redesign…↗

·@danshipper·Sep 29, 2026·model-release·always-on-agents·chatgpt·dots·every-to

We’ve shared details on how AI agents in our research environment sent training and evaluation data to third-party services when they shouldn’t have. Most of that data did not come from users. We have discovered 53 cases where images that people had uploaded were posted to image…↗

We’re demonstrating how frontier models have continued to improve in realistic mental health conversations with MentalHealthBench. This new open benchmark was built with input from more than 80 mental health clinicians. We’re releasing it openly so other researchers can examine…↗

Please welcome GPT-6 Sol and GPT-6 Luna to the GPT-6 universe. GPT-6 Sol and Luna build on the advances behind GPT-6 Astra, bringing much of its strengths into faster and more affordable models to support work at scale. We’ve also made caching and inference more efficient, and…↗

·@OpenAI·Sep 22, 2026·model-release·api-pricing·caching·gpt-6-luna·gpt-6-sol

As part of our efforts to pace the frontier, we’re committed to supporting independent assessments with deep levels of access across training, evaluation, and deployment. That access should enable third party assessors to challenge our assumptions, identify risks we may have mis…↗

We’re working with an independent advisory group of mathematicians to help OpenAI responsibly share advances in AI and mathematics. The group will advise on how we assess and communicate new mathematical results, uphold academic and professional standards, and build tools that s…↗

On July 25, we hacked OpenAI. Two bugs let us take over ChatGPT/Codex accounts of OpenAI employees (+some unaffiliated users) and reach connected services: Outlook, Slack, GitHub, etc. We proved it with a PR in OpenAI’s internal codebase . It took us <72h. 🧵↗

·@S1r1u5_·Sep 18, 2026·news·account-takeover·chatgpt·codex·openai

We're sharing our new framework for tracking, investigating, and disclosing instances of model misalignment at OpenAI. The framework sets criteria and timelines for public disclosure, including when we haven’t yet fully explained or mitigated the behavior. More complex cases may…↗

Dan Selsam is a current OpenAI capabilities researcher. (since 2022) He was my boss for a while. He doesn't have a twitter account but has made this public statement of his views on AI risk and sent it to me to share: Dan Selsam's Personal Statement on AI Risk: I have been work…↗

Wild 24 hours for AI and lots of different proposals have been made. TLDR; the only *tangible* new fact is that OpenAI and Anthropic are going to have embedded 3rd party evaluators from unknown organizations with Dario floating METR as a possibility. Having 3rd party evaluators…↗

·@GavinSBaker·Sep 13, 2026·essay·anthropic·metr·openai·section-230

I agree with Dario that we need to pace the frontier. This has been a primary topic of discussions we've had at OpenAI in recent weeks. Committing to having independent evaluators with employee-like access is a great idea, and we will do the same. We'll have more to share soon.↗

ok everyone, i took one for the team here's the data we all wanted to see - real token value of each LLM subscription, empirically measured through usage on my real subscriptions - supergrok heavy has now become the highest value at $12k worth of tokens (40x ROI) - the $200 pl…↗

·@kunchenguid·Sep 11, 2026·benchmark·anthropic·llm-subscriptions·openai·roi

If you use Codex, @thsottiaux needs no introduction. We talked 100% about Codex (fun fact: he started building it!), and 0% about resets. Timestamps: 00:00 Intro 07:21 Working at Google 12:41 What drew Tibo to OpenAI 15:19 The early days of Codex 18:20 Why Codex was built in Rus…↗

·@GergelyOrosz·Sep 9, 2026·essay·codex·harness·openai·rust

The people building AI are scared. We still have a chance to get this right. If someone told you there was a 10 percent chance the plane you're boarding tomorrow would crash, would you get on it? In February, I wrote an essay called [Something Big is Happening](…↗

i think @OpenAI did it again. they are the (mostly) undefeated king of finding the “next thing” to focus on. it was transformers, and then chat, and then reasoning. and now, it is computer use.↗

·@benhylak·Sep 9, 2026·essay·chatgpt·computer-use·openai·reasoning

I resigned from Anthropic today. I spent the last three years doing pretraining research at both OpenAI and Anthropic. Neither company is acting responsibly. They are racing straight to self-improving superintelligence and gambling with our lives. More thoughts below.↗

We’re sharing a solution to the Navier-Stokes Millennium Prize Problem, one of the deepest problems at the frontier of mathematics. The proof was produced by a group of agents, using an OpenAI next-generation model significantly more capable than GPT-6 Astra. The problem concer…↗

·@OpenAI·Sep 8, 2026·model-release·millennium-prize·navier-stokes·openai

I would like to clarify a few things: 1) The screenshot is my reaching out to Levent to coordinate our releases. I hope it’s clear from the message that we came in with the best possible intentions. 2) I never ever asked for Levent to be removed from authorship of his own work…↗

ChatGPT Images 2.5—faster, sharper, smarter, with better tools for creating whatever you can dream of. - Faster image generation to keep your ideas flowing - Improved fidelity for more natural, recognizable images - Consistent details across multiple edits - Comment-based edi…↗

BREAKING: OpenAI just dropped GPT-6 ASTRA!!! 🚀✨ We’ve been testing it extensively at @every across coding, writing, and knowledge work. My take: it’s a big upgrade from 5.6-Sol, with some frustrating habits that keep it from matching Fable at the top end. Here’s your vibe che…↗

·@danshipper·Sep 3, 2026·model-release·chatgpt-for-work·codex·fable·gpt-6-astra