Everyone is tired of reading AI slop. Anthropic says Opus 5.5 writes more naturally and actually follows instructions, so we ran an eval. We tested it against the new versions of OpenAI's GPT-6 Luna and Sol to see whether models can solve a problem AND write a decent explanation…↗
·@braintrust·Sep 23, 2026·benchmark·anthropic-opus-5-5·model-evaluation·openai-gpt-6-luna·openai-gpt-6-sol
