bookmarks
5 bookmarks

Introducing SWE-2, our closest model yet to the frontier. On leading evals, it scores on par with recent frontier models – at up to 70% lower cost. We scaled RL to multiple trillions of parameters, with a refined recipe that pushes the Pareto curve on both capabilities & cost.↗

Today we're announcing Base Labs, a dedicated research organization focused on advancing open-source AI. We believe in a healthy, open frontier model ecosystem. To enable this, we are working on: - Blue-sky research on continual learning, the science of RL, and how models learn…↗

🎙️ 🚨 Latest State of Agentic Coding w/ special guest @mariozechner (Flask) is out: - Vibe-checking Fable & GLM 5.2 - @mitsuhiko teaches us RL - How harnesses smooth out model jank - Loops, surviving AI fomo, and more Also on YT & Spotify ⬇️↗

·@bentlegen·Jul 13, 2026·media·agentic-coding·fable·flask·glm-5-2

Introducing SWE-1.7, the most capable model we’ve trained yet. It scores within a few points of the strongest frontier models at a fraction of the cost, and is now available at 1000 tok/s. RL is not hitting its limit: after refining our recipe, we keep seeing gains as we scale↗

·@cognition·Jul 8, 2026·model-release·inference-speed·model·rl·swe-1-7

.@polymath_labs is training world generation models to automate the creation of RL environments. Traditionally, RL environment generation has been bottlenecked by human data. Superintelligence will never be achieved by human data alone. Polymath is building the core technology…↗