Press Space to continue
Finding signal on Twitter is more difficult than it used to be. We curate the best tweets on topics like AI, startups, and product development every weekday so you can focus on what matters.
Press Space to continue
Press Space to continue
6.1 Sol only second in performance to Astra, but for a fraction of the cost - try it out, it’s my primary workhorse!
The AI pricing war is heating up. GPT-6.1-Sol is the 2nd most capable model after Astra (which is still the frontier model for challenging problems), and it's a ridiculous outlier on the Pareto front. @GertLabs tested GPT-6.1-Sol in 100 unsaturated coding and engineering environments and it's an unnecessarily strong answer to Opus/Sonnet 5.5. It's cheaper and pretty much smarter all around than both models with the exception of chemistry, which is a domain that recent Anthropic models perform surprisingly well in. While our bench measures hard problems, I think a lot of people are feeling that recent mid-tier models are already smart enough for most of their coding work. The speed at which frontier labs are releasing measurable upgrades at competitive prices is pretty wild.... Pricing wars == good for everyone