Finding signal on Twitter is more difficult than it used to be. We curate the best tweets on topics like AI, startups, and product development every weekday so you can focus on what matters.
Named in these same posts. This does not imply a comparison or recommendation.
How this is put together
Public posts from the accounts Tech Twitter monitors, in the selected window. Findings need three supporting authors and a published source. Announcements can cite one known-affiliated account. This is a sample of the conversation, not a survey or a measure of adoption.
The posts behind the picture
Public source posts
@omarsar0
Interesting results here. This is why I expect more agent workloads to run on blended models.
Pareto 26.9 from @TheUnbiasedCo sends requests to several frontier and open models and keeps the best answer.
In the new eval of 30 agent tasks, Pareto tied GPT-6 Astra for first place at about 1/3 the cost per successful task.
It also finished tasks faster than DeepSeek V4 Pro and GLM 5.3 Flash.
SITUATION EXPLAINED: GPT-6 Sol Max is two spots behind Astra at a fifth of the cost.
• Fourth in Code Arena at $8 per million tokens, blended input and output
• On par with Opus 5 Max while costing less than half as much
• On web dev it sits behind Opus 5 Max, Fable 5.1 Max, and Astra
• Third on games
@schisofrenia: "The Pareto is just getting constantly transformed every week."
I find it fascinating how many people think this is a bad time to start a company.
Grok 4.7, GPT-6 Sol/Luna, Opus 5.5, Astra, Gemini 3.8 Live, Muse, Instinct, Jev, Agent APIs etc. That's the last 3 weeks (crazy progress on personal agents/voice AI).
Because arbitrage means the same thing trades at 2 prices in 2 markets (and it's your wedge).
And usually you have to hunt for those gaps, because they're rare and they close the moment anyone notices.
What's different now is that they're opening faster than anyone can build into them!
I think this has got to be the greatest time for arbitrage in history.
GPT 6 Sol definitely feels better than GPT 5.6 and I still prefer ChatGPT as my harness but in my humble opinion Opus 5.5 is the best model available right now.
I prefer it over both Fable and Astra.
What an insane day in AI. The frontier models just became substantially cheaper, with the Opus 5.5 price cuts, and now with GPT-6 Sol and Luna dropping token prices by 50%.
The rate at which the cost per task (on a like-for-like basis) drops in AI is unlike any other type of technology in history. And every time the cost of AI drops, the use-cases you can deploy agents against dramatically increase. This is Jevons paradox applied to agents.
These improvements will directly lead to broader diffusion of AI in the economy as we can use agents to process all of our data, scan our code for security issues, read through all log data to make decisions, have agent swarms in workflows, and much more. The cost of tokens is directly correlated to these use-cases being opened up at scale.
Hot take: this is what pacing looks like.
None of today's releases were Astra or Fable tier. This is intentional.
The point of "pacing" isn't to stop iteration and improvement. The goal is to prevent the development bigger models from spiraling out of control.
Opus and Sol class models are a great place for our focus to go right now. Lots of opportunity for real wins without as much risk :)
GPT-6 Sol is now available to all Perplexity users. On our Wide-And-Deep-Research (WANDR) evals, it outperforms Opus 5 at one-fifth the price. Sol will become the “Light” Effort orchestrator for Computer users, while Astra remains the orchestrator for "High" effort. Congrats to @OpenAI for consistently launching pareto-optimal models!
Cognition is giving away 50 $200 Devin Max plans to celebrate the new model launches! ⚡
Now available in Devin:
• GPT-6 Astra, Sol, Luna + others
• Claude Opus 5.5, Fable 5.1 + others
• SWE-2 (Free until October 15)
• Fusion Frontier harness (Fable, Astra, Sol, Opus)
• Gemini 3.8 Flash + others
• Grok 4.7 + others
• Kimi K3 + others
• Inkling
• DeepSeek V4.1 Flash + others
• GLM-5.3 Flash + others
• Cloud agents on Linux, macOS, and Windows
To be eligible, reply below with what you're building (or something you'd like to build with Devin)! We will choose winners in 24 hours.
GPT-6 Sol and Luna are big improvements on intelligence, alignment, work output, coding, computer use, and more over their 5.6-family predecessors.
They are also half the price per token, and even less per task!
GPT-6 Sol and GPT-6 Luna from @OpenAI are live on OpenRouter!
Half the price of their GPT-5.6 predecessors, Sol at $2/M input and $10/M output, Luna at $0.10/M input and $0.50/M output. On AutomationBench each one tops its predecessor's best score at a fraction of the cost per task.
What each one is for 🧵
Please welcome GPT-6 Sol and GPT-6 Luna to the GPT-6 universe.
GPT-6 Sol and Luna build on the advances behind GPT-6 Astra, bringing much of its strengths into faster and more affordable models to support work at scale.
We’ve also made caching and inference more efficient, and we’re passing the savings directly to you: 50% lower API prices for Sol and Luna compared with GPT‑5.6 promotional pricing.
It's a double release day just like old times!
GPT-6 Sol and GPT-6 Luna are rolling out, for some of my mutuals they are already live. OpenAI announcement incoming.
If these prices rumoured for GPT-6 Sol are real and it's token efficient, inject it into my veins!
(and let me buy another $200 sub!)
$2.50 / 1M input
$15 / 1M output