Finding signal on Twitter is more difficult than it used to be. We curate the best tweets on topics like AI, startups, and product development every weekday so you can focus on what matters.
Named in these same posts. This does not imply a comparison or recommendation.
How this is put together
Public posts from the accounts Tech Twitter monitors, in the selected window. Findings need three supporting authors and a published source. Announcements can cite one known-affiliated account. This is a sample of the conversation, not a survey or a measure of adoption.
The posts behind the picture
Public source posts
@mattyp
I'm building a video editor by making Cursor edit my video
This is a Cloud Agent with the editor loaded
Cursor is using Grok 4.7 to edit the video via MCP, then testing the UI
At each failure, the agent reflects then fixes the product. The goal is a clean end-to-end edit
With the right feedback loops, you can autonomously build software
I find it fascinating how many people think this is a bad time to start a company.
Grok 4.7, GPT-6 Sol/Luna, Opus 5.5, Astra, Gemini 3.8 Live, Muse, Instinct, Jev, Agent APIs etc. That's the last 3 weeks (crazy progress on personal agents/voice AI).
Because arbitrage means the same thing trades at 2 prices in 2 markets (and it's your wedge).
And usually you have to hunt for those gaps, because they're rare and they close the moment anyone notices.
What's different now is that they're opening faster than anyone can build into them!
I think this has got to be the greatest time for arbitrage in history.
openai constantly throwing errors about their unavailable compute will drive me crazy enough that i'll only use grok 4.7 and gemini flash
i'm so tired of their bs 😤😤😤😤
Grok 4.5 was an incredible model for the price: fast, pleasant to use, reliable, solid default model.
Grok 4.6 was a (forgivable) step in the wrong direction IMO: slower and more expensive, using way more tokens per task for a slight edge in intelligence. I get it, though. They have to climb benchmarks.
Grok 4.7 is much harder to forgive. They claimed it would be more token-efficient, and it's less by 30 to 80%. It scores worse than Grok 4.6 in various benchmarks. It's slower, it's less pleasant to use, and real-world costs come out to more than 2x above Grok 4.6, putting it over Astra's costs in real-world use.
Considering how much they've been hyping this model release up, I have to say I'm disappointed. The benchmarks don't tell the whole story, and it is pleasant to use in various real-world engineering tasks, but it feels so 2025 still.
The frontend capabilities are unacceptably bad. The 3D capabilities are nonexistent. It gets stuck in random Gemini-style loops all the time.
This was a very disappointing release. I hope that the SpaceXAI team can acknowledge that and impress us with the next one.
SITUATION EXPLAINED: Grok 4.7 beats GPT-6 Astra on real-world tasks, and it's 5x cheaper.
• Same price as 4.6, $2 input and $6 output per million tokens, well under Sol, Fable, and Astra
• A new larger base model, which Elon has put at 2.1 trillion parameters, up from 1.5 trillion, with extra training on SpaceX data
• Trained to natively understand the Grok Bot harness, the same move Anthropic made with Claude Code
• Leads on electrical engineering and Harvey's legal benchmark, beats Astra Max on GDPval, trails on software engineering
• @elonmusk: "Grok 4.7 places SpaceXAI as third after Anthropic and OpenAI for agentic coding. When factoring in that Grok is significantly faster and lower cost, it's a great choice for your everyday workhorse"
@theojaffee: "SpaceX has a huge amount of data on real world hardware problems, real world engineering. So I bet Grok models are going to be better at rocketry engineering than any of the other models out there."
It happened. Grok 4.7 dropped
Better intelligence than Opus 5. Half the price
Fully baked into my favorite AI agent harness at the moment: Grok Bot
If you haven't tried using cloud cursor agents inside Grok Bot, now is by far the best time to do it
Choose a project you want to work on, connect your github, ask a grok bot to do work on it
It will spin up Cursor cloud agents and write code in the cloud. Lightning fast and incredibly smart
I recommend using a project management tools like Linear or Notion to make a bunch of tasks first, then have cloud agents just tear through them all 1 by 1.
You'll get a massive amount of work done without much oversight.
Big opportunity to lock in right now and get ahead of the curve with new tech
Take my steps up above and get to it
Grok 4.7 benchmark TLDR:
> big coding upgrade: ranks #4 on AA’s Coding Agent Index with Grok Build, just behind Fable 5.1, Astra and Opus 5
> but just +2 points on the Intelligence Index, still behind Sol and Opus 5
> with the gains in both coding and knowledge work it’s likely turned especially for Cursor and Grok Bot
> idk about Elon’s “will exceed all current models” but on real world engineering, we must see how it performs in action
i am looking forward to grok 4.7 and 4.8. grok 4.6 is pretty much the only model that’s steerable, fast, and doesn’t over engineer the world by default