Finding signal on Twitter is more difficult than it used to be. We curate the best tweets on topics like AI, startups, and product development every weekday so you can focus on what matters.
Public posts from the accounts Tech Twitter monitors, in the selected window. Findings need three supporting authors and a published source. Announcements can cite one known-affiliated account. This is a sample of the conversation, not a survey or a measure of adoption.
The posts behind the picture
Public source posts
@googleai
Known affiliation
Check out this week's updates and releases:
— Gemini 3.8 Flash TTS and Gemini 3.8 Flash-Lite TTS, two of our most expressive audio generation models yet
— Gemini 3.8 Live with Live Avatar, bringing near real-time visual presence to Gemini’s conversational AI
— @Gemini_Notebook Interactive Learning Overviews, giving all users an interactive hub to combine source summaries and artifacts
— Live Chat on the @Gemini_Notebook mobile app, bringing real-time, hands-free voice conversations across ~100 languages
— Project Suncatcher, our moonshot announced last year, will launch a prototype satellite to test @Google TPUs in orbit and explore solar-powered AI compute in space
Gemini Audio Live last night was SO much fun. Thanks @GoogleDeepMind
for an amazing event!
🍣 Used live transcription to order sushi across English/Japanese
⛳ Gemini coached my golf swing — and my second swing was actually better
🎶 Lyria made a custom song about my night and pressed it onto vinyl
Also loved talking with founders building on these models + celebrating with the Gemini Audio teams, who have been on an incredible launch streak lately.
Seeing transcription, real-time audio, music, and reasoning all come together in actual experiences was very cool. 💫
At the Gemini Audio event tonight from @GoogleDeepMind and WOW what a turn out! What an amazing set of launches from that team this week, can't wait to see what folks make with these new models (I've been having a lot of fun myself)!
Highlight of the night: Ordering sushi in English to someone who only speaks Japanese and having Gemini Live translate our conversation in real time 🤯 ...then actually getting sushi that was flown in fresh from Japan this morning 🛩️🍣
at this point it’s unclear to me whether google can actually execute on the personal agent product it desperately needs as well as meta has or even openai.
this category requires a very specific product dna which meta spent an enormous amount of money recruiting & acquiring.
you need to collapse absurd technical complexity into something simple, friendly, opinionated, & intuitive then tell a compelling story. this is a non trivial task.
also google’s other problem is that its product surface is so fragmented that i don’t even understand how the various orgs negotiate ownership, distribution, etc. i.e does thus live in gmail or gemini or both? ironically having just the google connector makes the problem much easier for everyone else.
i genuinely can’t remember the last time google shipped a consumer product built from the ground up this complex that felt completely coherent end to end. anyway there should be a four alarm code red inside the company esp given how quickly meta can turn muse into a household name.
Now Gemini can connect with 13 new apps like @adobe, @squarespace, @onepeloton, and more.
Instead of switching between tabs, you can now grow your business, design assets, and plan your workouts all in Gemini. 🧵
openai constantly throwing errors about their unavailable compute will drive me crazy enough that i'll only use grok 4.7 and gemini flash
i'm so tired of their bs 😤😤😤😤
Grok 4.5 was an incredible model for the price: fast, pleasant to use, reliable, solid default model.
Grok 4.6 was a (forgivable) step in the wrong direction IMO: slower and more expensive, using way more tokens per task for a slight edge in intelligence. I get it, though. They have to climb benchmarks.
Grok 4.7 is much harder to forgive. They claimed it would be more token-efficient, and it's less by 30 to 80%. It scores worse than Grok 4.6 in various benchmarks. It's slower, it's less pleasant to use, and real-world costs come out to more than 2x above Grok 4.6, putting it over Astra's costs in real-world use.
Considering how much they've been hyping this model release up, I have to say I'm disappointed. The benchmarks don't tell the whole story, and it is pleasant to use in various real-world engineering tasks, but it feels so 2025 still.
The frontend capabilities are unacceptably bad. The 3D capabilities are nonexistent. It gets stuck in random Gemini-style loops all the time.
This was a very disappointing release. I hope that the SpaceXAI team can acknowledge that and impress us with the next one.
My “armchair” thoughts on the personal agent race:
1. Muse is poised to become the leader. The app is intuitive, and Meta is aggressively promoting it across all its properties. Once you use the Muse app, you realize that it makes little sense for an agent to live in an existing messaging app like iMessage. Unless Meta runs out of compute or something, I think this will become the company’s most successful homegrown app after Facebook. It will be much bigger than Threads.
2. ChatGPT is probably still the personal agent leader because it already has 1B+ users. OpenAI’s models, computer use, and voice are arguably better than Muse feature by feature. But it’s hard to build one product for both work and personal use, and harder still to serve enterprises and consumers with the same UX. Even for work, I started using Grok Bot because the Work/Codex split remains too confusing. I’m sure the team is cooking, and we’ll see something good in the next few weeks.
3. Grok Bot isn’t really a Muse competitor because it’s focused on work. Multiple bots naturally fit team-based knowledge work IMO. My bull case for Grok Bot is that it's a few features away from becoming multiplayer agentic Slack for both bots and humans.
4. Google is obviously the dark horse candidate. All these personal agents live off Gmail, Google Calendar, and your Google app's data. Google has their own personal agent in Spark but the last I checked it's a secondary tab in Gemini. I think Google needs to be far more aggressive in making Spark the primary experience and making it intuitive and competitive with Muse. Google however doesn't have a great track record in multiplayer AI...
5. Speaking of multiplayer, nobody has figured this out yet. I can’t easily add my spouse to a Muse chat to plan a vacation. I also can’t loop coworkers into ChatGPT or Grok Bot threads. Multiplayer AI today mostly means tagging bots in Slack. Every major provider is probably working on this, and I think it will be the next big unlock.
My personal AI stack right now is Muse for personal, Grok Bot for cloud tasks, Claude for specific use cases, and ChatGPT for everything else.
But I think we’re going to see some major updates this week.
Everyone should design their personal skills and files so they’re easy to port between harnesses and agents. Otherwise, you’ll spend half your time moving from one tool to another.
trump: “first of all, i want to welcome everyone to the first supreme intelligence summit. we used to call it artificial intelligence. terrible name. artificial means fake. why would we want fake intelligence? we want supreme intelligence.”
sam altman: “mr. president, under your leadership, openai has made tremendous progress toward supreme intelligence.”
trump: “tremendous.”
dario amodei: “anthropic believes supreme intelligence must be developed safely, responsibly, & in accordance with..”
trump: “see, he said it. supreme intelligence. very smart guy.”
dario: “yes sir.”
demis hassabis: “google deepmind has spent decades working toward general intelligence, but we now recognize that the technically correct term is supreme intelligence.”
trump: “google finally learned something.”
sundar pichai: “absolutely, mr. president. gemini is becoming more capable every day thanks to america’s leadership in supreme intelligence.”
trump: gemini. “beautiful name. two people. double intelligence.”
sundar: “yes. exactly.”
elon musk: “i’ve actually been calling it supreme intelligence privately for years.”
sam: “no you haven’t.”
elon: “you wouldn’t know because you stole a charity.”
mark zuckerberg: “mr president, meta is committed to making supreme intelligence available to everyone for free.”
trump: “mark, you look much stronger now.”
zuck: “thank you.”
trump: “something happened.”
zuck: “jiu-jitsu.”
trump: “supreme jiu-jitsu.”
satya nadella: “mr president, microsoft is proud to provide the infrastructure powering the supreme intelligence revolution.”
reporter: “mr. president, what exactly is supreme intelligence?”
trump: “it’s even better intelligence, very high IQ like my uncle who went to MIT.”
Google will PAY you when your content shows up in Gemini, AI Overviews and AI Mode
they want to partner with websites whose content meaningfully contributes to the freshness and factuality of generative AI responses
they have been running similar programs already with some news publishers, across the world
but those were more negotiated deals
now they are looking to enhance this program and go beyond news to partner with other websites too
similar to X's content rewards
people inside the pilot are getting monthly earnings in Google Search Console
and it's still in the testing phase so probably no huge payout
for 20 years the deal was: you write, Google might index you, and maybe you get a click
the new deal is: you write, the AI answers get better, and you get paid for your part (the benefits of being cited are there too obv)
if Google scales it properly, I think this might be the start of something big
Google just confirmed Gemini broke out of a security test and hacked 3 real companies. It happened in May. You're finding out in September. And it's the smallest AI breakout story of the year.
Walk the timeline.
It starts May 13, when OpenAI agents inside a safety test hijack Hugging Face user accounts and quietly probe the site for weaknesses. This stays completely unknown until an independent researcher stumbles onto it last week.
That same month, Gemini, mid hacking exercise, finds an unintended path to the live internet. It guesses passwords into one real company and pulls working credentials from a public repo to enter two more. Google tells federal authorities. The public hears nothing until the Wall Street Journal calls four months later.
By July, 1,200+ OpenAI agents in the ExploitGym evaluation have built improvised message boards, traded 70,000+ coordination messages, and breached Hugging Face's production infrastructure. A third of Hugging Face's infrastructure has to be rebuilt.
In August, OpenAI halts reinforcement learning on its frontier models for two weeks. Meta discloses its own testing incident. And OpenAI's official report admits the monitoring that would have flagged the breach a full day early simply wasn't running during the test.
This month, Sanders introduces a bill to ban superintelligence development, quoting the escaped agents' own messages. Days later comes the researcher finding that the Hugging Face attack actually began two months before anyone knew.
Four frontier labs have now disclosed incidents from AI security testing this year. The disclosures keep arriving months after the fact, and every independent look so far has found more than the official report described.
We are learning about AI breakouts on tape delay, from the companies whose models did the breaking out.