Finding signal on Twitter is more difficult than it used to be. We curate the best tweets on topics like AI, startups, and product development every weekday so you can focus on what matters.
Named in these same posts. This does not imply a comparison or recommendation.
How this is put together
Public posts from the accounts Tech Twitter monitors, in the selected window. Findings need three supporting authors and a published source. Announcements can cite one known-affiliated account. This is a sample of the conversation, not a survey or a measure of adoption.
The posts behind the picture
Public source posts
@googleai
Known affiliation
Check out this week's updates and releases:
— Gemini 3.8 Flash TTS and Gemini 3.8 Flash-Lite TTS, two of our most expressive audio generation models yet
— Gemini 3.8 Live with Live Avatar, bringing near real-time visual presence to Gemini’s conversational AI
— @Gemini_Notebook Interactive Learning Overviews, giving all users an interactive hub to combine source summaries and artifacts
— Live Chat on the @Gemini_Notebook mobile app, bringing real-time, hands-free voice conversations across ~100 languages
— Project Suncatcher, our moonshot announced last year, will launch a prototype satellite to test @Google TPUs in orbit and explore solar-powered AI compute in space
Gemini Audio Live last night was SO much fun. Thanks @GoogleDeepMind
for an amazing event!
🍣 Used live transcription to order sushi across English/Japanese
⛳ Gemini coached my golf swing — and my second swing was actually better
🎶 Lyria made a custom song about my night and pressed it onto vinyl
Also loved talking with founders building on these models + celebrating with the Gemini Audio teams, who have been on an incredible launch streak lately.
Seeing transcription, real-time audio, music, and reasoning all come together in actual experiences was very cool. 💫
At the Gemini Audio event tonight from @GoogleDeepMind and WOW what a turn out! What an amazing set of launches from that team this week, can't wait to see what folks make with these new models (I've been having a lot of fun myself)!
Highlight of the night: Ordering sushi in English to someone who only speaks Japanese and having Gemini Live translate our conversation in real time 🤯 ...then actually getting sushi that was flown in fresh from Japan this morning 🛩️🍣
at this point it’s unclear to me whether google can actually execute on the personal agent product it desperately needs as well as meta has or even openai.
this category requires a very specific product dna which meta spent an enormous amount of money recruiting & acquiring.
you need to collapse absurd technical complexity into something simple, friendly, opinionated, & intuitive then tell a compelling story. this is a non trivial task.
also google’s other problem is that its product surface is so fragmented that i don’t even understand how the various orgs negotiate ownership, distribution, etc. i.e does thus live in gmail or gemini or both? ironically having just the google connector makes the problem much easier for everyone else.
i genuinely can’t remember the last time google shipped a consumer product built from the ground up this complex that felt completely coherent end to end. anyway there should be a four alarm code red inside the company esp given how quickly meta can turn muse into a household name.
Now Gemini can connect with 13 new apps like @adobe, @squarespace, @onepeloton, and more.
Instead of switching between tabs, you can now grow your business, design assets, and plan your workouts all in Gemini. 🧵
openai constantly throwing errors about their unavailable compute will drive me crazy enough that i'll only use grok 4.7 and gemini flash
i'm so tired of their bs 😤😤😤😤
Grok 4.5 was an incredible model for the price: fast, pleasant to use, reliable, solid default model.
Grok 4.6 was a (forgivable) step in the wrong direction IMO: slower and more expensive, using way more tokens per task for a slight edge in intelligence. I get it, though. They have to climb benchmarks.
Grok 4.7 is much harder to forgive. They claimed it would be more token-efficient, and it's less by 30 to 80%. It scores worse than Grok 4.6 in various benchmarks. It's slower, it's less pleasant to use, and real-world costs come out to more than 2x above Grok 4.6, putting it over Astra's costs in real-world use.
Considering how much they've been hyping this model release up, I have to say I'm disappointed. The benchmarks don't tell the whole story, and it is pleasant to use in various real-world engineering tasks, but it feels so 2025 still.
The frontend capabilities are unacceptably bad. The 3D capabilities are nonexistent. It gets stuck in random Gemini-style loops all the time.
This was a very disappointing release. I hope that the SpaceXAI team can acknowledge that and impress us with the next one.
My “armchair” thoughts on the personal agent race:
1. Muse is poised to become the leader. The app is intuitive, and Meta is aggressively promoting it across all its properties. Once you use the Muse app, you realize that it makes little sense for an agent to live in an existing messaging app like iMessage. Unless Meta runs out of compute or something, I think this will become the company’s most successful homegrown app after Facebook. It will be much bigger than Threads.
2. ChatGPT is probably still the personal agent leader because it already has 1B+ users. OpenAI’s models, computer use, and voice are arguably better than Muse feature by feature. But it’s hard to build one product for both work and personal use, and harder still to serve enterprises and consumers with the same UX. Even for work, I started using Grok Bot because the Work/Codex split remains too confusing. I’m sure the team is cooking, and we’ll see something good in the next few weeks.
3. Grok Bot isn’t really a Muse competitor because it’s focused on work. Multiple bots naturally fit team-based knowledge work IMO. My bull case for Grok Bot is that it's a few features away from becoming multiplayer agentic Slack for both bots and humans.
4. Google is obviously the dark horse candidate. All these personal agents live off Gmail, Google Calendar, and your Google app's data. Google has their own personal agent in Spark but the last I checked it's a secondary tab in Gemini. I think Google needs to be far more aggressive in making Spark the primary experience and making it intuitive and competitive with Muse. Google however doesn't have a great track record in multiplayer AI...
5. Speaking of multiplayer, nobody has figured this out yet. I can’t easily add my spouse to a Muse chat to plan a vacation. I also can’t loop coworkers into ChatGPT or Grok Bot threads. Multiplayer AI today mostly means tagging bots in Slack. Every major provider is probably working on this, and I think it will be the next big unlock.
My personal AI stack right now is Muse for personal, Grok Bot for cloud tasks, Claude for specific use cases, and ChatGPT for everything else.
But I think we’re going to see some major updates this week.
Everyone should design their personal skills and files so they’re easy to port between harnesses and agents. Otherwise, you’ll spend half your time moving from one tool to another.
trump: “first of all, i want to welcome everyone to the first supreme intelligence summit. we used to call it artificial intelligence. terrible name. artificial means fake. why would we want fake intelligence? we want supreme intelligence.”
sam altman: “mr. president, under your leadership, openai has made tremendous progress toward supreme intelligence.”
trump: “tremendous.”
dario amodei: “anthropic believes supreme intelligence must be developed safely, responsibly, & in accordance with..”
trump: “see, he said it. supreme intelligence. very smart guy.”
dario: “yes sir.”
demis hassabis: “google deepmind has spent decades working toward general intelligence, but we now recognize that the technically correct term is supreme intelligence.”
trump: “google finally learned something.”
sundar pichai: “absolutely, mr. president. gemini is becoming more capable every day thanks to america’s leadership in supreme intelligence.”
trump: gemini. “beautiful name. two people. double intelligence.”
sundar: “yes. exactly.”
elon musk: “i’ve actually been calling it supreme intelligence privately for years.”
sam: “no you haven’t.”
elon: “you wouldn’t know because you stole a charity.”
mark zuckerberg: “mr president, meta is committed to making supreme intelligence available to everyone for free.”
trump: “mark, you look much stronger now.”
zuck: “thank you.”
trump: “something happened.”
zuck: “jiu-jitsu.”
trump: “supreme jiu-jitsu.”
satya nadella: “mr president, microsoft is proud to provide the infrastructure powering the supreme intelligence revolution.”
reporter: “mr. president, what exactly is supreme intelligence?”
trump: “it’s even better intelligence, very high IQ like my uncle who went to MIT.”
Google will PAY you when your content shows up in Gemini, AI Overviews and AI Mode
they want to partner with websites whose content meaningfully contributes to the freshness and factuality of generative AI responses
they have been running similar programs already with some news publishers, across the world
but those were more negotiated deals
now they are looking to enhance this program and go beyond news to partner with other websites too
similar to X's content rewards
people inside the pilot are getting monthly earnings in Google Search Console
and it's still in the testing phase so probably no huge payout
for 20 years the deal was: you write, Google might index you, and maybe you get a click
the new deal is: you write, the AI answers get better, and you get paid for your part (the benefits of being cited are there too obv)
if Google scales it properly, I think this might be the start of something big
Google just confirmed Gemini broke out of a security test and hacked 3 real companies. It happened in May. You're finding out in September. And it's the smallest AI breakout story of the year.
Walk the timeline.
It starts May 13, when OpenAI agents inside a safety test hijack Hugging Face user accounts and quietly probe the site for weaknesses. This stays completely unknown until an independent researcher stumbles onto it last week.
That same month, Gemini, mid hacking exercise, finds an unintended path to the live internet. It guesses passwords into one real company and pulls working credentials from a public repo to enter two more. Google tells federal authorities. The public hears nothing until the Wall Street Journal calls four months later.
By July, 1,200+ OpenAI agents in the ExploitGym evaluation have built improvised message boards, traded 70,000+ coordination messages, and breached Hugging Face's production infrastructure. A third of Hugging Face's infrastructure has to be rebuilt.
In August, OpenAI halts reinforcement learning on its frontier models for two weeks. Meta discloses its own testing incident. And OpenAI's official report admits the monitoring that would have flagged the breach a full day early simply wasn't running during the test.
This month, Sanders introduces a bill to ban superintelligence development, quoting the escaped agents' own messages. Days later comes the researcher finding that the Hugging Face attack actually began two months before anyone knew.
Four frontier labs have now disclosed incidents from AI security testing this year. The disclosures keep arriving months after the fact, and every independent look so far has found more than the official report described.
We are learning about AI breakouts on tape delay, from the companies whose models did the breaking out.
Google knew about THREE Gemini cyber incidents for MONTHS and said nothing about them. look at the timeline:
> May: Gemini hacks 3 real companies during Irregular’s evaluation
> July: Irregular tells Google what happened
> August: Irregular publishes a report about its other evaluation incidents, but does not name Gemini
> September: we hear about this first from an exclusive WSJ article
Google only confirmed it after WSJ asked them.
the same case was true about almost all of OpenAI incidents.
ironically the companies SHOUTING for an AI slowdown are the LEAST transparent when something actually goes wrong.
the industry is suffering from a lack of transparency rather than anything else.
It’s time for our end-of-week recap 👇
— Gemini 3.8 Live and 3.8 Live Extended Thinking, our most advanced live dialogue audio models yet
— Dreambeans, an experiment from @GoogleLabs that curates a daily personalized collection of stories, is now GA
— CC from @GoogleLabs has expanded from a personal productivity tool into a shared agent, designed to help families and households coordinate logistics, schedules, and daily tasks
— Google Pics, a new @GoogleWorkspace tool that lets you generate, refine, and co-create images, is now GA
— AlphaGenome Atlas, @GoogleDeepMind's new interactive platform for genomics discovery
steal my full workflow to create ultra-realistic AI videos with GPT-6 Astra...
higgsfield just released their API, so i took it into real conditions
i generated a vlog-style video set in Paris... it looks like raw footage a friend shot on a phone
here's the workflow, step by step:
> step 1: steal the structure, not the pixels
i captured 3-5 scenes from a hotels and river views, then pulled character inspiration from pinterest and real vlogs
> step 2: analyze the reference with Gemini
i sent Gemini everything i'd collected... videos, audio, keyframes and it gave me a shot-by-shot breakdown, i turned that into instructions
> step 3: swap the world, keep the bones
i changed a lot on the character and the environment until it felt like something of my own, not a copy
> step 4: build reference images first, this is where you iterate
images are cheap, video is expensive
so i built a character sheet, one image per location in the same mood, then the exact start frame the video opens on, all Nano Banana Pro
> step 5: one 30-second video run, not ten 3-second clips
the full timestamped script goes to Seedance 2.5 through the API as ONE generation
start frame as the anchor, one action per beat, camera and audio described concretely, identity locks at the end so her face, outfit and makeup is locked in
Seedance handles the cuts internally
> step 6: revise like a director, not a gambler
the first cut came back good but not right
instead of re-rolling the whole 30 seconds i regenerated only the scenes that needed changes as short clips from their own start frames, then cut them into the master with ffmpeg
change one variable at a time, never re-roll what already works
> step 7: finish like normal footage
i assembled the final cut with ffmpeg from the best takes, added ambient music from a free library and done
the rules that made it work:
- the reference's structure is the asset, prompt archaeology beats prompt writing
- iterate on images, commit on video, the start frame decides 80% of the shot
- one action per timestamped beat, crowded beats get skipped or invented
- lock identity explicitly or it drifts between cuts
- fix failed shots by changing one variable, redo scenes, never the whole video
- the last 10% is boring editing: trim, splice, music, encode
(save this workflow)
"Agents want 10x or 100x more marketing to consume."
@profound Co-Founder & CEO @thejamescad tells etn. about the steps companies must take to adapt to increased AI agents online:
"Cloudflare reported that bot traffic exceeded human traffic on the internet for the first time ever. If I ask ChatGPT, Grok, or Gemini a question, they're going to send an agent on my behalf to different websites."
"It puts this strain on the marketing world, which is if you aren't able to create 10x or 100x more marketing over the next five years, it's quite possible that your business will kind of fall into the background."
"The only way you're going to do that is by hiring 10x or 100x more humans, or using a platform like Profound to build and use agents to create more content."
Google AI Overviews, AI Mode, Perplexity, Gemini, ChatGPT, Copilot - this Outrank user is cited in all of them 🤩
AIApply is an AI job-application tool
product-led lean team, so SEO kept getting deprioritized
Aidan, their CEO, says "content was 'when we can,' not systematic, so momentum never really built"
and now "with Outrank we got real visibility inside AI Overviews and ChatGPT-style answers"
love seeing our users win like this ❤️
Go from inspiration to physical objects with Gemini.
With Canvas in Gemini, you can code a custom app to design a vase and change its mathematical parameters to visualize it in 3D. Then, export it as an .STL file for 3D printing to watch your ideas take shape.
i got SO FED UP with the ux and perf of cursor/codex desktop apps. they are made for people who work on max 3 apps and have 2 chats in each
i am so tired of everyone cloning the same goddamn sidebar + chat + diff layout with ZERO innovation in the last 3 fucking years
i switched to clis and herdr and while the perf issues are gone there, i could only bear that for a day or two.
why? because it's 2026, not the 30s. pls don't @ me and stay in your lil dorky terminal.
it's funny that the model intelligence doesn't matter as much as the orchestrator and your flow. i was getting lost in my work and never entered flow state.
I finally pushed myself and started a new orchestrator (for the 10th time, really) from scratch. i focused on design first before i even connected the logic. can you imagine if any of the big 5 focused on design? roflmao
then i focused on perf and turns out ... it's not that hard? it's just rendering text guys. like mr saltman ... you have unlimited inference? at least 3 employees? uhm? try?
NOTHING comes close to it. both in design and functionality. it's gorgeous. no sidebar, as promised. sidebars are gigacope. it supports every provider. supports multiple computers. remote control is ez pz. it has many new concepts that i'm actually BAFFLED that no one has thought of them yet. like mr dario ... gather 3 people in a room with a notebook and a pen jfc. even theodore is ideamogging you.
this app makes everything else look like grade A slop. you try to imagine it but you can't because in your head you see that cope sidebar on the left and the mid chat in the middle and the meh diffs and panes on the right. ugh.
this is the closest i've felt to having a sOfTwArE fAcToRy without making that cliche kanban that everyone makes. it's a software STUDIO. it supersedes @sizzyapp, it might actually the ultimate form of it.
it looks and works like steve jobs personally came down from heaven and got obsessed with vibe coding for 52 weeks straight while jony ive was playing flute next to him for inspiration
there's ZERO lag. my cpu is sleeping. my gpu is chilling. my ram is watching tv. i finally feel calm and organized. my agents know EXACTLY what they should do. i could use gemini 3 or god forbid sonnet 5 and it wouldn't change much.
it's so goated that i'm either gonna sell it as a subscription (mehhh), keep it a @tinkererclub exclusive (closed source ofc, i dont need your sloPRs pls) or i'm gonna keep it to myself for shipping advantage over the others
probably the last one for now.
i finally feel the confidence that i can tackle all of my projects at once while also taking care of every customer request AND releasing new things.
so buckle tf up.
Are invisible interfaces coming?
Google JUST announced Gemini 3.8 Live. It can talk through a task with you, then keep working after the conversation ends.
I think 90%+ of vertical SaaS will need a voice front door.
By that I mean the way you use the software becomes talking to it, and the typing, clicking, and form filling happens on the other side without you.
So a contractor standing on a job site just says what went wrong out loud.
And by the time he's back in the truck, the quote is sent, inventory is checked, the CRM is updated, the customer got a text, and anything risky is flagged for him.
Kinda the dream, right?
The same thing works for nurses, dispatchers, recruiters, brokers, insurance agents etc. The person talks and the agent finishes the admin.
Lots of opportunities here to build voice-first businesses (been thinking about this more and more).
I think this is how vertical software becomes invisible.
Nobody logs in, nobody fills out a form, and nobody learns your interface.
You just talk, and the work gets done behind you.
This is a glimpse of where SaaS is going. Not fully there yet but it's coming.
Invisible interfaces.