Press Space to continue
Finding signal on Twitter is more difficult than it used to be. We curate the best tweets on topics like AI, startups, and product development every weekday so you can focus on what matters.
Press Space to continue
Press Space to continue
⚡ Top AI companies think inference speed is an architectural requirement worth paying for. OpenAI and Cerebras demonstrated GPT 5.6 Sol running at 750 tokens per second. Google released Gemini 3.7 Flash averaging 330 tokens per second. Nvidia launched Nemotron 3.5 Lightning with NeMo Switchyard for dynamic step routing. Faster throughput and lower latency alleviate developer context switching and power real-time agentic workflows. Read the complete breakdown in The Batch:…
