Finding signal on Twitter is more difficult than it used to be. We curate the best tweets on topics like AI, startups, and product development every weekday so you can focus on what matters.
3,431 tokens/s on Gemma 4 31B! Recent benchmarks on NVIDIA's | Tech Twitter
Press Space to continue
3,431 tokens/s on Gemma 4 31B!
Recent benchmarks on NVIDIA's Groq 3 LPX accelerator show a median output of ~3,400 tokens/s across both 10k and 100k input sequences, the fastest 100k-context speed recorded for the model to date.