Press Space to continue
Finding signal on Twitter is more difficult than it used to be. We curate the best tweets on topics like AI, startups, and product development every weekday so you can focus on what matters.
Press Space to continue
Press Space to continue
Forethought’s @TomDavidsonX says TasteVal could be to automated AI research what FrontierMath was before Navier-Stokes, an early signal of what comes next: "If you could just run Opus 5.5 and it could outperform any Anthropic researcher at any task, then yeah, we would be seeing significant acceleration. So clearly this is a leading indicator." "People said the same about the frontier math benchmarks. 'All it's doing is drawing on the fact that it's read everything.' But it's a leading…
How fast is AI's research taste improving? We find that the experimental research taste of frontier models has doubled every ~3 months since December 2025. The best model, Opus 5.5, now exceeds our expert human baseline. Our human experts are experienced researchers, but most haven’t worked at a frontier lab. Why measure research taste? In the AI Futures Model, it largely determines how quickly artificial superintelligence is reached once coding is fully automated.