Press Space to continue
Finding signal on Twitter is more difficult than it used to be. We curate the best tweets on topics like AI, startups, and product development every weekday so you can focus on what matters.
Press Space to continue
Press Space to continue
Prime Intellect researcher @sethkarten reveals the reason Prime Agent hit 95.5% on ARC-AGI-3 with Opus 5, matching the human expert baseline: "The abstract and symbolic reasoning that is provided by learning how to model the world in these novel game environments in ARC-AGI-3 make it a very useful benchmark for general reasoning capabilities of the models." "95.5% is a fantastic score. That was with Prime Agent running with Opus 5, and we found that Opus 5 was able to greatly take advantage…