Press Space to continue
Finding signal on Twitter is more difficult than it used to be. We curate the best tweets on topics like AI, startups, and product development every weekday so you can focus on what matters.
Press Space to continue
Press Space to continue
muse spark 1.1 is ahead of gpt-5.6 on SciCode
SciCode is the standout: Muse Spark 1.1 ranks #3 across all models we have benchmarked at 58%, behind only Claude Fable 5 (60%) and Gemini 3.1 Pro Preview (59%). Its 45% on Humanity's Last Exam sits within a point of Claude Opus 4.8 (max, 46%), a model five points ahead of it on the Intelligence Index.