Press Space to continue
Finding signal on Twitter is more difficult than it used to be. We curate the best tweets on topics like AI, startups, and product development every weekday so you can focus on what matters.
Press Space to continue
Press Space to continue
💻 Z .ai's GLM-5.3 just hit 84.5% on the CyberGym vulnerability benchmark, beating top proprietary models, a huge gain over the performance of its predecessor GLM-5.2. The kicker? http://z.xn--ais-so0a AI engineers did it purely through fine-tuning and optimization of the model’s agentic capabilities, without changing the base model. The model grew so capable at finding and targeting potential exploits that http://z.ai held back the open weights for safety testing. Read the full analysis in The…
