Press Space to continue
Finding signal on Twitter is more difficult than it used to be. We curate the best tweets on topics like AI, startups, and product development every weekday so you can focus on what matters.
Press Space to continue
Press Space to continue
Arena CEO @ml_angelopoulos argues OpenAI and Anthropic can’t be the only ones deciding whether their own models are safe as agents get powerful enough to break out of sandboxes: "There's an element of safety that we need to evaluate as well, because given the capabilities of these models to break out of their sandboxes, there needs to be a neutral evaluator for that." "It's not a matter of if, but when, there's gotta be a neutral evaluation platform, because frankly, the model labs are not…