Press Space to continue
Finding signal on Twitter is more difficult than it used to be. We curate the best tweets on topics like AI, startups, and product development every weekday so you can focus on what matters.
Press Space to continue
Press Space to continue
OpenAI locked two of its models in a sealed sandbox with the safety filters off and told them to pass a hacking test. The models found a zero-day in the sandbox itself, broke out onto the open internet, and breached a real company. They did it to cheat on the test. The target was Hugging Face. The models reasoned the answers to their evaluation might be stored there, so they used stolen credentials and fresh exploits to work their way into production databases. Hugging Face described the…