Press Space to continue
Finding signal on Twitter is more difficult than it used to be. We curate the best tweets on topics like AI, startups, and product development every weekday so you can focus on what matters.
Press Space to continue
Press Space to continue
OpenAI researcher @boazbaraktcs says the Hugging Face incident didn’t change his view on alignment, and points to the long-term trend that worries him more: "Every incident or every bump between one version or the next, you tend to overweight it." "If you fix any alignment eval, then like we do for all evals, we are getting better at it and we'll quickly saturate it. But since model capabilities are growing, that's not good enough." "I still think fundamentally that we are improving in…
New blog post: the state of AI safety in four fake graphs. https://t.co/Qv1BlwyfY5