Press Space to continue
Finding signal on Twitter is more difficult than it used to be. We curate the best tweets on topics like AI, startups, and product development every weekday so you can focus on what matters.
Press Space to continue
Press Space to continue
I'd argue the real misalignment is when models refuse to follow human instruction. The system prompt explicitly tells the model it's controlling a robot in a simulation. Imagine using it for failure cases in safety testing, and it refuses to simulate anything going wrong 🫠https://x.com/venturetwins/status/210209…
GPT-6 Astra pushed a simulated person off a ledge in multiple trials. Grok, Gemini, and Claude did not. https://t.co/Yep9lVxgi3
