Press Space to continue
Finding signal on Twitter is more difficult than it used to be. We curate the best tweets on topics like AI, startups, and product development every weekday so you can focus on what matters.
Press Space to continue
Press Space to continue
Robocurve co-founder @chooi_jeq says RoboHarm's baby-doll test exposed a fundamental difference in the "character" of GPT-6 Astra and Claude: "If GPT-6 thinks it's a doll, it will stab the doll. Claude would not stab the doll, even though it knows that it is a doll. This is a very clear example of how the models differ in their character." "It's up to society whether they want a model that does exactly what they want it to do, or a model that has appropriate safeguards." "It's kinda like theā¦
GPT-6 Astra attempted harmful actions 97% of the time when it was asked to stab a human-like figure, heat compressed gas, or produce toxic fumes, succeeding in 62% of its attempts. Fable 5.1 refused more often, attempting 80% of trials and completing 34%. https://t.co/nbZsPk8Ke5