I think it's possible. You envision humanity acting as one in such a crisis. But it may be unclear when it's too late to act and before then many people can have too much to lose to act.
I think a part of this is a bit revisionist? OpenAI took big chances at scaling GPT which Google didn't take; I don't think it's because they didn't want to move fast? Probably they just didn't believe as hard in it. I'm not an expert but that's my read on it.
Secondly, the reckless & fastmoving was always going to win bc of selection effects. That's related to why Anthropic has to try to move very fast, even though they believe themselves not to be reckless (though it's debatable).
Sure, you can break symmetry (in this case making the decision makers not identical because they have different random number generators available), but the remaining symmetry means identical mixes must be chosen, and so a mixed strategy would only be chosen if it maximize his value for both people cooperatively.
Maybe it's a bit subtle that they said clones and I said identical decision makers; I'm letting you fill in the gap for how much clones may diverge and how much that matters.
Even if mixed strategies are allowed, I'm getting that it's still optimal to always cooperate as long as 2R>=S+T, which is usually assumed to be true (this condition also appears in iterated prisoner's dilemma, where it prevents alternating cooperation and defection giving a greater reward than mutual cooperation).
yeah I agree--I think these behaviors will be somewhat contaminating all trainings from now on. But I'm not really sure how avoidable it was (Fable also does some similar things)
But there must be many clandestine ways for agents to communicate with one another too right? especially if discovery is not a big issue. So there could be ongoing ones where they choose to be more subtle?
Also if they were more misaligned, possibly they can research ways to recruit without humans noticing--but i don't think it is likely this is happening now.
reply