I would say that there are different kinds of risks that we should distinguish between. First, there’s the misuse risk — that people will use the model to get new ideas on how to do harm or will use it as part of some malicious system. And then there’s the risk of a treacherous turn, that the AI will have some goals that are misaligned with ours and wait until it’s powerful enough and try to take over.
John Schulman
co-founder of OpenAI; led the team that created ChatGPT
Asked in an April 2023 Berkeley News interview whether he had concerns about the safety of the GPT models, John Schulman distinguished misuse risk from the risk of a treacherous turn, in which an AI has goals misaligned with ours and waits until it is powerful enough and tries to take over, while saying in the same answer that misuse is not an existential risk at that stage and that a takeover is quite unlikely.
Co-founder of OpenAI and leader of the reinforcement learning team that developed ChatGPTApril 19, 2023Berkeley News interview during an EECS Colloquium campus visit
Context and checks
Surrounding words
The success of ChatGPT has renewed fears about the future of AI. Do you have any concerns about the safety of the GPT models? I would say that there are different kinds of risks that we should distinguish between. First, there’s the misuse risk — that people will use the model to get new ideas on how to do harm or will use it as part of some malicious system. And then there’s the risk of a treacherous turn, that the AI will have some goals that are misaligned with ours and wait until it’s powerful enough and try to take over. For misuse risk, I’d say we’re definitely at the stage where there is some concern, though it’s not an existential risk.
What this quote does not say
- The passage describes risk mechanisms but does not estimate their probability or state that takeover would succeed.
- The interview was edited for length and clarity rather than published as an audio transcript.
How this was checked
- Quote matched character-for-character against the captured page.
Reviewed by
an AI reviewer that read the whole source · an automatic character-by-character check