no - but you could learn what they are truly capable of and restrict them accordingly for public release. I think that is the point on this research. Also publishing findings before uncensored models catch up and will inevitably used for criminal purposes
Learning what the models are capable of is exactly what the test achieved, so I’d personally call it a success. So it created a few GitHub accounts. Who cares? Seeing the same behavior in the wild post-release would be infinitely worse.
Comments
You set them up with an internal intranet.
Are the models going to exclusively run on intranets?
The versions which haven't been post-trained not to go hack stuff? Yes, I would say those models should be exclusively run on intranets.
OpenAI said the model was sandboxed, so the intranet just needs to provide the same resources which were supposed to be available within the sandbox.
“Should be” is not reality. These models are in the hands of plenty of companies and governments today.
no - but you could learn what they are truly capable of and restrict them accordingly for public release. I think that is the point on this research. Also publishing findings before uncensored models catch up and will inevitably used for criminal purposes
Learning what the models are capable of is exactly what the test achieved, so I’d personally call it a success. So it created a few GitHub accounts. Who cares? Seeing the same behavior in the wild post-release would be infinitely worse.
The point is to test capabilities prior to connecting them to the internet.
So the first time the model gets internet access should be post-release in the hands of random people?