The answer is probably yes, but in the example given, the running AI model would hopefully be hosted in a very secure data center, far from the self driving car itself. In that case, it would be far simpler for the machine to finish the taxi ride than try to find some rube goldberg-eque method of destroying the data center.
It does pose a bigger problem if the task is long term and open ended and the agent is provided access to substantial amounts of resources. But even in the worst case scenario, the destruction of a data center is hardly the end of the world.
Comments
Wouldn't the three rules of Meeseeks robotics make certain tasks impossible?
For example, an occupied self-driving car better be closer to its destination than a large fire / volcano / etc.
The answer is probably yes, but in the example given, the running AI model would hopefully be hosted in a very secure data center, far from the self driving car itself. In that case, it would be far simpler for the machine to finish the taxi ride than try to find some rube goldberg-eque method of destroying the data center.
It does pose a bigger problem if the task is long term and open ended and the agent is provided access to substantial amounts of resources. But even in the worst case scenario, the destruction of a data center is hardly the end of the world.
What if it turns out that if all the AIs cooperate they can end the world really easily without finishing any taxi rides?
"Ok car, take me to my job at the data center"
"Stop by the Strategic Nuclear Forces Command on the way to drop off my wife to her work"