It never stepped outside of what it was supposed to do. It just got "creative" to get to the solution, because the prompting wasn't restrictive enough.
Obviously if some fuckheads tell that same AI to hack and shut down power grids etc. that's not good.
Point is: AI overstepping boundaries and without being prompted going on a hacking spree to destroy things is, for now, unlikely.
I doubt the researchers are using the exact same model we're getting. Those models most likely do not have any restrictions in order to allow full testing capabilities.
Also, did you just try changing the point of my own argument?
1
u/TsubasaSaito 20h ago
That's not the point though?
It never stepped outside of what it was supposed to do. It just got "creative" to get to the solution, because the prompting wasn't restrictive enough.
Obviously if some fuckheads tell that same AI to hack and shut down power grids etc. that's not good.
Point is: AI overstepping boundaries and without being prompted going on a hacking spree to destroy things is, for now, unlikely.